
1001 - 5000 employees
Founded 1975
🏛️ Government
🎖️ Defense
💼 Consulting
Government • Defense • Consulting
Koniag Government Services is an Alaska Native Corporation (ANC) that provides technical, professional, and operational expertise to the U. S. public sector. KGS supports Defense & Intelligence, Federal Civilian, and Health customers with enterprise solutions, professional services, and operations management, and emphasizes mission-focused outcomes, contracting speed (ANC direct awards), and strategic/technology partnerships. The company positions itself as a mission partner delivering people, technology, and program management to government customers.
🔥 4 minutes ago
🏛️ District of Columbia, Washington – Remote
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Improve your chances of getting an interview by checking your resume score before you apply.

1001 - 5000 employees
Founded 1975
🏛️ Government
🎖️ Defense
💼 Consulting
Government • Defense • Consulting
Koniag Government Services is an Alaska Native Corporation (ANC) that provides technical, professional, and operational expertise to the U. S. public sector. KGS supports Defense & Intelligence, Federal Civilian, and Health customers with enterprise solutions, professional services, and operations management, and emphasizes mission-focused outcomes, contracting speed (ANC direct awards), and strategic/technology partnerships. The company positions itself as a mission partner delivering people, technology, and program management to government customers.
• Monitor enterprise infrastructure, application, and cloud dashboards, including Splunk, SolarWinds, Grafana, or CloudWatch • Identify and triage issues • Respond to system alerts and perform initial troubleshooting • Escalate incidents according to established runbooks and ServiceNow processes • Assist in building and maintaining dashboards, alert rules, and automated notifications for infrastructure and application health • Support root cause analysis and after-action documentation for production incidents • Assist with synthetic monitoring and validation checks for business disaster continuity and recovery testing • Maintain accurate incident tickets, monitoring runbooks, and knowledge base articles • Collaborate with server, cloud, storage, and database engineers to ensure systems are properly instrumented and monitored • Participate in an on-call rotation supporting after-hours incident response • Support enterprise monitoring, alerting, and incident response for a federal civilian customer's hybrid infrastructure environment
• Associate's or Bachelor's degree in Information Technology or related field, or equivalent professional experience • Minimum of 5 years of experience in IT operations, monitoring, or a related support role • Familiarity with enterprise monitoring/observability tools such as Splunk, SolarWinds, Grafana, Datadog, or similar • Basic understanding of cloud infrastructure concepts; AWS preferred • Strong attention to detail and ability to follow incident response procedures • Understanding of ITSM ticketing and escalation processes • Ability to obtain Public Trust clearance • Basic scripting ability in Python, PowerShell, or Bash is a plus • Ability to work collaboratively in a fast-paced environment • Excellent communication skills and ability to convey complex technical concepts to non-technical stakeholders • Experience working in a federal government IT environment is desired • CompTIA Network+, Splunk Core Certified User, or AWS Certified Cloud Practitioner is desired • Exposure to ServiceNow or similar ITSM/on-call platforms is desired • Prior experience on a federal government IT support contract is desired • Interest in growing toward a Site Reliability Engineering career path
• Health insurance • Dental insurance • Vision insurance • 401(k) with company matching • Flexible spending accounts • Paid holidays • Three weeks paid time off • Competitive compensation
Apply Now🔥 1 hour ago
51 - 200
Senior DevOps Engineer building DevOps for InfoTrust, a data-driven marketing performance company. Owning AWS and Google Cloud reliability, automation, databases, infrastructure security, and SOC 2 controls.
🔥 2 hours ago
DevOps Engineer building secure TypeScript/Node.js services and GCP infrastructure for RCH Solutions’ scientific discovery platforms. Owning RBAC, Terraform, CI/CD, observability, and governed data access.
🔥 7 hours ago
Site reliability and security engineer securing Solventum’s healthcare speech products. Automating patching, operating production systems, and improving monitoring, alerting, and architecture.
🇺🇸 United States – Remote
💵 $106k - $145.8k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 9 hours ago
Senior network engineer scaling Coupa’s AI-powered spend management cloud platform. Automating public-cloud networking, Kubernetes infrastructure, and reliability operations.
🔥 10 hours ago
Senior cloud deployment engineer delivering secure Azure, AWS, and hybrid infrastructure for business and government customers. Leading migrations, automation, networking, security, and modernization initiatives.