Senior Site Reliability Engineer, Microsoft

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Koniag Government Services

Koniag Government Services

1001 - 5000 employees

Founded 1975

🏛️ Government

🎖️ Defense

💼 Consulting

Government • Defense • Consulting

Koniag Government Services is an Alaska Native Corporation (ANC) that provides technical, professional, and operational expertise to the U. S. public sector. KGS supports Defense & Intelligence, Federal Civilian, and Health customers with enterprise solutions, professional services, and operations management, and emphasizes mission-focused outcomes, contracting speed (ANC direct awards), and strategic/technology partnerships. The company positions itself as a mission partner delivering people, technology, and program management to government customers.

📋 Description

• Design, configure, develop, integrate, test, document, and sustain Microsoft Azure capabilities • Support Azure cloud platform engineering, AKS operations, infrastructure automation, CI/CD, observability, platform reliability, and AI-enabled delivery practices • Improve reliability, observability, incident response, performance, capacity, and operational readiness • Provide site reliability, production operations, and cloud platform engineering support for Microsoft Azure capabilities • Implement and improve CI/CD pipelines, infrastructure as code, container platform operations, monitoring, alerting, and secure deployment automation • Translate business, mission, security, accessibility, and operational requirements into practical technical solutions • Support platform architecture, backlog refinement, implementation planning, release readiness, and production transition activities • Develop reusable patterns, configuration standards, automation, documentation, and support procedures • Troubleshoot complex issues across platform configuration, code, data, APIs, identity, security, performance, and user experience • Collaborate with cybersecurity, privacy, data, infrastructure, QA, and change management teams • Maintain technical documentation, design decisions, implementation notes, test evidence, and operational runbooks • Work with government stakeholders, architects, engineers, product owners, cybersecurity, operations, and business users to deliver secure, reliable, accessible, and maintainable digital services

🎯 Requirements

• Bachelor's degree in Computer Science, Information Systems, Software Engineering, Data Analytics, Cybersecurity, or a related discipline, or equivalent work experience • 7+ years of experience in site reliability, production operations, and cloud platform engineering • Hands-on experience with Microsoft Azure implementation, configuration, development, integration, testing, or operations • Experience working with Agile delivery teams and translating stakeholder needs into maintainable technical outcomes • Strong hands-on knowledge of Microsoft Azure capabilities, implementation patterns, administration, development, integration, and lifecycle management • Ability to design and implement secure, supportable, upgrade-aware solutions • Experience with APIs, identity and access controls, data management, testing, monitoring, troubleshooting, and release coordination • Ability to document technical designs, configuration decisions, operational procedures, test results, and risks • Ability to obtain a Public Trust clearance • Experience with DevSecOps practices, CI/CD pipelines, automated testing, infrastructure as code, or platform release automation • Knowledge of NIST controls, FISMA, FedRAMP-authorized services, audit evidence, and least privilege • Experience improving platform governance, reuse, documentation, observability, and operational readiness • Familiarity with Microsoft 365, ServiceNow, Atlassian, Snowflake, data catalog, CRM, or enterprise integration ecosystems • Microsoft Azure or Kubernetes certification (desired)

🏖️ Benefits

• Health, dental and vision insurance • 401K with company matching • Flexible spending accounts • Paid holidays • Three weeks paid time off • Extraordinary benefits package • Competitive compensation

Apply Now

Similar Jobs

🔥 5 hours ago

Gormat

11 - 50

🔒 Cybersecurity

🏛️ Government

🎖️ Defense

Cloud DevOps Engineer developing and integrating cloud-based solutions. Improving system performance, configuration, reliability, and release processes with up to 25% travel.

🔥 7 hours ago

Zigabyte

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

DevSecOps Engineer building secure CI/CD pipelines and cloud-native infrastructure. Integrating cybersecurity, automation, and compliance controls for consulting solutions.

🔥 9 hours ago

Identiq

51 - 200

💳 Fintech

🛍️ eCommerce

🔒 Cybersecurity

Founding Site Reliability Engineer building observability, incident management, and SRE practices for Incident IQ’s K-12 district workflow platform. Defining SLIs, SLOs, and reliability automation.

🔥 9 hours ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Site Reliability Engineer improving reliability, performance, and scalability across Akamai’s distributed cloud and edge platform. Automating operations, strengthening observability, and leading incident response.

🔥 10 hours ago

Jabil

10,000+ employees

🚘 Automotive

🎖️ Defense

🏥 Healthcare

Lead Fortinet network security and site reliability engineering for Jabil’s manufacturing infrastructure. Securing networks, virtualization, storage, and production test environments.