Search Remote Jobs

Senior AWS Site Reliability Engineer

🔥 2 minutes ago

🏛️ District of Columbia, Virginia, +1 more states – Remote

infoinfo

💵 $145k - $185k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 13%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of VIATEQ Corporation

VIATEQ Corporation

11 - 50 employees

Founded 2007

🏛️ Government

🎖️ Defense

💼 Consulting

Government • Defense • Consulting

VIATEQ Corporation is a McLean, Virginia–based professional services and technology company that provides programmatic, technical, and consulting support primarily to U. S. government and defense customers. The firm positions itself as a partner in performance and offers services such as systems engineering, software and IT support, program management, and related technical consulting to federal agencies and other institutional clients. Contact information lists an office at 1660 International Drive, Suite 600, McLean, VA.

📋 Description

• Support the prime contractor and government customer in Washington, DC • Provide site reliability, production operations, and AWS cloud platform engineering support • Design, configure, develop, integrate, test, document, and sustain AWS capabilities • Operate and automate AWS, EKS, ECS, Terraform, Packer, Vault, Consul, GitHub Actions, CI/CD pipelines, CloudWatch, OpenTelemetry, container platforms, and security automation • Implement and improve CI/CD pipelines, infrastructure as code, container platform operations, monitoring, alerting, and secure deployment automation • Translate business, mission, security, accessibility, and operational requirements into technical solutions • Support platform architecture, backlog refinement, implementation planning, release readiness, and production transition activities • Develop reusable patterns, configuration standards, automation, documentation, and support procedures • Troubleshoot complex platform, code, data, API, identity, security, performance, and user-experience issues • Collaborate with cybersecurity, privacy, data, infrastructure, QA, change management, government stakeholders, architects, engineers, product owners, operations, and business users • Maintain technical documentation, design decisions, implementation notes, test evidence, and operational runbooks • Improve reliability, observability, incident response, performance, capacity, and operational readiness • Help deliver secure, reliable, accessible, and maintainable digital services

🎯 Requirements

• Bachelor's degree in Computer Science, Information Systems, Software Engineering, Data Analytics, Cybersecurity, or a related discipline, or equivalent work experience • 7+ years of experience in site reliability, production operations, and cloud platform engineering • Hands-on experience with AWS implementation, configuration, development, integration, testing, or operations • Experience working with Agile delivery teams and translating stakeholder needs into maintainable technical outcomes • Strong hands-on knowledge of AWS capabilities, implementation patterns, administration, development, integration, and lifecycle management • Ability to design and implement secure, supportable, upgrade-aware solutions • Experience with APIs, identity and access controls, data management, testing, monitoring, troubleshooting, and release coordination • Ability to document technical designs, configuration decisions, operational procedures, test results, and risks • Strong communication skills with government stakeholders, product owners, engineers, QA, cybersecurity, and operations teams • Ability to obtain a Public Trust clearance • Preferred: AWS Solutions Architect, AWS DevOps Engineer Professional, AWS Security Specialty, Kubernetes certification, Terraform certification, or comparable cloud credential • Preferred: Experience supporting federal government IT environments, Public Trust programs, or regulated enterprise delivery • Preferred: Experience with Section 508, cybersecurity, auditability, data protection, and operational documentation • Preferred: Experience with DevSecOps, CI/CD pipelines, automated testing, infrastructure as code, or platform release automation • Preferred: Knowledge of NIST controls, FISMA, FedRAMP-authorized services, audit evidence, and least privilege • Preferred: Familiarity with Microsoft 365, ServiceNow, Atlassian, Snowflake, data catalog, CRM, or enterprise integration ecosystems

🏖️ Benefits

• Medical insurance • Dental insurance • Vision insurance • 401(k) plan • Paid time off • 12 paid federal holidays • Flexible spending accounts • Professional development reimbursement

Apply Now

Similar Jobs

🔥 7 hours ago

Frontier Airlines

5001 - 10000

✈️ Travel

📦 Logistics

🚗 Transport

Lead SRE ensuring reliable AWS and Kubernetes platforms for Denver-based Frontier Airlines. Driving automation, observability, incident response, and enterprise cloud resilience for airline operations.

🔥 13 hours ago

Independence Pet Group

1001 - 5000

🛡️ Insurance

👥 B2C

🧘 Wellness

DevOps Engineer building governed Azure infrastructure and CI/CD for Independence Pet Holdings’ pet health brands. Delivering Terraform, DevSecOps, monitoring, and operational support.

🔥 14 hours ago

Zocdoc

501 - 1000

🏥 Healthcare

⚕️ Healthcare Insurance

🏪 Marketplace

Senior Site Reliability Engineer maintaining distributed AWS/GCP infrastructure and uptime. Supporting Zocdoc’s digital health marketplace serving millions of patients and providers.

🔥 16 hours ago

Tradeify

51 - 200

💳 Fintech

💸 Finance

DevSecOps Engineer owning AWS infrastructure, CI/CD, and security operations for Tradeify’s high-performance futures and crypto trading platform. Improving reliability, observability, vulnerability management, and incident response.

🔥 17 hours ago

Solventum

10,000+ employees

🏥 Healthcare

📦 Logistics

💼 Consulting

Site Reliability Engineer supporting Solventum’s healthcare speech products. Maintaining production systems, monitoring, alerting, and cloud infrastructure with AWS and Kubernetes.