
10,000+ employees
đŒ Consulting
đ„ Healthcare
đŠ Logistics
Consulting âą Healthcare âą Logistics
Peraton is a mission-focused enterprise that supports national security initiatives through advanced IT and cyber services. They provide capabilities in areas such as cyber defense, cloud operations, engineering, and intelligence. With a commitment to solving complex challenges, Peraton integrates data-driven technologies to ensure mission success for their military and government clients.
đ„ 40 minutes ago
đșđž United States â Remote
đ” $104k - $166k / year
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor
đ» Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
đŒ Consulting
đ„ Healthcare
đŠ Logistics
Consulting âą Healthcare âą Logistics
Peraton is a mission-focused enterprise that supports national security initiatives through advanced IT and cyber services. They provide capabilities in areas such as cyber defense, cloud operations, engineering, and intelligence. With a commitment to solving complex challenges, Peraton integrates data-driven technologies to ensure mission success for their military and government clients.
âą Design, develop, and maintain reliability solutions and SRE utilities using Python in AWS environments âą Build automation scripts, APIs, and utilities in Python to reduce toil and improve platform reliability âą Implement observability and monitoring solutions using Grafana and AWS CloudWatch âą Build and optimize Terraform Infrastructure as Code for AWS resources and SRE solutions âą Develop CI/CD pipelines and automated testing âą Define SRE standards, best practices, guidelines, and metrics such as SLIs and SLOs âą Apply version control, code reviews, test-driven development, and documentation practices âą Participate in incident management and an on-call rotation âą Provide technical support for SRE tools and troubleshoot production issues âą Collaborate with teams to reduce incident recurrence through proactive detection and pattern analysis âą Stay current with AWS services, SRE methodologies, and cloud-native development technologies âą Collaborate with cross-functional teams in Agile and Scaled Agile frameworks âą Produce clear, blameless postmortems with actionable items and documented failure scenarios
âą Must be a U.S. Citizen with the ability to obtain and maintain the required Public Trust level Clearance âą Bachelor's Degree and 8 years of experience, or a High School diploma or equivalent and 12 years of experience âą 5+ years of advanced Python development experience building enterprise-grade, highly available tools, APIs, and utilities for AWS âą 7+ years of software development experience focused on reliability and platform engineering âą 3+ years of hands-on experience developing solutions in AWS environments âą Deep understanding of AWS services including EC2, VPC, S3, Lambda, IAM, CloudFormation, EventBridge, and Step Functions âą Experience with AWS resource cost optimization âą 3+ years applying SRE principles including observability, toil automation, SLIs/SLOs, and reliability engineering âą Expert-level proficiency with Terraform IaC, including module development and state management âą Strong experience with CI/CD pipelines, automated testing frameworks, and DevOps practices âą Experience with Grafana, AWS CloudWatch, and AWS Canary âą Experience defining, implementing, and managing SLOs/SLIs and error budgets âą Familiarity with conducting RCAs and producing postmortem documentation âą Working experience in Agile and Scaled Agile environments âą Familiarity with ITSM processes, resilience testing, and chaos engineering practices
âą Employees may be eligible for overtime âą Employees may be eligible for shift differential âą Employees may be eligible for a discretionary bonus âą Equal opportunity employer, including disability and protected veterans
Apply Nowđ„ 8 hours ago
DevOps Process Analyst optimizing development and operations workflows for SSI Group, a healthcare software company. Improving processes, reporting, risk management, and cross-functional delivery.
đ„ 10 hours ago
Senior SRE improving reliability, observability, and cloud infrastructure for CentralReachâs autism and IDD care software platforms. Driving SLO adoption, incident response, automation, and production performance.
đșđž United States â Remote
đ” $160k - $180k / year
đ° Private equity on 2018-03
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ„ 10 hours ago
Site Reliability Engineer operating AWS and Snowflake infrastructure for CentralReachâs autism and IDD care software. Improving reliability, connectivity, deployments, and incident response across data platforms.
đșđž United States â Remote
đ” $135k - $160k / year
đ° Private equity on 2018-03
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ„ 10 hours ago
DevOps Engineer II automating CI/CD, containers, and infrastructure for Navy readiness and training systems. Supporting DoD cybersecurity compliance across development, test, and production environments.
đșđž United States â Remote
đ” $115k - $125k / year
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ„ 11 hours ago
Senior Site Reliability Engineer building Laravelâs multi-region Kubernetes infrastructure and observability systems. Establishing SRE practices, SLOs, and automation across global developer products.