Search Remote Jobs

Senior Site-Reliability Engineer

Job not on LinkedIn

🔥 0 minutes ago

⚔️ Virginia – Remote

infoinfo

đź’µ $105k - $140k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź‘» Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Entarian

Entarian

1001 - 5000 employees

Founded 1993

🚀 Aerospace

🎖️ Defense

🏛️ Government

Aerospace • Defense • Government

Entarian is a diversified engineering and federal-technology solutions firm (formerly ERT) that delivers end-to-end science, engineering, and mission-support capabilities across space, defense, and civilian federal markets. The company provides satellite mission operations and ground systems, space and Earth science data processing, enterprise and digital engineering (including data analytics and AI/ML), network and command-and-control integration, and layered cyber defense to support resilient, secure government missions. Entarian works with U. S. federal agencies and defense partners on programs that include satellite calibration/validation, search-and-rescue operations, weather forecasting support, and IT modernization; it recently rebranded after acquiring Sev1Tech and is Macquarie Capital–backed.

đź“‹ Description

• Design, implement, and maintain scalable infrastructure using Infrastructure as Code practices • Develop and maintain automation scripts using PowerShell, Python, Ruby, and other scripting languages for OS provisioning, configuration management, and operational tasks • Implement and manage Terraform, Puppet, and/or Chef across hybrid environments • Monitor system health, performance, and availability using industry-standard tools and practices • Establish and enforce SLAs, SLOs, and error budgets for production services • Participate in on-call rotation and respond to incidents with a focus on rapid restoration and root cause analysis • Collaborate with development teams to improve deployment pipelines and release processes • Document operational procedures, runbooks, and architectural decisions • Conduct post-mortem reviews and implement corrective actions to prevent recurrence • Ensure reliability, availability, and performance of Windows-based production environments • Bridge development and operations to deliver highly available services while maintaining operational excellence

🎯 Requirements

• 5+ years of experience in Systems Administration, DevOps, or Site-Reliability Engineering roles • Strong expertise in Windows Server environments (2016+), including Active Directory, IIS, and MS SQL • Strong cloud skills; AWS experience preferred • Advanced proficiency in scripting, including module development and integration with REST APIs • Hands-on experience with Terraform for infrastructure provisioning • Hands-on experience with Puppet or Chef for configuration management • Experience with monitoring and observability platforms such as Prometheus, Grafana, Datadog, or New Relic • Solid understanding of networking concepts including DNS, TCP/IP, load balancing, and VPN • Strong problem-solving skills and ability to troubleshoot complex issues across multiple technology layers • Bachelor’s degree in Computer Science, Information Technology, or related field, or equivalent professional experience (desired qualification) • Certifications such as AWS Solutions Architect, Microsoft Certifications, or HashiCorp Certified: Terraform Associate • Experience with Docker and Kubernetes • Familiarity with GitLab Pipelines, Jenkins, or GitHub Actions • Knowledge of security best practices and compliance frameworks • Experience with ELK Stack or Splunk • Security clearance must be clearable

🏖️ Benefits

• Medical, dental, and vision insurance • Life, AD&D, and disability insurance • Paid time off • 11 company holidays • 401(k) retirement plan with company matching • Additional employee benefits and wellness resources

Apply Now

Similar Jobs

🔥 5 minutes ago

URBN (Urban Outfitters, Anthropologie Group, Free People & Nuuly)

10,000+ employees

👥 B2C

đź›’ Retail

đź‘— Fashion

DevOps Engineer scaling Nuuly’s GCP, Kubernetes, and Kafka infrastructure for its fashion rental platform. Automating deployments, improving reliability, and optimizing event-driven systems.

🔥 1 hour ago

Innosphere

51 - 200

đź’Ľ Consulting

📦 Logistics

📣 Marketing

Senior Site Reliability Engineer building reliable AWS infrastructure and CI/CD systems. Supporting Innosphere’s distributed technology staffing and software development teams.

🔥 1 hour ago

Harris Computer

10,000+ employees

🏥 Healthcare

đź’Ľ Consulting

📦 Logistics

Platform & DevSecOps Architect building automated CI/CD, security, and AI infrastructure for STChealth’s public-health software. Enabling compliant delivery across Kubernetes and legacy systems.

🔥 1 hour ago

Harris Computer

10,000+ employees

🏥 Healthcare

đź’Ľ Consulting

📦 Logistics

DevSecOps Engineer securing CI/CD pipelines, cloud infrastructure, containers, and vulnerability management. Supporting STChealth’s technology platform for immunization data exchange and public health.

🔥 1 hour ago

Capgemini

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Mainframe DevOps consultant guiding Capgemini clients through Endevor-to-DBB/Git/IDD migrations. Installing, troubleshooting, validating applications, and delivering technical training.