Search Remote Jobs

Senior DevOps Engineer/Site Reliability Engineer

đź•’ June 2

🗽 New York – Remote

infoinfo

đź’µ $165k - $215k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

đź‘» Ghost score 7%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Stellar Cyber

Stellar Cyber

51 - 200 employees

đź”’ Cybersecurity

🤖 Artificial Intelligence

🏢 Enterprise

đź’° $38M Series B on 2021-11

Cybersecurity • Artificial Intelligence • Enterprise

Stellar Cyber is a company that provides an automation-driven security operations platform seamlessly integrating next-generation SIEM, Network Detection and Response (NDR), and Open Extended Detection and Response (XDR). Their platform utilizes advanced AI to quickly detect and correlate cybersecurity threats across various security tools, offering comprehensive threat intelligence and automated incident response to enhance security operations for enterprises, MSSPs, and MSPs. With a focus on reducing security operation costs and improving threat response times, Stellar Cyber helps organizations protect their entire attack surface, including on-premises, cloud, and IT/OT environments.

đź“‹ Description

• Administer and maintain Kubernetes clusters and containerized workloads. • Manage cloud infrastructure across OCI, AWS, GCP, or Azure environments. • Develop and maintain CI/CD pipelines for reliable application deployments. • Implement and manage Infrastructure as Code (IaC) using Terraform and Helm. • Build automation tooling and operational workflows using Python, Go, or Bash. • Drive observability initiatives including monitoring, logging, tracing, and alerting improvements. • Monitor, troubleshoot, and resolve production incidents while participating in on-call rotations. • Support and optimize distributed data platforms including Kafka, Elasticsearch, Spark, Redis, and MongoDB. • Improve platform reliability, scalability, and operational efficiency using SRE best practices. • Collaborate with cross-functional teams across multiple time zones. • Perform Linux system administration and networking troubleshooting. • Contribute to incident response processes, postmortems, and reliability improvements. • Support GitOps and deployment workflows using tools such as ArgoCD and GitHub Actions. • Evaluate and implement AI-assisted operational tooling for auto-remediation, alert correlation, and operational intelligence.

🎯 Requirements

• 5+ years of experience in DevOps, SRE, or Platform Engineering roles. • Strong expertise with Kubernetes, Docker, and container orchestration. • Hands-on experience managing production cloud environments. • Strong Infrastructure as Code experience with Terraform and Helm. • Experience with CI/CD tools and deployment automation. • Advanced troubleshooting skills in Linux systems, networking, and distributed systems. • Experience with observability platforms including Prometheus, Grafana, Loki, Alertmanager, and Elastic Stack. • Strong programming and scripting skills in Python, Bash, or Go. • Experience supporting high-availability production systems and on-call operations. • Knowledge of incident management and reliability engineering practices. • Familiarity with data platform technologies such as Kafka, Spark, Elasticsearch, Redis, or MongoDB. • Understanding of AI-driven operational tooling and automated remediation concepts. • Excellent communication, collaboration, and problem-solving skills. • Resides on the East Coast

🏖️ Benefits

• Pre-IPO Stock Options • Medical, Dental & Vision care • 401(k) • Employee Assistance Program • Employee Discount Program • Life Insurance • Paid time off • Referral Program • Rewards and Recognition Program

Apply Now

Similar Jobs

đź•’ June 2

Agilent Technologies

10,000+ employees

🍽️ Food & Beverage

🏥 Healthcare

đź’Ľ Consulting

DevOps Software Engineer designing and maintaining CI/CD pipelines and cloud infrastructure for Agilent’s CrossLab Connect team. Supporting application development and optimizing deployment processes.

đź•’ June 1

Wikimedia Foundation

501 - 1000

🤝 Non-profit

📚 Education

📱 Media

Senior Site Reliability Engineer at Wikimedia Foundation maintaining scalable infrastructure for API services. Collaborating with engineering teams to ensure high reliability and performance.

đź•’ June 1

Trivelta

201 - 500

đź’Ľ Consulting

📦 Logistics

🎲 Gambling

Senior DevOps Engineer at Trivelta, specializing in CI/CD pipelines and cloud infrastructure automation. Leading development in a rapidly scaling social-first gaming technology company.

đź•’ May 30

Quevera

51 - 200

đź’Ľ Consulting

📦 Logistics

🎖️ Defense

Platform Engineer/DevOps Engineer responsible for deploying and maintaining OpenShift environments at Quevera. Join a people-first technology company supporting mission-critical customers.

đź•’ May 30

Zealogics Inc

501 - 1000

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

DevOps Engineer maintaining build systems and improving site reliability for multinational team with Fortune 500 clients. Collaborating with software teams and enforcing good development practices.