Lead DevOps Engineer

🔥 2 minutes ago

🏄 California – Remote

infoinfo

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Sphera

Sphera

1001 - 5000 employees

Founded 1989

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Sphera is a leading provider of enterprise sustainability management software, data, and consulting services. It offers solutions that enable organizations in sectors like Chemicals, Oil & Gas, Industrials, and Financial Services to manage their sustainability goals effectively. Sphera's offerings include the SpheraCloud platform, which focuses on Environment, Health, Safety & Sustainability (EHS&S) management, operational compliance, and risk management. The company's software and consulting services are designed to help organizations improve their safety, mitigate risks, reduce costs, and build resilience. By providing an integrated 360-degree view of sustainability and performance management, Sphera assists companies in meeting regulatory reporting requirements and achieving their sustainability objectives.

📋 Description

• Provide technical leadership for DevOps and platform engineering across supported products, prioritizing the core platform and AI initiatives • Own and improve CI/CD pipelines with integrated code-quality and security scanning • Configure repository and pipeline permissions and enforcement controls • Design, build, and maintain cloud deployment automation, including auto-scaling, blue/green and canary delivery, and one-click rollback • Implement and manage Terraform infrastructure and use Azure CLI for emergency and ad-hoc operations • Build, deploy, and manage Docker workloads on Azure Container Apps, Azure Web Apps, and Function Apps • Operate Azure API Management, Elasticsearch, and Databricks for reliability, security, performance, and cost efficiency • Build and operate infrastructure and pipelines for AI, machine learning, and data workloads, including MLOps • Partner with Engineering and Security on network and firewall changes and compliance • Develop self-service platform capabilities and internal developer tooling or golden paths • Automate system administration, provisioning, configuration, maintenance, and disaster recovery • Build observability with the NOC team using metrics, logs, traces, and alerting • Embed security and compliance into pipelines and infrastructure through DevSecOps automation, auditing, and tooling • Guide critical projects, solve complex problems, and mentor team members • Partner with FinOps on cloud and AI resource cost optimization • Establish and maintain business-focused KPIs and SLOs for operational excellence

🎯 Requirements

• Excellent communication and collaboration skills across multiple disciplines and levels • Demonstrated technical leadership or mentorship on complex, cross-team initiatives • Strong analytical skills with a proven ability to drive reliability and performance improvements • Proven success building and optimizing CI/CD for an Azure-based SaaS application, including SonarCloud and Black Duck scanning • Experience configuring and enforcing repository and pipeline permissions • Hands-on experience with Azure DevOps and/or GitHub Actions • Strong Terraform infrastructure-as-code experience • Proficiency with Azure CLI • Strong scripting skills in PowerShell, Bash, or Python • Experience with Docker on Azure Container Apps, Azure Web Apps, and Function Apps • Experience operating Azure API Management, Databricks, and Elasticsearch • Experience managing network and firewall configurations and change-request processes • Strong knowledge of Microsoft/Windows and Linux server environments • Working knowledge of version control and Git-based workflows • Experience within an agile software development lifecycle • Strong understanding of security principles and secure-by-design practices • Working knowledge of Microsoft SQL Server and IIS configuration and management • Working knowledge of local and wide-area networking, including switches, routers, firewalls, VPNs, VNets, and application gateways • Bachelor’s degree in Computer Science or a similar area of study, or equivalent practical experience • Preferred: experience supporting AI/ML or generative-AI workloads in production, including MLOps, model deployment and serving, and monitoring • Preferred: familiarity with Azure Machine Learning, Azure OpenAI, or common ML frameworks • Preferred: advanced Databricks and Elasticsearch tuning and scaling • Preferred: GPU-based compute, vector databases, RAG architectures, internal developer platforms, Kubernetes, observability stacks, additional PaaS and delivery tooling, data platforms, and relevant Azure certifications

🏖️ Benefits

• Equal Opportunity Employer • Inclusive environment committed to diversity • Opportunity to work on AI, machine learning, data, and cloud infrastructure initiatives • Mentoring and professional growth through guiding and growing team members

Apply Now

Similar Jobs

🔥 5 minutes ago

CACI International Inc

10,000+ employees

🎖️ Defense

🏛️ Government

🔒 Cybersecurity

Cloud/DevOps specialist securing AWS assessment infrastructure for CACI’s federal cybersecurity program. Building IaC, CI/CD, and authorized cloud-exploitation capabilities for continuous federal cyber assessments.

🇺🇸 United States – Remote

💵 $90.3k - $189.6k / year

🔥 Funding within the last year

💰 $500M Post-IPO Debt on 2026-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 1 hour ago

MaintainX

501 - 1000

☁️ SaaS

🏭 Manufacturing

🏢 Enterprise

Site Reliability Engineer improving reliability, observability, and developer autonomy for MaintainX’s industrial work execution platform. Building tooling and standards for resilient, self-service operations.

🇺🇸 United States – Remote

💵 $120k - $249.3k / year

💰 $150M Series D on 2025-08

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 2 hours ago

Rimutee

11 - 50

🤝 B2B

👥 HR Tech

Senior DevOps Engineer designing AWS cloud infrastructure, IaC, CI/CD, and disaster recovery for ReKluti’s international software clients. Remote, full-time role focused on secure, highly available platforms.

🗣️🇪🇸 Spanish Required

🔥 8 hours ago

IREN

201 - 500

🤖 Artificial Intelligence

🤝 B2B

⚡ Energy

Azure DevOps Lead Engineer owning Azure architecture, governance, and reliability for IREN’s renewable-powered AI cloud and data centers. Leading DevOps teams and compliance automation.

🔥 17 hours ago

CACI International Inc

10,000+ employees

🎖️ Defense

🏛️ Government

🔒 Cybersecurity

Senior DevOps Engineer operating secure OCI infrastructure and CI/CD automation for CACI’s federal Oracle Fusion Cloud HCM modernization program. Supporting compliant, resilient environments across federal agencies.

🇺🇸 United States – Remote

💵 $105.1k - $231.1k / year

🔥 Funding within the last year

💰 $500M Post-IPO Debt on 2026-02

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)