Search Remote Jobs

Cloud / DevOps Engineer, Infra & IaC

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

💵 $75 - $110 / hour

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 20%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Weekday (YC W21)

Weekday (YC W21)

11 - 50 employees

Founded 2021

💼 Consulting

👥 HR Tech

☁️ SaaS

Consulting • HR Tech • SaaS

Weekday is a modern recruitment platform that combines AI technologies with a vast database of potential candidates, aiming to streamline the hiring process for companies in India. They offer various services, including a proactive outreach approach that helps employers connect with top talent, as well as tools for candidates to easily apply for jobs. Weekday's emphasis on candidate engagement through multiple channels, including email, WhatsApp, and phone calls, sets it apart in the competitive landscape of recruitment agencies.

📋 Description

• Collaborate with research and engineering teams to identify knowledge gaps and improve AI model performance across cloud infrastructure, DevOps, Kubernetes, and Infrastructure-as-Code domains • Design realistic and technically challenging tasks covering Kubernetes troubleshooting, AWS service integration, infrastructure automation, and production operations • Develop accurate, detailed reference solutions for complex infrastructure engineering scenarios • Review and evaluate AI-generated technical solutions for correctness, reliability, scalability, security, and adherence to production best practices • Provide clear, structured written feedback highlighting technical gaps, incorrect assumptions, and opportunities for improvement • Create detailed evaluation criteria, rubrics, and benchmarks for assessing Kubernetes troubleshooting, IaC architecture, AWS integrations, and CI/CD reasoning • Develop scenarios involving cluster failures, infrastructure automation, deployment workflows, service integrations, and operational reliability • Work closely with other technical subject matter experts to maintain consistency, accuracy, and quality across evaluation datasets • Translate practical production experience into structured guidance that can be used to improve AI-generated infrastructure solutions • Establish high-quality standards for AI systems working with complex Cloud, DevOps, Kubernetes, AWS, and IaC problems • Transform practical engineering knowledge into structured tasks, reference solutions, evaluation frameworks, and technical feedback to improve next-generation AI models

🎯 Requirements

• 4+ years of professional experience in Cloud Infrastructure, DevOps, Site Reliability Engineering, Platform Engineering, or a closely related field • Strong hands-on experience managing Kubernetes in production environments, including diagnosing, troubleshooting, and resolving cluster failures and operational issues • Experience with Kubernetes beyond simply writing manifests or consuming managed Kubernetes control planes • Proven production experience with Infrastructure-as-Code, particularly Terraform and/or AWS CDK • Strong practical knowledge of AWS cloud services, including production integration with AWS Lambda, API Gateway, and DynamoDB • Experience designing, implementing, and maintaining CI/CD pipelines for production workloads • Strong understanding of cloud architecture, infrastructure automation, deployment strategies, observability, reliability, and operational best practices • Demonstrated career progression with increasing ownership and responsibility in infrastructure, DevOps, or platform engineering • Ability to commit reliably to 40 hours per week during standard weekdays • Excellent written and verbal communication skills, with the ability to explain complex technical concepts and engineering decisions clearly • Strong analytical and troubleshooting abilities, particularly when diagnosing distributed systems and infrastructure failures • Experience working with large-scale cloud infrastructure or highly distributed systems • Familiarity with Kubernetes networking, security, storage, scaling, and cluster lifecycle management • Experience implementing infrastructure security and reliability best practices • Knowledge of AWS architecture patterns and cloud-native application design • Experience with GitOps, containerization, monitoring, logging, and observability platforms • Familiarity with modern DevOps and platform engineering methodologies • Experience reviewing or evaluating technical documentation, engineering solutions, or AI-generated outputs

Apply Now

Similar Jobs

🔥 2 hours ago

Hyatt

10,000+ employees

🍽️ Food & Beverage

✈️ Travel

🛒 Retail

Network DevOps Engineer automating Hyatt’s global hospitality network infrastructure. Building IaC, CI/CD pipelines, and multi-vendor network testing systems.

🔥 4 hours ago

Guild Mortgage

1001 - 5000

🏗️ Construction

💸 Finance

🏠 Real Estate

DevOps Manager leading engineers, CI/CD, cloud platforms, and infrastructure automation for Guild Mortgage’s mortgage banking operations. Driving secure, predictable technology delivery across enterprise initiatives.

🔥 6 hours ago

Leidos

10,000+ employees

🏥 Healthcare

💼 Consulting

📦 Logistics

DevSecOps Engineer securing Leidos defense systems and JADC2 infrastructure. Automating CI/CD, cyber remediation, and compliant systems administration for DoD mission platforms.

🔥 6 hours ago

Garner Health

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior SRE owning AWS, Kubernetes, and Terraform reliability for Garner’s AI-powered healthcare platform. Automating operations, leading incident response, and upholding HIPAA compliance.

🔥 6 hours ago

Skimmer

11 - 50

☁️ SaaS

🤝 B2B

⚡ Productivity

Senior DevOps Engineer building Azure infrastructure, CI/CD, and observability for Skimmer’s pool-service platform. Owning deployment reliability for a new product line.