Cloud DevOps Engineer

🔥 0 minutes ago

🌐 Argentina, Brazil, +3 more countries – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 14%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Azumo

Azumo

51 - 200 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Azumo is a leading software development company that specializes in nearshore services. The company offers a range of solutions, including software development, dedicated teams, staff augmentation, and virtual CTO services. Azumo is particularly known for its expertise in artificial intelligence, mobile app development, data engineering, and cloud services. The company prides itself on delivering high-quality, scalable, and innovative software solutions tailored to the specific needs of various industries, including fintech, game development, healthcare, and media. With a focus on assembling talented developers from Latin America, Azumo ensures time zone alignment and seamless communication with clients in North America. Their commitment to quality and client satisfaction is evidenced by numerous awards and positive client testimonials from prestigious organizations like Facebook, Twitter, and Discovery Channel.

📋 Description

• Own production infrastructure for AI systems, including clusters, deployment pipelines, and monitoring • Provision, upgrade, network, and configure production Kubernetes clusters • Build reproducible infrastructure as code with Terraform or equivalent and detect infrastructure drift • Build and maintain reliable CI/CD pipelines and standardized container image workflows • Implement monitoring and alerting, investigate incidents, determine root causes, and make preventive changes • Provide infrastructure for AI workloads, including inference services and their scaling and cost profiles • Analyze infrastructure costs and capacity needs, including projected costs at higher traffic levels • Secure Linux, Kubernetes, containers, and service meshes, including secrets, access, and audit trails • Develop internal tooling and documentation and support developers and QA during release cycles • Work within client environments, repositories, cloud accounts, and change processes when required • Operate within Azumo's SOC 2-certified environment and accommodate engagement requirements such as HIPAA • Use AI-assisted engineering tools and automated codebase audits to assess security, cost, and architecture findings

🎯 Requirements

• 5+ years as a DevOps, SRE or systems engineer running production infrastructure • Linux administration, networking, Git, and scripting in Bash plus Python or Go • Production Kubernetes experience, including provisioning, upgrading and debugging clusters • Infrastructure as code using Terraform or equivalent, including state, modules and environment parity • Experience building and maintaining CI/CD pipelines using GitHub Actions, GitLab or equivalent • Responsibility for container images • Cloud deployment experience on AWS, Azure or GCP, including managed Kubernetes services and associated cost models • Monitoring and incident response experience using Datadog, CloudWatch or equivalent • Experience taking incidents from alert through root cause analysis to preventive changes • Security hardening of Linux, containers and Kubernetes • Infrastructure cost and capacity management experience • Active use of AI-assisted coding tools such as Claude Code, Cursor, or GitHub Copilot in delivery work • Clear written and spoken English at C1 or above • Ability to explain technical trade-offs directly to a client • Bachelor's degree in Computer Science, a related field, or equivalent professional experience • Preferred: service mesh and traffic management experience with Istio, Linkerd or equivalent • Preferred: Helm, Kustomize or equivalent templating and environment configuration • Preferred: production-scale database operations with PostgreSQL, MongoDB, RDS, DynamoDB or equivalent • Preferred: running or scaling inference workloads and evaluating cost and latency trade-offs against hosted APIs • Preferred: delivery under compliance regimes such as SOC 2 or HIPAA • Preferred: cloud certifications, open-source infrastructure contributions, or published technical writing

🏖️ Benefits

• 100% remote-first culture (work anywhere in Latin America) • Paid time off (PTO) • U.S. Holidays • Solid AI Training and certification • Mentored career development • Profit sharing • $US remuneration • Maternity coverage

Apply Now

Similar Jobs

🔥 13 hours ago

Coderio

201 - 500

💼 Consulting

📦 Logistics

📣 Marketing

Performance & Observability Engineer configurando y analizando pruebas de rendimiento para soluciones bancarias. Coderio desarrolla soluciones digitales escalables para empresas globales.

🗣️🇪🇸 Spanish Required

AWS

Grafana

Jenkins

Kafka

Kubernetes

OpenShift

Prometheus

RabbitMQ

ServiceNow

🕒 Yesterday

Allata

201 - 500

🚘 Automotive

🏥 Healthcare

📦 Logistics

DevOps Engineer automating AWS and Azure infrastructure at Allata, a global AI consulting and technology services firm. Improving CI/CD, observability, reliability, and cloud operations.

AWS

Azure

Cloud

Docker

Python

Terraform

🕒 6 days ago

Paired

1 - 10

💼 Consulting

🎯 Recruiter

👥 HR Tech

Technical Operations Engineer maintaining reliable infrastructure, integrations, and data for a subscription CPG e-commerce agency. Automating observability, testing, and incident response with AWS, Terraform, and Datadog.

AWS

Cypress

JavaScript

MongoDB

NoSQL

Python

SQL

Terraform

TypeScript

🕒 September 11

Talan

1001 - 5000

💼 Consulting

🏢 Enterprise

☁️ SaaS

DevOps Engineer maintaining Talan’s Fixed Income pricing and distribution platform. Supporting Linux, containers, AWS, automation, CI/CD, and observability for Talan’s technology consulting clients.

Ansible

AWS

Cloud

Docker

Grafana

Java

Jenkins

Linux

Prometheus

.NET

🕒 September 4

Ecosistemas

501 - 1000

💼 Consulting

📣 Marketing

🏥 Healthcare

DevOps Engineer Senior gestionando CI/CD, Kubernetes e infraestructura multi-cloud. Automatización, observabilidad y seguridad para un cliente estadounidense.

🗣️🇪🇸 Spanish Required

AWS

Cloud

Docker

Google Cloud Platform

Kubernetes

Python

Terraform