
51 - 200 employees
💼 Consulting
🏥 Healthcare
📦 Logistics
Consulting • Healthcare • Logistics
Azumo is a leading software development company that specializes in nearshore services. The company offers a range of solutions, including software development, dedicated teams, staff augmentation, and virtual CTO services. Azumo is particularly known for its expertise in artificial intelligence, mobile app development, data engineering, and cloud services. The company prides itself on delivering high-quality, scalable, and innovative software solutions tailored to the specific needs of various industries, including fintech, game development, healthcare, and media. With a focus on assembling talented developers from Latin America, Azumo ensures time zone alignment and seamless communication with clients in North America. Their commitment to quality and client satisfaction is evidenced by numerous awards and positive client testimonials from prestigious organizations like Facebook, Twitter, and Discovery Channel.
🔥 0 minutes ago
🌐 Argentina, Brazil, +3 more countries – Remote
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
👻 Ghost score 14%
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
💼 Consulting
🏥 Healthcare
📦 Logistics
Consulting • Healthcare • Logistics
Azumo is a leading software development company that specializes in nearshore services. The company offers a range of solutions, including software development, dedicated teams, staff augmentation, and virtual CTO services. Azumo is particularly known for its expertise in artificial intelligence, mobile app development, data engineering, and cloud services. The company prides itself on delivering high-quality, scalable, and innovative software solutions tailored to the specific needs of various industries, including fintech, game development, healthcare, and media. With a focus on assembling talented developers from Latin America, Azumo ensures time zone alignment and seamless communication with clients in North America. Their commitment to quality and client satisfaction is evidenced by numerous awards and positive client testimonials from prestigious organizations like Facebook, Twitter, and Discovery Channel.
• Own production infrastructure for AI systems, including clusters, deployment pipelines, and monitoring • Provision, upgrade, network, and configure production Kubernetes clusters • Build reproducible infrastructure as code with Terraform or equivalent and detect infrastructure drift • Build and maintain reliable CI/CD pipelines and standardized container image workflows • Implement monitoring and alerting, investigate incidents, determine root causes, and make preventive changes • Provide infrastructure for AI workloads, including inference services and their scaling and cost profiles • Analyze infrastructure costs and capacity needs, including projected costs at higher traffic levels • Secure Linux, Kubernetes, containers, and service meshes, including secrets, access, and audit trails • Develop internal tooling and documentation and support developers and QA during release cycles • Work within client environments, repositories, cloud accounts, and change processes when required • Operate within Azumo's SOC 2-certified environment and accommodate engagement requirements such as HIPAA • Use AI-assisted engineering tools and automated codebase audits to assess security, cost, and architecture findings
• 5+ years as a DevOps, SRE or systems engineer running production infrastructure • Linux administration, networking, Git, and scripting in Bash plus Python or Go • Production Kubernetes experience, including provisioning, upgrading and debugging clusters • Infrastructure as code using Terraform or equivalent, including state, modules and environment parity • Experience building and maintaining CI/CD pipelines using GitHub Actions, GitLab or equivalent • Responsibility for container images • Cloud deployment experience on AWS, Azure or GCP, including managed Kubernetes services and associated cost models • Monitoring and incident response experience using Datadog, CloudWatch or equivalent • Experience taking incidents from alert through root cause analysis to preventive changes • Security hardening of Linux, containers and Kubernetes • Infrastructure cost and capacity management experience • Active use of AI-assisted coding tools such as Claude Code, Cursor, or GitHub Copilot in delivery work • Clear written and spoken English at C1 or above • Ability to explain technical trade-offs directly to a client • Bachelor's degree in Computer Science, a related field, or equivalent professional experience • Preferred: service mesh and traffic management experience with Istio, Linkerd or equivalent • Preferred: Helm, Kustomize or equivalent templating and environment configuration • Preferred: production-scale database operations with PostgreSQL, MongoDB, RDS, DynamoDB or equivalent • Preferred: running or scaling inference workloads and evaluating cost and latency trade-offs against hosted APIs • Preferred: delivery under compliance regimes such as SOC 2 or HIPAA • Preferred: cloud certifications, open-source infrastructure contributions, or published technical writing
• 100% remote-first culture (work anywhere in Latin America) • Paid time off (PTO) • U.S. Holidays • Solid AI Training and certification • Mentored career development • Profit sharing • $US remuneration • Maternity coverage
Apply Now🔥 13 hours ago
Performance & Observability Engineer configurando y analizando pruebas de rendimiento para soluciones bancarias. Coderio desarrolla soluciones digitales escalables para empresas globales.
🗣️🇪🇸 Spanish Required
AWS
Grafana
Jenkins
Kafka
Kubernetes
OpenShift
Prometheus
RabbitMQ
ServiceNow
🕒 Yesterday
DevOps Engineer automating AWS and Azure infrastructure at Allata, a global AI consulting and technology services firm. Improving CI/CD, observability, reliability, and cloud operations.
AWS
Azure
Cloud
Docker
Python
Terraform
🕒 6 days ago
Technical Operations Engineer maintaining reliable infrastructure, integrations, and data for a subscription CPG e-commerce agency. Automating observability, testing, and incident response with AWS, Terraform, and Datadog.
AWS
Cypress
JavaScript
MongoDB
NoSQL
Python
SQL
Terraform
TypeScript
🕒 September 11
DevOps Engineer maintaining Talan’s Fixed Income pricing and distribution platform. Supporting Linux, containers, AWS, automation, CI/CD, and observability for Talan’s technology consulting clients.
🇦🇷 Argentina – Remote
💰 Secondary Market on 2020-07
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Cloud
Docker
Grafana
Java
Jenkins
Linux
Prometheus
.NET
🕒 September 4
DevOps Engineer Senior gestionando CI/CD, Kubernetes e infraestructura multi-cloud. Automatización, observabilidad y seguridad para un cliente estadounidense.
🗣️🇪🇸 Spanish Required
AWS
Cloud
Docker
Google Cloud Platform
Kubernetes
Python
Terraform