Senior DevOps Engineer

Job not on LinkedIn

🔥 1 minute ago

🇦🇷 Argentina – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Aiphoria

Aiphoria

51 - 200 employees

Founded 2022

💼 Consulting

📦 Logistics

📣 Marketing

💰 $34M Series A - Aiphoria on 2025-07

Consulting • Logistics • Marketing

Aiphoria is a provider of AI-driven virtual employees and a proprietary platform that automates multi-channel customer-facing and back-office tasks. Its Aiphoria Pros are multilingual, multi-modal agents (voice, chat, email, calls) that handle support, sales, collections, legal, HR and marketing workflows for enterprise customers in industries such as banking, telecom and e‑commerce. The company offers cloud and on‑prem deployments, analytics and open-architecture integrations to reduce human labor, improve response times and boost customer satisfaction.

📋 Description

• Deploy, operate, and evolve a microservices-based platform running in Kubernetes clusters across AWS, GCP, and on-premises Rancher • Operate and support GPU-based ML inference services using Triton Inference Server and vLLM on RunPod, Scaleway, and Nebius • Build and maintain Docker images for all microservices and ensure a stable service lifecycle • Maintain and scale development and production Kubernetes clusters • Participate in deployment debugging, incident investigation, and performance troubleshooting • Develop, maintain, and evolve custom Helm charts for each service • Design and operate CI/CD pipelines using GitHub and GitLab for on-premises customer deployments • Ensure platform compliance with SOC 2 requirements and improve security and compliance processes • Manage cluster access via NetBird VPN and implement role-based access control using group policies • Deploy and manage infrastructure using Terraform and Ansible • Develop and continuously improve observability systems using Grafana, Prometheus, and the ELK stack • Continuously optimize infrastructure across IaC, IAM, observability, and CI/CD • Work with Python, Kubernetes, Linux, Docker, GitHub CI/CD, PostgreSQL, ClickHouse, Kafka, Superset, Terraform, and Ansible

🎯 Requirements

• Minimum 5 years of experience in a DevOps and/or Site Reliability Engineering role • Strong hands-on experience with Linux system administration • Extensive experience deploying, operating, and scaling Kubernetes in both cloud and bare-metal environments • Deep expertise and practical experience with at least one major cloud provider, preferably Google Cloud Platform • Experience with ML inference on GPU/CPU is a strong plus • Proven experience implementing SRE practices and building observability stacks using Grafana, Prometheus, and Loki • Strong adherence to GitOps, Infrastructure as Code (IaC), and CI/CD principles • Advanced expertise in Terraform, Ansible, and Python • Ability to work in high-uncertainty environments and rapidly learn new technologies and patterns • Proactive mindset and ability to debug and understand the product beyond DevOps tasks • Strategic thinking for selecting technologies and architectural approaches based on long-term goals

🏖️ Benefits

• Fully remote • 21 vacation days + public holidays + 5 sick days • Private English lessons via Preply • Fast career progression • Startup pace with enterprise stability — real clients, real revenue, no bureaucracy • Cutting-edge tech stack • High engineering bar and real ownership • Opportunity to work on award-winning AI products • Exposure to Speech Technologies, NLP, Generative AI, LLMs, diffusion models, and voice-first agentic architecture

Apply Now

Similar Jobs

🕒 5 days ago

Stefanini LATAM

10,000+ employees

💼 Consulting

📦 Logistics

📣 Marketing

Senior DevOps designing CI/CD, cloud, container, and infrastructure automation. Supporting Stefanini, a global IT services company, with secure and scalable software delivery.

🗣️🇪🇸 Spanish Required

Cloud

DNS

Docker

Grafana

Java

JavaScript

Kubernetes

Linux

Node.js

OpenStack

Prometheus

Python

Spring

Spring Boot

SpringBoot

🕒 6 days ago

Talan

1001 - 5000

💼 Consulting

🏢 Enterprise

☁️ SaaS

Senior DevSecOps Engineer building secure AWS/EKS infrastructure for Talan, an international technology and business transformation consultancy. Driving Kubernetes, Terraform, CI/CD, IAM and observability practices.

AWS

Cloud

Docker

Java

Kubernetes

Linux

Python

Terraform

🕒 6 days ago

Azumo

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Cloud DevOps Engineer owning Kubernetes, CI/CD, observability, and secure infrastructure for Azumo’s production AI systems. Supporting client deployments across Latin America.

AWS

Azure

Cloud

DynamoDB

Google Cloud Platform

Kubernetes

Linux

MongoDB

Postgres

Python

Terraform

Go

🕒 6 days ago

Coderio

201 - 500

💼 Consulting

📦 Logistics

📣 Marketing

Performance & Observability Engineer configurando y analizando pruebas de rendimiento para soluciones bancarias. Coderio desarrolla soluciones digitales escalables para empresas globales.

🗣️🇪🇸 Spanish Required

AWS

Grafana

Jenkins

Kafka

Kubernetes

OpenShift

Prometheus

RabbitMQ

ServiceNow

🕒 September 23

Allata

201 - 500

🚘 Automotive

🏥 Healthcare

📦 Logistics

DevOps Engineer automating AWS and Azure infrastructure at Allata, a global AI consulting and technology services firm. Improving CI/CD, observability, reliability, and cloud operations.

AWS

Azure

Cloud

Docker

Python

Terraform