Lead Data Platform Engineer

Vaga não está no LinkedIn

🕒 Junho 26

🗽 New York – Remoto

info

💵 $125.000 - $174.333 / ano

⏰ Tempo Integral

🟠 Sênior

🏗️ Engenheiro de Plataforma

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Coupa Software

Coupa Software

1001 - 5000 funcionários

Fundada em 2006

💼 Consultoria

📦 Logística

🏥 Saúde

Consulting • Logistics • Healthcare

A Coupa Software é uma fornecedora líder de soluções de gestão de gastos corporativos (Business Spend Management — BSM). Sua plataforma foca em otimizar e transformar os gastos diretos e indiretos em Procurement (Compras), Finanças, Cadeia de Suprimentos e TI. A Coupa utiliza inteligência artificial (IA) e insights de dados em larga escala para impulsionar eficiências de custos, gerenciar relacionamentos com fornecedores e mitigar riscos. Com produtos que abrangem áreas como faturamento, pagamentos, gestão de despesas e colaboração na cadeia de suprimentos, a Coupa atende a uma ampla gama de setores, incluindo automotivo, saúde, varejo e outros. Sua comunidade e seu ecossistema de parceiros abrangentes permitem que as organizações descubram economias ocultas e melhorem a conformidade (compliance), promovendo crescimento e resiliência em um cenário econômico em constante mudança.

Descrição

• Manage end-to-end **Data pipeline **(ETL jobs) within agreed SLAs. • Manage AWS core and **big data services** (S3, IAM, EMR, Redshift, etc..). • Running applications in containers (ECS, Docker). • Lead Day 2 operational lifecycle for ML and GenAI infrastructure. This includes designing, deploying, and maintaining high-availability production LLM serving platforms, implementing automated scaling, self-healing, and infrastructure-as-code patterns. Focus on proactive reliability, model performance observability, and continuous cost optimization for high-compute AI workloads. • Collaborate closely with our product development and engineering teams to create AI-driven features. • Drive cloud operations consistency by automating platform maintenance, standardizing infrastructure configurations (IaC), and implementing robust release management processes to minimize drift across multi-cloud environments. • Manage AWS infrastructure using code (Terraform, Chef, etc..). • Administering applications running in Linux operating system. • Enable application and system monitoring for better observability. • Application and infrastructure support for ETL jobs and data pipelines including participating in an on-call rotation for after-hours emergencies. • Collaborate with platform and Dev teams to plan and deploy product releases and patch Linux/ECS clusters. • Ability to participate in design reviews, code reviews, and troubleshooting incidents. • Ability to operate in a high-pressure environment and troubleshoot complex issues quickly while successfully handling multiple priorities. • Ability to record, write, and review RCAs.

🎯 Requisitos

• Bachelor's Degree and at least 8+ years of experience managing Big Data technologies and Data Pipelines. • Sound knowledge and experience in Linux administration and troubleshooting. • 5+ years of experience in managing cloud infrastructure and platforms, such as AWS and Azure. • Familiar with the current engineering landscape in the generative AI space and have a strong interest in AI and related technologies. • Strong expertise in MLOps and production-grade LLM operations. Proven track record in managing high-availability model inference clusters, automating model lifecycle management, and implementing advanced observability (latency, throughput, and error rate monitoring) specifically for AI workloads. • Have Bash or Python scripting experience. • Experience with containerization, Amazon ECS, EKS/ Azure AKS. • Experience with tools like Chef, Ansible, Jenkins, Rundeck, or equivalent. • Experience with source control systems such as Git and operating in complex branching strategies. • Experience with Infrastructure as Code products like Terraform, helm charts. • Good understanding of DNS and Load balancers setup and troubleshooting. • Experience in Big Data platforms/Data lakes and managing Business Intelligence tools (like looker..). • Knowledge in ApacheSpark architecture and troubleshooting Java applications. • Basic understanding of MySQL Server and general database knowledge. • Excellent written and verbal communication with a passion for solving the problem. • Confidence in your ability to own and deliver projects and issues to resolution on your own & can think and act globally. • Deep experience in Day 2 cloud operations, including automated incident remediation, capacity planning, and managing large-scale production cloud environments with a focus on performance and reliability.

Candidatar-se

Vagas Similares

🕒 Junho 26

SitusAMC

5001 - 10000

💼 Consultoria

📦 Logística

🏠 Imobiliário

Platform Engineer specializing in CI/CD automation and cloud technologies within SitusAMC's Cloud development team. Requires extensive experience with AWS, Kubernetes, and Release automation.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $135.000 - $165.000 / ano

💰 Private equity em 2020-05

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

🏗️ Engenheiro de Plataforma

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 26

Cytora

51 - 200

💼 Consultoria

📦 Logística

🤖 Inteligência Artificial

Senior Platform Engineer developing Cytora’s serverless Underwriting Productivity Suite for insurance applications. Collaborating with platform team on infrastructure and CI/CD pipelines.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $100.000 - $140.000 / ano

💰 Series B em 2019-04

⏰ Tempo Integral

🟠 Sênior

🏗️ Engenheiro de Plataforma

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 25

EvenUp

51 - 200

💼 Consultoria

🏥 Saúde

⚖️ Jurídico

Senior Frontend Engineer leading development of scalable AI systems at EvenUp, a vertical SaaS company. Collaborating with cross-functional teams to deliver quality software solutions.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $190.000 - $249.000 / ano

⏰ Tempo Integral

🟠 Sênior

🏗️ Engenheiro de Plataforma

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 25

NVIDIA

10.000+ funcionários

🏥 Saúde

🏭 Manufatura

🤖 Inteligência Artificial

Senior HPC Support Engineer handling AI hardware and software solutions for NVIDIA's compute and GPU platforms. Resolving sophisticated customer issues and maintaining a high level of customer happiness.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 25

nDeavour Consulting

1 - 10

💼 Consultoria

📦 Logística

📣 Marketing

Senior Data Engineer building analytical data platform on GCP for Mobile Wave Solutions. Collaborating on AI features and data governance for reliable data consumption.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

🏗️ Engenheiro de Plataforma

🗣️🇺🇸🇬🇧 Inglês obrigatório