Forward Deployment Engineer, Generative AI

🕒 Maio 18

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Tiger Analytics

Tiger Analytics

1001 - 5000 funcionários

Fundada em 2011

🏥 Saúde

📦 Logística

📣 Marketing

Healthcare • Logistics • Marketing

A Tiger Analytics é uma consultoria líder em IA e analytics, especializada em aplicar ciência de dados e Machine Learning para gerar insights estratégicos de negócio em diversos setores. Oferecemos serviços de estratégia de dados, engenharia de IA e business intelligence, viabilizando a tomada de decisão orientada por dados e a transformação digital de nossos clientes. A Tiger Analytics colabora com parceiros de tecnologia de ponta, como Microsoft, Google Cloud e AWS, para entregar soluções de última geração. Atendemos a uma ampla gama de segmentos, incluindo bens de consumo embalados (CPG), saúde e serviços financeiros, ajudando as empresas a operacionalizar insights e a se diferenciar com tecnologias de IA e Machine Learning.

Descrição

• The Forward Deployment Engineer (FDE) drives the on-site deployment, integration, and scaling of our enterprise Generative AI solutions. • This role embeds directly within customer engineering teams to operationalize Large Language Models (LLMs) and retrieval systems across multi-cloud environments (AWS, Azure, GCP). • You will bridge the gap between AI research and production-grade cloud infrastructure. • You will collaborate with cross-functional teams and business partners and will have the opportunity to drive current and future strategy by leveraging your analytical skills as you ensure business value and communicate the results.

🎯 Requisitos

• AI Solution Deployment: Deploy, fine-tune, and optimize large-scale Gen AI models and LLM orchestration frameworks within customer cloud environments. • Infrastructure Engineering: Architect scalable infrastructure for AI workloads utilizing GPU/TPU orchestration, high-performance storage, and low-latency networking. • Data & Retrieval Pipelines: Design and implement high-throughput data ingestion pipelines and Vector Database architectures for Retrieval-Augmented Generation (RAG). • Multi-Cloud Management: Build agnostic, resilient cloud deployments across AWS, Azure, and GCP using Infrastructure as Code (IaC). • Technical Advocacy: Act as the primary technical consultant, guiding enterprise clients through AI safety, prompt engineering patterns, and inference cost optimization. • Product Collaboration: Feed edge-case deployment insights back to core AI research and platform engineering teams to improve product robustness. • Technical Requirements- AI Frameworks: Hands-on experience with LLM orchestration tools (LangChain, LlamaIndex, AutoGen) and deep learning frameworks (PyTorch, Hugging Face). • Vector Databases: Production experience setting up and querying vector stores (Milvus, Pinecone, Qdrant, Chroma, or pgvector). • Model Operations (LLMOps): Proficiency in model serving frameworks (vLLM, TGI, Triton Inference Server) and evaluation tools. • Cloud & Containers: Advanced knowledge of cloud AI primitives (AWS Bedrock/SageMaker, Azure OpenAI, GCP Vertex AI) and Kubernetes (K8s) for GPU workloads. • IaC & Automation: Mastery of Terraform or OpenTofu to provision complex multi-cloud compute environments. • Programming: Strong coding skills in Python (preferred) or Go, with an emphasis on writing clean, concurrent code. • Soft Skills- AI Consultation: Ability to manage customer expectations around LLM non-determinism, hallucinations, and performance trade-offs. • Rapid Adaptability: Passion for keeping pace with the weekly advancements in the Generative AI landscape. • Critical Debugging: Exceptional skill in isolating errors across complex software layers, from GPU drivers up to prompt engineering logic. • Mobility: Willingness to travel to client sites to lead high-stakes, on-site deployment sprints.

🏖️ Benefícios

• This position offers an excellent opportunity for significant career development in a fast-growing and challenging entrepreneurial environment with a high degree of individual responsibility. • Tiger Analytics provides equal employment opportunities to applicants and employees without regard to race, color, religion, age, sex, sexual orientation, gender identity/expression, pregnancy, national origin, ancestry, marital status, protected veteran status, disability status, or any other basis as protected by federal, state, or local law.

Candidatar-se

Vagas Similares

🕒 Maio 18

decircle

1 - 10

📣 Marketing

📦 Logística

💼 Consultoria

DevOps Engineer for M0, a stablecoin platform optimizing AWS infrastructure and CI/CD pipelines. Collaborating with product teams and ensuring security and performance of cloud-native applications.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 14

NVIDIA

10.000+ funcionários

🏥 Saúde

🏭 Manufatura

🤖 Inteligência Artificial

Senior Network Reliability Engineer maintaining NVIDIA's cloud and datacenter networks. Engaging in global support and driving operational improvements across teams.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 14

Avaya

5001 - 10000

💼 Consultoria

📣 Marketing

📦 Logística

Site Reliability Engineer at Avaya driving stability and performance across Azure and GCP platforms. Collaborating with DevOps and Security teams to manage incidents and optimize operations.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $129.000 - $143.000 / ano

💰 Post-IPO Debt em 2022-06

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 13

Wikimedia Foundation

501 - 1000

🤝 Sem Fins Lucrativos

📚 Educação

📱 Mídia

Senior Site Reliability Engineer with Wikimedia Foundation supporting platform for Wikipedia. Focus on operational tasks, collaboration, and continual improvement of infrastructure reliability.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $113.082 - $175.725 / ano

💰 $2.500.000 Grant em 2019-09

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 13

WEX

5001 - 10000

🏥 Saúde

📦 Logística

✈️ Turismo

SRE Architect driving AI-Powered Reliability Engineering strategy and enforcing enterprise-wide SRE standards. Overseeing the architecture and implementation of mission-critical systems for WEX.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $200.600 - $250.400 / ano

💰 $310.000.000 Post-IPO Debt em 2020-06

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório