Senior ML / AI Engineer

Job not on LinkedIn

🔥 22 hours ago

🇺🇸 United States – Remote

💵 $150k - $210k / year

⏰ Full Time

🟠 Senior

🤖 AI Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Nonstop Administration & Insurance Services

Nonstop Administration & Insurance Services

51 - 200 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Nonstop Administration & Insurance Services is a company that provides innovative health insurance solutions designed to lower costs and improve access to healthcare for both employers and employees. They focus on offering first-dollar coverage, reducing or eliminating out-of-pocket expenses like copays and premiums. Their solutions are tailored for organizations with 50 to 500 employees, providing financial predictability and significant savings on premiums through a high-tech administration platform. Nonstop also provides tools for real-time claims data and financial reporting, aiding employers in managing healthcare benefits effectively. They position themselves as an ally for employers, brokers, and employees, aiming to make healthcare more accessible and affordable.

📋 Description

• Design and ship agentic systems using dynamic tool-calling agents, structured Pydantic outputs, governed tool catalogs, per-step verification, grounding checks, LLM-as-judge, human-in-the-loop, and bounded auditable loops • Build the RAG layer, including document ingestion, chunking, embeddings, hybrid retrieval, reranking, access scoping, and injection/poisoning defenses • Operate self-hosted inference at scale with vLLM and embedding/rerank services on EKS GPU nodes • Optimize inference throughput using continuous batching, prefix caching, quantization, and KV-cache tuning • Enforce fair-share concurrency across tenants • Own the MLOps and governance plane, including offline evaluation harnesses, golden sets, quality scoring, canary/gray releases, auto-rollback, cost/budget governance, and observability • Make systems production-safe through durable state, circuit breakers, retries, PHI-safe logging/redaction, RBAC, fail-closed defaults, and horizontal AWS scalability via Terraform • Extend the platform to new domains from problem framing through governed, evaluated, deployed services • Mentor engineers on building initiatives on the shared platform foundation • Maintain clean, typed, tested code and participate in thoughtful design reviews • Balance autonomy and determinism for high-stakes tasks

🎯 Requirements

• 5+ years building and shipping ML/AI systems in production, including hands-on LLM application work in the last 1–2 years • Strong Python (typed, tested, production-grade) and solid software-engineering fundamentals • Comfortable across an async web service, a data layer, and infrastructure • Practical depth in prompting, structured/function-calling outputs, RAG, embeddings, vector search, retrieval quality, agent/tool-use loops, grounding, hallucination control, eval sets, and guardrails • Cloud and MLOps experience on AWS or equivalent, including containers, Kubernetes, Terraform, CI/CD, observability, and model-serving cost/performance tuning • Track record of owning reliability, including state durability, failure handling, scaling, and debugging production incidents • Clear written communication and judgment under ambiguity • Experience self-hosting/optimizing open-weight models such as vLLM or TGI on GPUs • Experience with embeddings/rerank serving, such as bge or Infinity • Experience with LangGraph, LangChain, or comparable agent frameworks and multi-agent orchestration • Regulated-data experience with HIPAA, SOC 2, ISO 27001, PHI/PII handling, RBAC, and audit trails • pgvector/Postgres, MongoDB/DocumentDB, SQS, and Cognito/OIDC • OpenTelemetry, Prometheus/Grafana, and frontend development with React/TypeScript • Amazon Bedrock or a multi-provider abstraction • On-premises or air-gapped deployment experience • Healthcare/insurance/claims domain experience, including X12 835, EOB, or benefits, is a strong plus • Must be authorized to work for any employer in the United States; this job does not offer visa sponsorship

🏖️ Benefits

• 100% coverage of medical, dental and vision benefits for employees and dependents • 401k match up to 4%

Apply Now

Similar Jobs

🔥 22 hours ago

ExpertVoice

51 - 200

📣 Marketing

🛒 Retail

🛍️ eCommerce

Frontend/AI Engineer building React interfaces and LLM-powered features for ExpertVoice, which helps consumer brands engage and reward influential product experts. Prototyping AI tools across frontend and backend integrations.

🔥 22 hours ago

Kognitiv Inc.

201 - 500

💼 Consulting

☁️ SaaS

🤝 B2B

Principal AI consultant architecting autonomous agents and Generative AI solutions for Kognitiv, a Workday partner. Leading executive workshops, technical roadmaps, compliance, and AI consulting engagements.

🔥 22 hours ago

FICO

1001 - 5000

💼 Consulting

🛡️ Insurance

🏥 Healthcare

AI engineering leader advancing FICO's global analytics software platform with LLM-powered fraud detection and decision automation. Leading teams, model governance, monitoring, and scalable responsible AI systems.

🕒 Yesterday

NewRocket

501 - 1000

💼 Consulting

🏥 Healthcare

🛡️ Insurance

ServiceNow AI engineer building intelligent applications, integrations, and workflows for NewRocket’s enterprise ServiceNow consulting clients. Developing automation, AI features, and platform solutions across ITSM, ITOM, CSM, and HRSD.

🕒 Yesterday

SmartLogic

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

Senior Staff Engineer leading AI adoption, technical standards, and cloud architecture. Delivering custom software projects for SmartLogic’s consultancy clients.