AI Engineer – LLM, VLM

Job not on LinkedIn

🔥 17 hours ago

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🗣️ LLM Engineer

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of SAI Group

SAI Group

51 - 200 employees

Founded 1997

📡 Telecommunications

⚡ Energy

🏗️ Construction

Telecommunications • Energy • Construction

SAI Group is a turnkey developer and contractor based in Salem, NH that provides wireless telecommunications, renewable energy, and construction services. The company offers full lifecycle site development for wireless networks — including RF engineering, DAS services, tower antenna/cable crews, electricians, installation, commissioning and optimization — and emphasizes in-house capabilities, quality control and schedule/cost management. SAI also develops renewable energy projects and infrastructure such as electric vehicle charging, solar power and energy storage solutions, and performs residential and commercial construction work. Founded in 1998, SAI highlights the ability to self-perform over 90% of development tasks and serves clients seeking integrated telecom and energy development services.

📋 Description

• Develop and deploy AI applications using LLMs and VLMs • Build RAG pipelines involving document ingestion, chunking, embeddings, retrieval, reranking, and generation • Work with models such as GPT, Claude, Gemini, Llama, Mistral, Qwen, and multimodal/VLM models • Develop multimodal solutions involving text, images, PDFs, charts, tables, and documents • Perform prompt engineering, supervised fine-tuning, LoRA/QLoRA, and model evaluation • Build AI agents and tool-calling workflows where appropriate • Optimize inference for latency, throughput, memory, and cost • Develop APIs and production services using Python, FastAPI, Docker, and cloud platforms • Implement evaluation frameworks to measure accuracy, hallucination, relevance, latency, and safety • Collaborate with ML engineers, software engineers, and product teams to take prototypes into production

🎯 Requirements

• Strong Python programming and software-engineering fundamentals • Hands-on experience with LLMs and/or VLMs • Strong understanding of Transformers, attention mechanisms, tokenization, embeddings, and inference • Experience with PyTorch and Hugging Face Transformers • Experience building RAG systems and vector-search solutions • Knowledge of prompt engineering and LLM evaluation • Experience with APIs, REST services, Git, Docker, and CI/CD • Familiarity with vector databases such as FAISS, Milvus, Pinecone, Weaviate, or pgvector • Understanding of cloud AI infrastructure, preferably AWS/Azure/GCP • Experience with multimodal models such as Qwen-VL, LLaVA, Gemini, GPT vision models, or similar • Understanding of image preprocessing and document/image understanding • Experience with OCR, document intelligence, image classification, object detection, or visual question answering is a plus • Ability to build pipelines combining vision + language + retrieval

Apply Now

Similar Jobs

🕒 August 25

Evnek

51 - 200

💼 Consulting

🤖 Artificial Intelligence

🤝 B2B

Forward Deployed Engineer delivering LLM, RAG, and cloud AI solutions for IT services clients. Owning discovery, client proposals, prototyping, and end-to-end delivery outcomes.

AWS

Azure

Cloud

Google Cloud Platform

🕒 July 27

QuantumLoopAi

51 - 200

🏥 Healthcare

🤖 Artificial Intelligence

☁️ SaaS

AI/ML & Prompt Engineer owning AI intelligence behind EMMA, a healthcare chatbot for patient access. Design and optimize LLM systems ensuring clinical safety in a remote role.

🇮🇳 India – Remote

💰 $2M Convertible note on 2025-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

🗣️ LLM Engineer

Azure

Docker

GraphQL

JavaScript

Microservices

MySQL

Next.js

Python

React

🕒 June 30

Weekday (YC W21)

11 - 50

💼 Consulting

👥 HR Tech

☁️ SaaS

Generative AI Engineer for Weekday's client responsible for developing AI agents. Remote role focusing on research and development in Generative AI.

Azure

Python

PyTorch

Scikit-Learn

SQL

Tensorflow

🕒 June 17

Prescience Decision Solutions (A Movate Company)

51 - 200

💼 Consulting

🏥 Healthcare

🏭 Manufacturing

GenAI Engineer designing, building, and deploying scalable AI solutions leveraging LLMs and agentic AI systems. Specializing in Data Science and Advanced Analytics for Fortune 500 companies.

AWS

Azure

Cloud

Docker

Google Cloud Platform

JavaScript

Kubernetes

Node.js

Python

React

SQL

TypeScript

🕒 May 7

Mactores

51 - 200

💼 Consulting

🏢 Enterprise

Generative AI Engineer at Mactores developing and deploying generative AI solutions using large language models. Collaborating with teams to implement AI-powered solutions for real-world problems.

AWS

Cloud

Python