
5001 - 10000 employees
đź Consulting
đĽ Healthcare
âď¸ Legal
Consulting ⢠Healthcare ⢠Legal
RWS Group is a leading provider of AI-powered language services and technology, specializing in translation and localization for a variety of industries. With a focus on enhancing global communication, RWS offers solutions that improve translation quality, manage multilingual content, and streamline the localization process. Their services cater to sectors such as aerospace, finance, legal, life sciences, and technology, leveraging advanced AI to ensure clients can effectively connect across languages and cultures.
đĽ 0 minutes ago
đ United Kingdom, Netherlands, +4 more countries â Remote
â° Full Time
đ Senior
đ¤ AI Engineer
đť Ghost score 10%
Improve your chances of getting an interview by checking your resume score before you apply.

5001 - 10000 employees
đź Consulting
đĽ Healthcare
âď¸ Legal
Consulting ⢠Healthcare ⢠Legal
RWS Group is a leading provider of AI-powered language services and technology, specializing in translation and localization for a variety of industries. With a focus on enhancing global communication, RWS offers solutions that improve translation quality, manage multilingual content, and streamline the localization process. Their services cater to sectors such as aerospace, finance, legal, life sciences, and technology, leveraging advanced AI to ensure clients can effectively connect across languages and cultures.
⢠Design and architect core platform components and evaluation systems ⢠Make technical decisions and maintain accountability for reliability, scalability, and long-term maintainability ⢠Set technical direction for building, evaluating, and deploying AI capabilities across the company ⢠Define a scalable platform vision beyond the immediate team ⢠Design reusable abstractions, SDKs, and services for model integration, prompt management, experimentation, and deployment ⢠Define company-wide AI evaluation strategies and methodologies, including automated metrics, human-in-the-loop workflows, test-set management, and benchmarking ⢠Build production-ready evaluation frameworks and developer tooling ⢠Establish AI observability standards for quality, performance, cost, and regression signals ⢠Build dashboards and reporting that turn signals into actionable decisions ⢠Drive testing discipline, reproducibility, sound experimental design, and statistically defensible model-quality measurement ⢠Provide technical leadership on ambiguous, high-impact problems ⢠Mentor engineers and raise standards through code and design reviews ⢠Contribute to model and system governance, documentation, dataset and test-set versioning, reproducibility, and responsible-AI checks ⢠Codify best practices into tooling and standards adopted by hundreds of developers ⢠Track LLM, evaluation research, and AI tooling developments ⢠Prototype, de-risk, and shepherd emerging techniques from experiment to supported platform capability ⢠Champion adoption of new platform capabilities across teams ⢠Partner with research, product, and localization leaders to align evaluation with customer needs ⢠Influence roadmap and technical strategy across engineering and product stakeholders ⢠Gather developer requirements and represent them in platform direction ⢠Communicate technical direction, trade-offs, and quality standards to technical and non-technical audiences
⢠Significant software engineering experience, typically 5+ years, building and operating production systems, tools, libraries, or services used by many engineers ⢠Excellent API design, reliability, and developer experience ⢠Experience with CI/CD and cloud infrastructure ⢠Proficiency in Python and/or another general-purpose language ⢠Strong testing discipline ⢠Hands-on experience building with LLMs or other ML systems, including prompt engineering, fine-tuning, retrieval, or model integration ⢠Understanding of LLM/ML failure modes and tradeoffs ⢠Proven experience designing and leading evaluation for AI/ML systems ⢠Experience defining metrics and methodology, building evaluation pipelines, managing test sets, and rigorously assessing model quality and regressions ⢠Strong command of accuracy, precision, recall, F1, automated evaluation, human evaluation, statistical significance, and methodological limitations ⢠Excellent written and verbal communication ⢠History of influencing technical direction across teams and mentoring engineers ⢠Comfort with ambiguity and ability to scope, prioritize, and sequence high-impact work with limited direction ⢠Preferred: experience evaluating NLP, machine translation, or content-generation systems using COMET, chrF++, BLEU, MetricX, or MQM-style human evaluation ⢠Preferred: experimentation and observability tooling, data/test-set versioning, and rigorous benchmarking workflows ⢠Preferred: AI governance and documentation including model cards, system cards, reproducibility, and responsible-AI considerations ⢠Preferred: familiarity with modern LLM ecosystems, orchestration frameworks, and vector stores ⢠Preferred: experience supporting multilingual or localization-focused enterprise products
⢠Equal employment opportunity and inclusive, discrimination-free work environment ⢠Opportunities to grow as an individual and excel in your career ⢠Work with global teams and experts in AI, data, content, language, and localization
Apply Nowđ 2 days ago
Senior AI Engineer building multi-agent AI platforms for Spyrosoft, a software engineering company. Designing secure, observable LLM systems with TypeScript, Go and Python.
đŹđ§ United Kingdom â Remote
đľ ÂŁ70k / year
â° Full Time
đ Senior
đ¤ AI Engineer
đŹđ§ UK Skilled Worker Visa Sponsor
đ 2 days ago
AI engineering leader guiding Blend360, a data consultancy, across production AI architecture and major client engagements. Setting technical standards, developing engineering leaders, and scaling the AI practice.
đŹđ§ United Kingdom â Remote
đ° $100M Private Equity Round on 2022-08
â° Full Time
đ Senior
đ¤ AI Engineer
đ 3 days ago
Applied AI Engineer building reliable agentic infrastructure for Dwelly, an AI-powered residential lettings operator. Automating agency workflows through LLM orchestration, evaluation, observability, and production systems.
đŹđ§ United Kingdom â Remote
đ° Seed on 2024-05
â° Full Time
đĄ Mid-level
đ Senior
đ¤ AI Engineer
đ 6 days ago
Senior Applied AI Engineer shipping customer-facing LLM and RAG features for Pleoâs spend-management platform. Owning evaluation, observability, and production deployment across engineering and data teams.
đŹđ§ United Kingdom â Remote
đ° $42.9M Debt Financing - Pleo on 2024-04
â° Full Time
đ Senior
đ¤ AI Engineer
đ 6 days ago
AI compiler engineer optimizing kernel generation and computational graphs for NVIDIA GPUs. Advancing compiler technology powering AI inference, training, and datacenter deployments.