
10,000+ employees
🏥 Healthcare
💼 Consulting
📦 Logistics
Healthcare • Consulting • Logistics
Thermo Fisher Scientific is a leading global supplier of scientific instrumentation, reagents and consumables, and software services. They support the life sciences, healthcare, and analytical chemistry sectors by providing robust solutions for laboratory research and production processes. Their innovative products and services encompass a range of applications, including diagnostics, lab workflow automation, and drug discovery.
🕒 July 3
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
🏥 Healthcare
💼 Consulting
📦 Logistics
Healthcare • Consulting • Logistics
Thermo Fisher Scientific is a leading global supplier of scientific instrumentation, reagents and consumables, and software services. They support the life sciences, healthcare, and analytical chemistry sectors by providing robust solutions for laboratory research and production processes. Their innovative products and services encompass a range of applications, including diagnostics, lab workflow automation, and drug discovery.
• Design simulation environments and digital twins for enterprise workflows • Post-train LLM agents using RLHF, DPO, GRPO, PPO, and emerging methods • Build pipelines that convert human-labeled traces and verifiable signals into training data • Architect multi-turn, tool-using agents with closed learning loops • Design reward functions and verifiers that resist reward hacking and reflect real task outcomes • Set the technical bar across the team — architecture, code review, engineering standards • Mentor researchers and engineers; drive technical direction through influence • Translate research into production; contribute to publications
• 7+ years in ML/AI research or engineering; 3+ years at senior/staff level • MS or PhD in Computer Science, Machine Learning, or related field (or equivalent) • 5+ years hands-on RL — environment design, reward engineering, policy optimization — with at least one production deployment LLM Post-Training • 3+ years fine-tuning LLMs with hands-on RL post-training (RLHF, DPO, GRPO, PPO) • Expert-level implementation of RLHF pipelines, reward modeling (Bradley-Terry), DPO, and KTO • Strong Python and software engineering skills — comfortable building production pipelines, not just notebooks • Deep expertise in MDPs, policy gradient methods (PPO, SAC), and temporal difference learning • Working knowledge of modern post-training and rollout-serving libraries (TRL, veRL, OpenRLHF, SkyRL)
• Health insurance • 401(k) matching • Flexible work hours • Paid time off • Remote work options
Apply Now🕒 July 1
Applied Scientist developing AI-driven solutions with a focus on deep learning and innovative architectures. Collaborating within a diverse team to leverage AI research for practical applications in business intelligence.
🇺🇸 United States – Remote
💵 $197.3k - $313.7k / year
⏰ Full Time
🔴 Lead
🧬 Research Scientist
🦅 H1B Visa Sponsor
🕒 July 1
Principal Scientist at Natera focusing on oncology translational research and integrating scRNA-seq and spatial biology into pipelines. Leading advanced projects with a team of scientists.
🇺🇸 United States – Remote
💵 $171.9k - $214.9k / year
⏰ Full Time
🔴 Lead
🧬 Research Scientist
🦅 H1B Visa Sponsor
🕒 July 1
Staff Applied Scientist leading R&D work on AI/ML models for grocery replenishment technology. Apply knowledge in machine learning and optimization to reduce food waste across global supply chains.
🇺🇸 United States – Remote
💵 $191.8k - $287.6k / year
⏰ Full Time
🔴 Lead
🧬 Research Scientist
🦅 H1B Visa Sponsor
🕒 July 1
Staff Applied Scientist responsible for end-to-end data science engine focusing on insurance vertical. Driving revenue efficiency and collaborating with stakeholders at Launch Potato.
🕒 June 26
Principal Scientist leading immunology-driven research for Jade Biosciences' pipeline. Integrating insights to enhance clinical development for autoimmune therapies.