Research Scientist – Interactive Avatars

🔥 2 minutes ago

🇪🇺 Europe – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🧬 Research Scientist

👻 Ghost score 11%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Synthesia

Synthesia

501 - 1000 employees

Founded 2017

📣 Marketing

💼 Consulting

📦 Logistics

🔥 Funding within the last year

💰 $200M Series E - Synthesia on 2025-10

Marketing • Consulting • Logistics

<Synthesia> Synthesia is a SaaS AI video platform that enables businesses to create studio-quality videos without cameras, microphones, actors, or studios by using AI avatars and synthetic voiceovers. The platform supports 160+ languages, one-click translation/localization, an AI screen recorder, brand management, collaboration and analytics, and enterprise-grade security (SOC 2 Type II, GDPR). It’s marketed primarily to teams and enterprises for training, sales enablement, marketing, knowledge management and internal communications, helping companies scale video production while reducing time and cost.

📋 Description

• Join a team of 40+ researchers and engineers in R&D working on generative AI, avatar-centric interactive video diffusion models • Contribute to research direction for dyadic interaction modeling and own well-scoped research problems end to end • Advance the perceptual layer of interactive agents by understanding user audio and video and generating contextually appropriate reactions • Post-train multimodal models to generate natural dyadic interactions from user audio and video inputs • Adapt diffusion models to conditioning signals such as conversational state, turn-taking, and listener cues • Build evaluation frameworks and test suites for tracking interaction quality • Partner with the data team to define data needs and shape high-quality datasets • Run rigorous experiments and share findings that inform technical decisions

🎯 Requirements

• Strong ML background and hands-on experience with diffusion models, ideally for video or avatar generation • Publications at top-tier venues such as CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, or SIGGRAPH in world models, dyadic interaction, or video diffusion, or equivalent demonstrated impact • Experience taking research ideas through to working implementations • Proficiency in PyTorch and modern ML tooling for large-scale training • Clear communication of hypotheses, experiments, and results • Experience with real-time or streaming generation, including autoregressive video diffusion (nice to have) • Distillation or other techniques for low-latency inference (nice to have) • Audio-driven facial, gesture, or full-body motion modeling (nice to have) • Conversational modeling, such as turn-taking, backchanneling, or listener-response generation (nice to have) • Experience mentoring students or junior researchers (nice to have) • Ability to work legally in the country of employment without visa sponsorship, as indicated in the application form

🏖️ Benefits

• Build production-scale video foundation models in a fast-growing Generative AI company • Work on human-centric video generation with real-world impact • Tackle hard problems in scaling, stability, and controllability • Influence the direction of next-generation synthetic human technology • Highly technical, high-ownership environment where work ships

Apply Now