
Research Scientist β Interactive Avatars
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Europe.
β’ Become part of a team comprising over 40 researchers and engineers in R&D focusing on generative AI and avatar-centric interactive video diffusion models.
β’ Play a pivotal role in shaping research direction for dyadic interaction modeling while managing well-defined research projects from inception to completion.
β’ Enhance the perceptual layer of interactive agents by analyzing user audio and video to generate contextually relevant responses.
β’ Perform post-training on multimodal models to create natural dyadic interactions based on user audio and video inputs.
β’ Modify diffusion models to incorporate conditioning signals including conversational state, turn-taking, and listener cues.
β’ Develop evaluation frameworks and testing suites to monitor interaction quality.
β’ Collaborate with the data team to identify data requirements and cultivate high-quality datasets.
β’ Conduct thorough experiments and disseminate findings that guide technical decision-making.
β’ Solid machine learning foundation and practical experience with diffusion models, preferably in video or avatar generation.
β’ Publications in prestigious venues like CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, or SIGGRAPH regarding world models, dyadic interaction, or video diffusion, or equivalent demonstrated impact.
β’ Proven ability to transform research ideas into functional implementations.
β’ Expertise in PyTorch and contemporary ML tools for large-scale training.
β’ Ability to clearly articulate hypotheses, experimental designs, and results.
β’ Experience with real-time or streaming generation, including autoregressive video diffusion (preferred).
β’ Familiarity with techniques for low-latency inference, such as distillation (preferred).
β’ Experience in audio-driven facial, gesture, or full-body motion modeling (preferred).
β’ Knowledge of conversational modeling techniques, including turn-taking, backchanneling, or listener-response generation (preferred).
β’ Experience in mentoring students or junior researchers (preferred).
β’ Eligibility to work legally in the country of employment without visa sponsorship, as indicated in the application form.
β’ Develop production-scale video foundation models within a rapidly growing Generative AI organization.
β’ Engage in human-centric video generation projects with meaningful real-world implications.
β’ Address challenging issues related to scaling, stability, and controllability.
β’ Shape the future of next-generation synthetic human technology.
β’ Work in a highly technical environment that emphasizes ownership and enables delivery of impactful work.
Humana
Index Analytics LLC
Surgo Health
Lucas James Talent Partners
Get handpicked remote jobs straight to your inbox weekly.