
Staff Research Engineer β Multimodal Generative Modelling
Posted Jul 17

Posted Jul 17
This is a fully remote position, open to applicants in United Kingdom.
β’ Define our strategic roadmap to introduce new model capabilities and enhance functionality for our clients, both in the short and long term.
β’ Suggest innovative multi-modal system architectures, particularly focusing on text and voice integration.
β’ Create and assess streaming and conversational systems aimed at achieving low-latency, interactive voice-video synthesis.
β’ Develop solutions that enhance emotional expressiveness and facilitate natural interactions.
β’ Execute and realize designs, overseeing the process from pretraining to post-training phases.
β’ Deploy models into production with optimized runtimes to serve our customers and subsequently address their feedback.
β’ In-depth knowledge of generative modeling, preferably in the context of sequential or multimodal data.
β’ Practical experience with large language models or comparable transformer-based architectures.
β’ Advanced proficiency in PyTorch, including distributed training and model optimization techniques.
β’ A comprehensive understanding of time-series modeling and tokenization, ideally within audio, speech, or video domains.
β’ Demonstrated experience in training deep learning models from start to finish, encompassing data preparation to evaluation.
β’ Strong foundational skills in general software engineering.
β’ Health insurance
β’ Flexible working arrangements
β’ Professional development opportunities
β’ Equipment allowances
SecurityScorecard
Oregon Health & Science University Foundation
LiveKit
Netflix
Get handpicked remote jobs straight to your inbox weekly.