
AI Research Engineer β Pre-training, LLM, Multi-Modal
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in Spain.
β’ Facilitate foundational pre-training for LLMs and Multi-Modal models on extensive, distributed servers featuring multi-node setups and thousands of NVIDIA GPUs.
β’ Create, prototype, and expand innovative architectures, tokenizers, and cross-modal alignment layers.
β’ Source, filter, and curate extensive textual and multi-modal datasets.
β’ Execute experiments both independently and collaboratively, analyze results, and refine training methodologies to achieve optimal performance.
β’ Investigate, debug, and resolve bottlenecks affecting model efficiency and stability during training sessions.
β’ Contribute to the enhancement of distributed training systems to promote scalability and hardware efficiency.
β’ A degree in Computer Science or a related discipline.
β’ Preference for a PhD in NLP, Machine Learning, or a related area, supported by a strong record in AI R&D (including notable publications in A* conferences).
β’ Practical experience in contributing to large-scale LLM or Multi-Modal pre-training operations on extensive, distributed servers equipped with thousands of NVIDIA GPUs.
β’ Familiarity and hands-on experience with large-scale, distributed training frameworks, libraries, and tools.
β’ Profound understanding of cutting-edge transformer and non-transformer modifications aimed at improving intelligence, efficiency, and scalability.
β’ Strong proficiency in PyTorch and Hugging Face libraries, with hands-on experience in model development, continuous pretraining, and deployment.
β’ Health insurance
β’ Remote work options
Cerence Inc.
Tether.to
Cotiviti
Next Step Systems - National IT Recruiting Firm
Get handpicked remote jobs straight to your inbox weekly.