
Machine Learning Engineering Intern
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in Germany, +1 more state.
• Develop product features for AI Agents as a member of the engineering team
• Ensure that evaluations accurately represent human judgment
• Create datasets and frameworks to evaluate the quality of AI agents
• Annotate and curate high-quality “golden datasets” derived from real-world conversations
• Design and enhance LLM-as-Judge evaluation logic within the Braintrust platform utilizing Python
• Analyze automated evaluation scores in comparison to human feedback
• Refine evaluation prompts and logic to enhance agent quality and performance
• Collaborate across both the back-end and front-end of the Zendesk Explore platform
• Currently a student
• Proficient in at least one of the following languages: Python, TypeScript, or JavaScript
• Sincere interest in Large Language Models (LLMs), Natural Language Processing (NLP), and the evaluation of AI agents
• Driven and eager to comprehend intricate AI systems
• Capable of articulating thoughts clearly and engaging in active listening during discussions
• Confident in working with multiple languages across the front-end and back-end of a SaaS application
• Open to learning new technologies and expanding knowledge and skills
• Familiarity with AWS and other cloud services (preferred)
• Experience with back-end data pipelines (preferred)
• Background knowledge and/or experience in the analytics industry (preferred)
• Fully flexible remote work arrangement
• Option to work remotely for part of the week
• Inclusive work environment
• Reasonable accommodations available for applicants with disabilities and disabled veterans
SumerSports
Airbnb
The Home Depot
Get handpicked remote jobs straight to your inbox weekly.