
PhD Research Fellow β Agents, Monitoring, AI Reliability, Model Evaluation
Posted Jul 3

Posted Jul 3
This is a fully remote position, open to applicants in Brazil.
β’ Create execution flow maps for agents, pinpointing essential areas for log collection and monitoring.
β’ Construct Python pipelines to facilitate the collection, storage, and processing of execution logs.
β’ Establish and execute metrics to assess agents' performance, quality, and consistency.
β’ Develop automated evaluation routines (evals) for various scenarios and use cases.
β’ Generate analyses and reports on agent behavior, highlighting failures and potential areas for enhancement.
β’ Assist in the implementation of dashboards or monitoring tools.
β’ Document the architecture, metrics, and outcomes achieved throughout the project.
β’ Modality: TECH 4 (PhD level)
β’ Education: PhD
β’ Degree fields: Data Science, Statistics, Computer Engineering, Computer Science, or related disciplines.
β’ Comprehensive benefits package to support your health and well-being.
β’ Opportunities for professional development and career advancement.
β’ Collaborative and innovative work environment.
CVS Health
One Impression
Volga Partners
Mercor
Get handpicked remote jobs straight to your inbox weekly.