Senior ML Specialist

Posted Sep 9

This is a fully remote position, open to applicants in Brazil.

πŸ“‹ Description

β€’ Take ownership of model strategy across the platform by assessing and selecting models for each task.

β€’ Utilize benchmarks and LLM-as-judge frameworks with statistically robust analysis to inform model decisions.

β€’ Develop and uphold evaluation methodologies, which include offline evaluation sets, judge calibration, regression benchmarks, and quality metrics.

β€’ Create classification and fine-tuned models for routing, categorization, and detection tasks.

β€’ Construct datasets, as well as train, evaluate, and deploy models.

β€’ Design, develop, and deliver agentic AI products, which encompass agent identities, reusable skills, tool integrations, and structured outputs.

β€’ Execute prompt and context engineering, which includes assembling system prompts, injecting live-data context, managing threads/memory, and producing Pydantic structured outputs.

β€’ Ensure output quality using deterministic enforcers, validation quality gates, and LLM-as-judge evaluators.

β€’ Identify hallucinations, inconsistencies, and model-version drift, implementing solutions through prompts, retrieval, guardrails, or model selection.

β€’ Build and enhance agent tools that query Databricks, MongoDB, and other external systems.

β€’ Connect agents with the larger ecosystem via the Model Context Protocol (MCP).

β€’ Manage deployed services using OpenTelemetry, monitoring quality, latency, and cost in Grafana.

β€’ Work collaboratively with data engineers, software engineers, and product teams on comprehensive AI solutions.

β€’ Elevate the team's machine learning fundamentals through reviews and knowledge sharing.

β€’ Rigorously assess research literature and apply validated methodologies.

β€’ Document and maintain the codebase, ensuring high standards of code quality and best practices.


⛳️ Requirements

β€’ Over 5 years of practical experience in machine learning, including training, evaluating, and deploying models.

β€’ Strong foundation in statistics and probability, encompassing hypothesis testing, confidence intervals, sampling, sample-size reasoning, bias/variance, and calibration.

β€’ Profound understanding of LLM and generative AI principles, including transformers, attention mechanisms, tokenization, embeddings, training pipelines, decoding strategies, scaling behaviors, and potential failure modes.

β€’ Direct experience with classical machine learning and natural language processing, including classification, clustering, feature engineering, text classification, embeddings, and semantic similarity.

β€’ Proficient in fine-tuning models using LoRA/PEFT, supervised fine-tuning, or full fine-tuning; capable of constructing training datasets.

β€’ Experience in designing AI evaluation methodologies, including offline evaluation sets, LLM-as-judge, inter-rater agreement, regression benchmarks, and significance testing.

β€’ Production-level Python proficiency, including PyTorch, scikit-learn, Hugging Face, modern asynchronous Python, FastAPI, Pydantic, testing, and continuous integration.

β€’ Practical experience with at least one major LLM provider API: OpenAI, Anthropic, or Google, including structured outputs and tool/function calling.

β€’ A Master's or PhD in Machine Learning, Statistics, Computer Science, Mathematics, or a related quantitative discipline, or equivalent demonstrated expertise through publications, competition results, released models, or research code.

β€’ Ability to work autonomously at a senior level, taking ownership of problems end to end.

β€’ Preferred: Publications or contributions to NLP/ML; a record of competitions on Kaggle or equivalent platforms.

β€’ Preferred: Experience with retrieval systems, agent evaluation, model cost/latency optimization, Databricks, Azure, Docker, observability stacks, MCP, and 0-to-1 product delivery.


🏝️ Benefits

β€’ An equal opportunity employer dedicated to diversity, inclusion, and a sense of belonging.

β€’ PJ contract arrangement.

People also viewed

Budderfly1 day ago

Service Excellence Lead HVAC Technician

US flagNorth Carolina OnlyFull-timeUncategorized$24 – $38/hour
ApplyView job
St. Croix Hospice1 day ago

Per Diem RN, Triage

US flagIllinois, +8 more statesPart-timeUncategorized$420/year
ApplyView job
GuestReady1 day ago

Owner Success Associate

BR flagBrazil OnlyFreelanceUncategorized
ApplyView job
Academy of Art University School of Game Development1 day ago

Curriculum Systems Specialist

US flagUnited States OnlyFull-timeUncategorized$22 – $27/hour
ApplyView job
Airbnb1 day ago

Disaster Response Coordinator

GB flagUnited Kingdom OnlyFull-timeUncategorizedΒ£40k – Β£45k/year
ApplyView job
Carrington Holding Company, LLC1 day ago

Loan Servicing Specialist

US flagUnited States OnlyFull-timeUncategorized$20 – $22/hour
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers