
Member of Technical Staff, Enterprise AI
Posted Sep 15

Posted Sep 15
This is a fully remote position, open to applicants in New York.
• Integrate into enterprise AI workflows as a technical research partner.
• Collaborate with domain specialists and enterprise teams to gain insights into real-world system behavior.
• Identify, formalize, and prioritize failure modes arising from deployed AI systems.
• Convert operational challenges into structured research inquiries and measurable technical issues.
• Generate analyses regarding system behavior, limitations, and potential areas for enhancement.
• Create high-quality datasets aimed at addressing model and system vulnerabilities.
• Develop evaluation protocols, quality standards, and structured assessment frameworks.
• Detect shortcomings in current datasets and evaluation comprehensiveness.
• Conduct rapid experimental cycles to test hypotheses and measure system enhancements.
• Construct and benchmark agentic workflows for robustness, reliability, and scalability.
• Assess AI systems functioning within intricate enterprise workflows.
• Analyze experimental outcomes and ascertain the significance and reproducibility of improvements.
• Refine datasets, evaluations, and system configurations based on research insights.
• Create lightweight tools for evaluation, data curation, experimentation, and swift iteration.
• Collaborate across research, engineering, product, domain, and enterprise-facing teams.
• Convert research findings into clear, decision-oriented recommendations.
• Contribute to reports, benchmarks, evaluation documentation, and technical analyses.
• Communicate complex findings to both technical and non-technical stakeholders.
• Master's degree in Computer Science, Machine Learning, Artificial Intelligence, or a closely related technical field.
• Strong judgment concerning research signal quality, data selection, and evaluation design.
• Experience in designing datasets, evaluation frameworks, or QA processes for machine-learning systems.
• Capability to translate ambiguous operational challenges into structured research and evaluation problems.
• Familiarity with reinforcement-learning environments, agentic systems, or AI system evaluation.
• Strong analytical abilities and aptitude for producing concise, actionable technical insights.
• Proven competence in executing effectively within rapid iteration cycles and high-ambiguity environments.
• Excellent written and verbal communication skills.
• Collaborative experience with research, product, engineering, and domain teams.
• Client-facing experience in technical or research-oriented settings is a plus.
• Experience in building internal research or evaluation tools is advantageous.
• Contributions to benchmarks, research publications, or open research initiatives are favorable.
• Exposure to enterprise AI deployments or forward-deployed research settings is highly valued.
• Fully remote work arrangement.
• Full-time engagement.
• Compensation ranging from $300,000 to $700,000 per year.
• Remote consulting opportunities through 24-MAG LLC.
Rich Products Australia
Coursedog
DaCodes.
DaCodes.
Get handpicked remote jobs straight to your inbox weekly.