Machine Learning Engineer — Model Evaluation & Experimentation
Mercor
$60 – $90/hour · as listed
Machine Learning Engineer — Model Evaluation & Experimentation
As a Machine Learning Engineer at Mercor, you'll work with a leading AI lab's team to evaluate and benchmark frontier AI models. Your core responsibility involves serving as a ground-truth expert, assessing model outputs and designing experiments that measure performance across complex tasks. This role combines deep technical judgment with hands-on AI evaluation work, requiring you to understand model behavior at an advanced level and contribute to the development of evaluation frameworks for next-generation AI systems.
This position suits experienced ML practitioners and software engineers with strong coding fundamentals and a solid grasp of machine learning concepts. You'll need the ability to think critically about model outputs, design rigorous tests, and work independently on technical problems. The role is based in the US and offers full-time, remote work at 35 hours per week, with compensation of $60–$90 per hour depending on experience.
From the listing: Mercor marketplace · full-time · 35h/week. Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models. ## 1\. Overview A leading AI lab is building the next generation of agentic evaluation benchmarks for frontier models and needs experienced machine learning practitioners to act as ground-truth experts for model evaluation and expe
Applications are completed on the listing’s own site. Pay is shown as posted by the source (“as listed”) or the aggregator’s estimate where marked — offers and availability change, and individual results vary.
Want help getting selected — and finding the best-paid work for your country? See how membership works →