NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Monitoring & Evaluation Specialist in United States of America
9 days ago
Apply with autofill
Apply with autofill
Rhoda-ai·9 days ago
9 days ago

Research Evaluation Lead

Mountain View, United States of AmericaFull-timeSenior · 6-10 yearsMonitoring & Evaluation Specialist

Sign up free to see how well your resume matches this role.

Boost your chances at rhoda-ai

How you compare FREE

?
Your scoreYour score: not yet known
→
17
Top 10%Top 10%: 17 out of 100

Top 10% of NextRaise users matched against Monitoring & Evaluation Specialist roles in United States.

Must-have skills for this role

  • robotics
  • ml experimentation
  • experimentation
  • execution

PDF or DOCX · no account needed

Apply faster with autofill FREErhoda-ai uses Ashby - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Translate research intent into clear eval protocols, trial plans, and success criteria.
  • Own execution end-to-end: model handoff → station readiness → pilot execution → QA → results.
  • Train and manage eval pilots; ensure consistency across people, shifts, and stations.
  • Distinguish model failures from hardware, setup, operator, or data-quality issues.
  • Maintain eval setups, resets, randomization, metadata, and experiment traceability.
  • Track quality, throughput, and bottlenecks; continuously improve the eval process.
  • Partner closely with Research, Robot Data, and Eval Platform teams.

What they're looking for

  • Some understanding of robotics / ML experimentation.
  • Computer science background or hands-on experience with coding
  • Rigorous, detail-oriented, and able to understand the intent behind an experiment, not just execute instructions.
  • Strong hands-on execution and ownership.

Nice to have

  • Experience in robotics testing, data collection, lab operations, or QA preferred.

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.

Mission

Own end-to-end robot model evaluation for Research. Turn research questions into consistent, high-quality, repeatable evals that enable fast iteration and trusted results.

Responsibilities

  • Translate research intent into clear eval protocols, trial plans, and success criteria.

  • Own execution end-to-end: model handoff → station readiness → pilot execution → QA → results.

  • Train and manage eval pilots; ensure consistency across people, shifts, and stations.

  • Distinguish model failures from hardware, setup, operator, or data-quality issues.

  • Maintain eval setups, resets, randomization, metadata, and experiment traceability.

  • Track quality, throughput, and bottlenecks; continuously improve the eval process.

  • Partner closely with Research, Robot Data, and Eval Platform teams.

What we’re looking for

  • Some understanding of robotics / ML experimentation.

  • Computer science background or hands-on experience with coding

  • Rigorous, detail-oriented, and able to understand the intent behind an experiment, not just execute instructions.

  • Strong hands-on execution and ownership.

  • Experience in robotics testing, data collection, lab operations, or QA preferred.

Success looks like

A researcher can hand off a model and research question and receive a trusted, standardized eval result with sufficient trials and QA.

Company

Rhoda-ai
Mountain View, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Rhoda Ai's careers site·first seen 11 Sept 2026·last verified 11 Sept 2026·How we source jobs

Similar jobs

  • Software Engineers: Paid Interview on AI Evaluation Tasks at teracUnited States–match not yet calculated
  • Quality Lead, Agentic AI Workflow Evaluation at innodataincIn Office, United States of America–match not yet calculated
  • Evaluation Associate at cityofphiladelphiaPhiladelphia, United States of America–match not yet calculated
  • Manager II, Engineering - Database Monitoring (AI) at DatadogNew York, United States of America–match not yet calculated
  • Management Consultants: Paid AI Output Evaluation at teracUnited States–match not yet calculated

Browse more jobs

  • Monitoring & Evaluation Specialist jobs in United States
  • Community Health Worker jobs in United States
  • Public Health Officer jobs in United States
  • Field Officer jobs in United States
  • Monitoring & Evaluation Specialist jobs in India
  • Monitoring & Evaluation Specialist jobs in France