NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Monitoring & Evaluation Specialist in United States of America
3 months ago
Apply with autofill
Apply with autofill
Mirendil·3 months ago
3 months ago

Member of Technical Staff, Model Evaluation

San Francisco, United States of AmericaFull-timeOn-siteSenior · 8-12 yearsMonitoring & Evaluation Specialist

Sign up free to see how well your resume matches this role.

Boost your chances at mirendil

How you compare FREE

?
Your scoreYour score: not yet known
→
16
Top 10%Top 10%: 16 out of 100

Top 10% of NextRaise users matched against Monitoring & Evaluation Specialist roles in United States.

PDF or DOCX · no account needed

Apply faster with autofill FREEmirendil uses Ashby - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

About this role

Mirendil

Mirendil is a tech-first company focused on solving core bottlenecks that unlock step-change acceleration across science and technology. Our first goal is to democratize frontier AI R&D across scientific disciplines. We are building a frontier AI research company and training our own models end-to-end.

The Role

We are looking for a research engineer to build the evaluation infrastructure that tells us whether our models are getting better in ways we care about. You'll own the frameworks, pipelines, and tooling that measure model behavior across capabilities. Some example areas you might work on (not limited to):

  • Design and build evaluation frameworks that measure model capabilities along realistic axes, beyond standard benchmarks.

  • Build automated eval pipelines and regression-detection systems that run continuously and surface signal quickly.

  • Develop agent-assisted workflows for humans to efficiently inspect model behavior.

  • Instrument training runs with observability tooling so researchers can understand what's changing in model behavior, and why.

  • Partner with post-training and RL teams to close the loop between eval signal and training decisions.

If you're excited about the hard problem of knowing whether a frontier AI system is actually improving, we'd love to hear from you.

We offer a base salary of $300,000–$400,000 USD and a meaningful equity grant, depending on experience and background, along with competitive benefits.

Company

Mirendil
San Francisco, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Mirendil's careers site·first seen 7 Aug 2026·last verified 8 Sept 2026·How we source jobs

Similar jobs

  • Remote Patient Monitoring Specialist at vironix-aiUnited States–match not yet calculated
  • Sr. Manager, Security — Continuous Monitoring v 2.0 at DatabricksRemote - California–match not yet calculated
  • NOSC Monitoring Analyst at gditUSA FL MacDill AFB–match not yet calculated
  • M&E Manager at MTR - Construction RecruitmentHandforth, United States of America–match not yet calculated
  • Cyber Monitoring Analyst at CACI InternationalHampton, United States of America–match not yet calculated

Browse more jobs

  • Monitoring & Evaluation Specialist jobs in United States
  • Community Health Worker jobs in United States
  • Public Health Officer jobs in United States
  • Field Officer jobs in United States
  • Monitoring & Evaluation Specialist jobs in India
  • Monitoring & Evaluation Specialist jobs in United Kingdom