← Back to jobs

Remote job

Model Evaluation Engineer — Benchmarks & Evals

AI Engineer Full-time Remote

Job details

Not specified Salary
Remote Eligibility
Not specified Experience
Full-time Employment

About this role

Role overview

Remote, full-time Model Evaluation Engineer role on an AI Research team. The position centers on designing, running, and analyzing benchmarks and evaluations that measure machine learning model behavior and progress.

Responsibilities

- Design and maintain benchmark suites and evaluation methodologies - Run experiments, analyze results, and communicate findings to the AI Research team - Iterate on evaluation infrastructure to support ongoing model development

Requirements

- Experience with ML model evaluation, benchmarking, or research engineering - Strong analytical skills and rigor in experimental design - Comfort operating as part of a fully remote, distributed team

Detected Oct 9, 2026
Last verified Oct 9, 2026

Hidden Jobs Access

Unlock application links

Read the full job details for free. An active Hidden Jobs Access subscription is required to open the original application link.

Weekly

FREE $6.99/week after trial
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
  • Cancel anytime before day 7

Monthly

$35.99 $17.99 /month
  • 35% cheaper than weekly
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching

Lifetime

$99.99 $49.99 /forever
  • One-time payment
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
Hidden Jobs gives subscribers direct access to original application links
Offer ends in 00:00:00 Your profile-fit rate expires at midnight