Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Researcher, Evaluations

epoch-ai
Companyepoch-ai
CategoryUncategorised
LocationRemote
RemoteRemote
EmploymentFull-time
LevelNot stated
SalaryNot stated by the employer
Posted30 Jun 2026
Last verified2 Aug 2026
SourceEmployer career page (lever)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
Epoch AI is looking for a researcher to evaluate frontier AI models on hard-to-grade tasks drawn from real-world scenarios. About the role We’re seeking a Researcher to lead a new effort evaluating how well frontier models perform on the kinds of open-ended tasks that make up real office work. You will curate a suite of realistic tasks to serve as a benchmark, design the grading rubrics for AI performance, and run newly-released models through the suite, assessing their performance both quantitatively and qualitatively. The focus is on how models handle messy, real-world work rather than on scientific knowledge or programming ability. The role makes heavy use of AI tools, but strong software engineering experience is not required. Comfort setting up AI-assisted automated workflows is a plus. If this role sounds interesting, we are also looking for researchers on multiple other teams.  Applications are rolling.
HOUSE AD1,962,144 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →