Sep 2026–Present
Remote · Part-time
Mercor
Mathematics Expert — AI Training & Evaluation
- Translate research papers in mathematics and computational finance into self-contained LLM benchmarks with reference solutions, grading rubrics, and source-grounded reasoning requirements.
- Build Python oracle implementations, numerical edge-case tests, and multi-stage integration tests; use repeated calibration runs to refine benchmark difficulty and reliability.