Skip to content
Reducto logo

Machine Learning Eval Engineer

ReductoAI Document company
San Francisco, United StatesMid
Andreessen Horowitz logo
Andreessen Horowitz
Benchmark logo
Benchmark
First Round Capital logo
First Round Capital
BoxGroup logo
BoxGroup
Y Combinator logo
Y Combinator
Data & AI

About the role

TL;DR

Play a key role in building evaluation systems for machine learning models.

  • As an ML Eval Engineer, you’ll play a key role in building the evaluation systems and benchmarks that make Reducto’s models better over time.
  • Key Responsibilities Design, build, and maintain evaluation benchmarks that reveal where our models perform well and where they fail.
  • Develop metrics, heuristics, and workflows to automatically identify new failure modes across large and messy real-world datasets.
  • Partner closely with other ML engineers to turn evaluation insights into model improvements and better training priorities.
  • Requirements Hold yourself to a high bar for quality and precision.
  • Enjoy solving complex problems and building from first principles.
  • Have strong Python skills and can independently build clean, reliable technical solutions.
  • Bonus points for product and frontend experience!
View original posting →

Required skills

Python

Tech stack

Python

Similar jobs

D

Senior Data Scientist

D

Analytics Engineer