Machine Learning Eval Engineer
ReductoAI Document company
San Francisco, United StatesMid
Andreessen Horowitz
Benchmark
First Round Capital
BoxGroup
Y Combinator
Data & AI
About the role
TL;DR
Play a key role in building evaluation systems for machine learning models.
- •As an ML Eval Engineer, you’ll play a key role in building the evaluation systems and benchmarks that make Reducto’s models better over time.
- •Key Responsibilities Design, build, and maintain evaluation benchmarks that reveal where our models perform well and where they fail.
- •Develop metrics, heuristics, and workflows to automatically identify new failure modes across large and messy real-world datasets.
- •Partner closely with other ML engineers to turn evaluation insights into model improvements and better training priorities.
- •Requirements Hold yourself to a high bar for quality and precision.
- •Enjoy solving complex problems and building from first principles.
- •Have strong Python skills and can independently build clean, reliable technical solutions.
- •Bonus points for product and frontend experience!
Required skills
Python
Tech stack
Python