Member of Technical Staff, Post-Training, RL Infra
MirendilAI Research company
San Francisco, United States$300K - $500KLead
Andreessen Horowitz
Kleiner Perkins
Nvidia
Data & AI
About the role
TL;DR
Build and optimize infrastructure for large-scale reinforcement learning training.
- •We are looking for engineers to help build the post-training stack for frontier reasoning models.
- •This role sits at the intersection of research and infrastructure.
- •Key Responsibilities Design and build reliable infrastructure for large-scale RL training Implement novel performance optimizations across the training stack Develop evaluation and benchmarking infrastructure to measure model progress, throughput, and uptime Build data collection and feedback pipelines that close the loop between human signal, reward modeling, and training Collaborate with multiple teams to rapidly iterate on RL algorithms and get experiments into production training runs Requirements Extensive experience in reinforcement learning and large-scale model training Strong background in performance optimization and infrastructure development Ability to collaborate effectively with cross-functional teams Passion for building the infrastructure that makes frontier RL research possible at scale
Domain expertise
ai
Benefits & perks
Equity Grant