Design, train, and deploy AI agents for core products using end-to-end ML tasks.
•Toloka AI is seeking a Senior Machine Learning Engineer to design, train, and deploy AI agents that power their core products.
•This role focuses on end-to-end ML tasks, including fine-tuning and Reinforcement Learning (RL), to build resilient agentic workflows.
•Key Responsibilities Train, fine-tune, and distill ML models (including RL approaches) to power autonomous AI agents.
•Build and operate agentic workflows in Python, handling complex reasoning and hybrid human-expert interactions.
•Own evaluation and benchmarking, selecting foundational models and establishing cost models.
•Manage the full ML lifecycle: design solutions, ship to production, and monitor real-time signals.
•Implement observability metrics tailored for agent logic, model performance, and system reliability.
•Requirements 3+ years of experience in ML, with a strong background in model training, fine-tuning, and Reinforcement Learning (RL). 1+ year in Agent Development, with practical experience building, evaluating, and launching autonomous AI agents.
•Proven experience working with agentic frameworks such as LangChain, LlamaIndex, or AutoGen.
•Practical knowledge of model distillation and adapting open-source models (e.g., Llama, Mistral).
•Advanced proficiency in Python with a drive to apply disciplined software engineering standards to ML.
•Ability to work across the entire chain, from research to production operations.