- Asapp 2
- Mountain View, CA
- Full-Time
- 17 days ago
Lead Machine Learning Engineer, Evaluations.
Before you go
Before you leave us, sign up for our email alerts
We don't do job spam, just the best digital jobs delivered straight to your inbox.
Lead Machine Learning Engineer, Evaluations: our view in 3 lines...
- The Role:Lead the evaluation platform for agentic AI systems, focusing on quality, safety and performance for production machine learning models.
- The Person:The person will design and own evaluation frameworks, build the data infrastructure behind them, partner with Research, Product and Platform teams, and mentor other engineers.
- Requirements:The role calls for deep experience building evaluation systems for modern ML, LLM and agentic systems, with Python, AWS, Kubernetes and/or Docker, plus a Bachelor’s Degree in CS or a related field.
About the role
The AI Engineering team is responsible for working closely with the research and modeling teams to create state-of-the-art NLP models for specific tasks, and deploy them in a production setting designed to serve our customers at scale. We are looking for a Machine Learning Engineer to help build and evaluate the core intelligence behind our agentic AI systems. This role will play a key part in designing and owning evaluation frameworks that ensure quality, safety, and performance across complex agentic systems.
We're looking for a Lead Machine Learning Engineer to own and grow the evaluation platform that measures quality, safety, and performance across ASAPP's agentic AI systems- the infrastructure that tells us, with confidence, whether a model or agent change is actually an improvement before it reaches customers.
This a hybrid role with 10-12 days of in-office presence per month to balance flexibility with collaboration.

