As a Member of Technical Staff on Models, you'll own the post-training loop that improves the performance of our AI Employees. You will be defining how we train models to do mission critical work in the real world.
**Areas you may work in:**
* Post-training and fine-tuning
* Turning agent traces and expert feedback into training data
* Reward modeling and graders for non-verifiable outcomes
* Distillation into smaller, faster, cheaper models
* Designing and running training experiments
**You may be a good fit if you:**
* Have trained or fine-tuned models and shipped the result into a product
* Have hands-on experience with SFT, preference optimization, or RLHF
* Can read traces and tell whether a metric measures the thing that matters
* Have strong software engineering fundamentals alongside ML depth
* Want to work in person in San Francisco
**Even better:**
* You've adapted open-weight models to a specialized domain
* You've built training data or evaluation infrastructure
* You contribute to open source projects
Visa sponsorship is not available for this role.