- Salary
- $130k – $180k/yr
- Location
- San Francisco, California, US
- Type
- Full-time
- Department
- IT
- Seniority
- Senior
- Source
- Y Combinator
Description
As a Member of Technical Staff on Models, you'll own the post-training loop that improves the performance of our AI Employees. You will be defining how we train models to do mission critical work in the real world.
Areas you may work in:
- Post-training and fine-tuning
- Turning agent traces and expert feedback into training data
- Reward modeling and graders for non-verifiable outcomes
- Distillation into smaller, faster, cheaper models
- Designing and running training experiments
You may be a good fit if you:
- Have trained or fine-tuned models and shipped the result into a product
- Have hands-on experience with SFT, preference optimization, or RLHF
- Can read traces and tell whether a metric measures the thing that matters
- Have strong software engineering fundamentals alongside ML depth
- Want to work in person in San Francisco
Even better:
- You've adapted open-weight models to a specialized domain
- You've built training data or evaluation infrastructure
- You contribute to open source projects