Member of Technical Staff (Models)
2 нед. назад
130k–180k USD / yearUSALeadOnsite
machine learningreinforcement learningsoftware engineeringmodel fine-tuning
Responsible for improving AI model performance through post-training, fine-tuning, and experimentation.
Другое
- As a Member of Technical Staff on Models, you'll own the post-training loop that improves the performance of our AI Employees. You will be defining how we train models to do mission critical work in the real world.
- Post-training and fine-tuning
- Turning agent traces and expert feedback into training data
- Reward modeling and graders for non-verifiable outcomes
- Distillation into smaller, faster, cheaper models
- Designing and running training experiments
- Have trained or fine-tuned models and shipped the result into a product
- Have hands-on experience with SFT, preference optimization, or RLHF
- Can read traces and tell whether a metric measures the thing that matters
- Have strong software engineering fundamentals alongside ML depth
- Want to work in person in San Francisco
- You've adapted open-weight models to a specialized domain
- You've built training data or evaluation infrastructure
- You contribute to open source projects