
Job Description
As a Member of Technical Staff on Models, you'll own the post-training loop that improves the performance of our AI Employees. You will be defining how we train models to do mission critical work in the real world.
Areas you may work in:
- Post-training and fine-tuning
- Turning agent traces and expert feedback into training data
- Reward modeling and graders for non-verifiable outcomes
- Distillation into smaller, faster, cheaper models
- Designing and running training experiments
You may be a good fit if you:
- Have trained or fine-tuned models and shipped the result into a product
- Have hands-on experience with SFT, preference optimization, or RLHF
- Can read traces and tell whether a metric measures the thing that matters
- Have strong software engineering fundamentals alongside ML depth
- Want to work in person in San Francisco
Even better:
- You've adapted open-weight models to a specialized domain
- You've built training data or evaluation infrastructure
- You contribute to open source projects
Optimize Your Resume for This Job
Get a match score and see exactly which keywords you're missing
Job Details
- Category
- Software
- Employment Type
- Full Time
- Location
- San Francisco, CA
- Posted
- Compensation
- $130,000 - $180,000 per year
About Jarmin
Jarmin.ai is your 24/7 ML eng employee that you hire. Just talk to Jarmin like any other employee and Jarmin handles your AI/ML work from there. Completely hand over full initiatives or individual tasks. Jarmin’s founding team is from Meta SuperIntelligence Labs (Staff Engineers), Apple (Machine Learning Engineer), AWS (scaling infra), Lockheed-Martin (Research), and JPMorganChase (SWE) with experience building AI/ML and automating complex work of Researchers, Data Scientists, ML Engineers with AI.
More Roles at Jarmin





Similar Software Roles



Found this role interesting?