Member of Technical Staff, Post-Training
Job Overview
Post-training is where a raw model becomes something people can actually rely on. You will design and run the fine-tuning and reinforcement learning pipelines that shape our flagship models, build the evaluations that tell us whether they got better, and decide what ships.
Responsibilities
- Design and run supervised fine-tuning and reinforcement learning pipelines at scale
- Build evaluations that measurably distinguish model quality, not just vibes
- Investigate regressions in model behaviour and drive them to root cause
- Set the technical direction for how flagship models are shaped before release
Qualifications
- Deep experience training or fine-tuning large language models
- Strong grounding in reinforcement learning from human or AI feedback
- Track record of building evaluations that survive contact with reality
- Comfortable owning ambiguous, high-stakes technical decisions
Compensation & Benefits
This position has an estimated salary range of $400,000.00 - $575,000.00 per year, plus potential equity and bonus.
Key skills
Compliance Information
Equal Opportunity: Northwind is an equal opportunity employer and complies with all applicable federal, state, and local fair employment practices laws.
Hiring Disclosure: We are committed to providing equal employment opportunities to all employees and applicants for employment.
Labor Law Reference: Fair Labor Standards Act (FLSA), Title VII of the Civil Rights Act of 1964