About the role
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
Responsibilities, requirements, fit, evidence and preparation are organized here. The original posting stays available for final verification.
What you’ll do
- Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
- You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models.
- Note: For this role, we conduct all interviews in Python. This role may require responding to incidents on short-notice, including on weekends.
- Implement and optimize post-training techniques at scale on frontier models
- Conduct research to develop and optimize post-training recipes that directly improve production model quality
- Design, build, and run robust, efficient pipelines for model fine-tuning and evaluation
- Develop tools to measure and improve model performance across various dimensions
- Collaborate with research teams to translate emerging techniques into production-ready implementations
- Debug complex issues in training pipelines and model behavior
- Help establish best practices for reliable, reproducible model post-training
What they’re looking for
Select a requirement to inspect fit, evidence or application context.
- A field relevant to the role as demonstrated through coursework, training, or professional experience
How they work
- Annual Salary: $350,000—$500,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time.
- However, some roles may require more time in our offices.
- We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Eligibility
Visa sponsorship: We do sponsor visas!
How to apply
- We encourage you to apply even if you do not believe you meet every single qualification.
About Anthropic
✓ Verified
We have limited verified information about this company. You can still explore its active opportunities and check the official source.
If this one isn’t right.
Related live opportunities you can compare without restarting your search.
Direct employer source
Aptiora keeps the source available for trust and final verification while the working experience stays here.