Anthropic

Research Engineer, Production Model Post-Training

$350K–$500KFull-time · San Francisco, CA +2 more
✓ Verified live on the employer's own system · added 496 days ago
Save search

What this role involves

MultimodalModel TrainingDeep LearningFrontierTuningPhysicsLLM

Skills & tools

HiringPythonTroubleshootingProgrammingDistributed SystemsOperationsMachine Learning
Apply on company site ↗ See your fit → free

Full job description

Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.

You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models.

Note: For this role, we conduct all interviews in Python. This role may require responding to incidents on short-notice, including on weekends.

- Implement and optimize post-training techniques at scale on frontier models

- Conduct research to develop and optimize post-training recipes that directly improve production model quality

- Design, build, and run robust, efficient pipelines for model fine-tuning and evaluation

- Develop tools to measure and improve model performance across various dimensions

- Collaborate with research teams to translate emerging techniques into production-ready implementations

- Debug complex issues in training pipelines and model behavior

- Help establish best practices for reliable, reproducible model post-training

- Thrive in controlled chaos and are energised, rather than overwhelmed, when juggling multiple urgent priorities

- Maintain clarity when debugging complex, time-sensitive issues

- Have strong software engineering skills with experience building complex ML systems

- Are comfortable working with large-scale distributed systems and high-performance computing

- Have experience with training, fine-tuning, or evaluating large language models

- Can balance research exploration with engineering rigor and operational reliability

- Are adept at analyzing and debugging model training processes

- Enjoy collaborating across research and engineering disciplines

- Can navigate ambiguity and make progress in fast-moving research environments

- Have a keen interest in AI safety and responsible deployment

We welcome candidates at various experience levels, with a preference for senior engineers who have hands-on experience with frontier AI systems. However, proficiency in Python, deep learning frameworks, and distributed computing is required for this role.

More jobs at Anthropic

Similar jobs near San Francisco, CA +2 more

Tell me when more Production Engineer, Network jobs post near San Francisco, CA We re-check every listing against the employer’s own board — no résumé needed.

Search Research Engineer, Production Model Post-Training jobs near San Francisco, CA +2 more → Browse all live jobs

This posting was published by Anthropic on their own careers system and is shown here with a direct link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.