Description
Anthropic is hiring a Research Engineer for its Post-Training team to implement, optimize, and scale post-training techniques such as Constitutional AI, RLHF, and other alignment methods on frontier models. The role involves developing fine-tuning and evaluation pipelines, measuring model performance, debugging training issues, and translating research into production-ready systems. Candidates should have strong software engineering and machine learning experience, familiarity with large language models, and proficiency in Python, deep learning frameworks, and distributed computing. The role is hybrid, requires at least a bachelor’s degree in a relevant field, and may involve short-notice incident response.

