💼Jobs 📝Blog 🧮Salary Calc 🌍Cost of Living 📋Tax Guide
👤Sign In Post a Job — from $99
Anthropic

Research Lead, Training Insights

OtherFull-TimeLead
Location
Worldwide
Job Type
Full-Time
Experience
Lead
Apply Now

Job Description

ABOUT THE ROLE Anthropic is a rapidly growing organization dedicated to creating reliable, interpretable, and steerable AI systems. Our mission is to ensure that AI is safe and beneficial for our users and for society as a whole. As a Research Lead on the Training Insights team, you will play a critical role in shaping the evaluation narrative for model releases and contributing to how Anthropic communicates about its models to both internal and external audiences. WHAT YOU'LL DO As a Research Lead, you will be responsible for developing the strategy for and leading execution on how we measure and characterize model capabilities across training and deployment. This will involve driving original research into new evaluation methodologies, leading a small team of researchers and research engineers, and taking a cross-organizational view to map the landscape of model evaluations at Anthropic. Your work will span the full lifecycle of model development, from researching and building new long-horizon evaluations to developing novel approaches to measuring emerging capabilities. Your responsibilities will include: - Building new novel and long-horizon evaluations - Developing novel measurement approaches for understanding how model capabilities emerge and evolve during RL training - Leading strategic evaluation coverage across the company - Shaping the evaluation narrative for model releases - Leading and mentoring a small team of researchers and research engineers - Designing evaluation frameworks that balance scientific rigor with the practical demands of production training schedules - Building and maintaining relationships across Anthropic's research organization to ensure evaluation insights inform training and deployment decisions - Contributing to the broader research community through publications, open-source contributions, or external engagement on evaluation best practices WHAT YOU'LL NEED To be successful in this role, you will need significant experience designing and running evaluations for large language models or similar complex ML systems. You should have led technical projects or teams, either formally or through sustained ownership of critical research directions. You should be equally comfortable designing experiments and writing code, and be able to synthesize information across multiple teams and workstreams to form a coherent picture of model capabilities. You should also be able to communicate complex technical findings clearly to both technical and non-technical audiences, and be results-oriented and thrive in fast-paced environments where priorities shift based on research findings. A deep understanding of AI safety and a desire to directly influence how capable AI systems are developed and deployed is also essential. WHY REMOTE As a remote employee, you will have the flexibility to work from anywhere and be part of a dynamic and growing organization that is committed to creating beneficial AI systems. BENEFITS The annual compensation range for this role is $180,000 - $250,000, depending on experience. In addition to a competitive salary, we offer a comprehensive benefits package, including health insurance, retirement savings, and paid time off.