仕事内容
<div class="content-intro"><h2><strong>About Anthropic</strong></h2>
<p>Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.</p></div><h2><strong>About the role</strong></h2>
<p>Anthropic's Safeguards organization builds the policies, evaluations, and enforcement systems that keep our models from contributing to catastrophic harm. We are hiring a manager to lead the research engineering team responsible for biological safety: the evaluations, datasets, and classifiers that govern how our models handle biological knowledge.</p>
<p>You will lead a team of research scientists and engineers who design and run capability evaluations against frontier models, curate training data for our safety classifiers, train and iterate on those classifiers alongside our ML engineers, and measure how they hold up against adversarial pressure in production traffic. You will set the technical direction for that work, decide where the team invests, and own the results.</p>
<p>This is a hands-on management role. Most of your time goes to growing and directing the team, but you will keep enough technical depth to review an eval design, interrogate a classifier's failure modes, and represent the work credibly to Research, Product, and Policy partners.</p>
<p>The core tension your team owns is precision: safeguards need to be robust against sophisticated actors while staying out of the way of the far larger population of legitimate researchers using Claude to accelerate life sciences work. Getting that tradeoff right is an empirical problem, and your team is the one measuring it.</p>
<h2><strong>Key responsibilities</strong></h2>
<ul>
<li>
<p>Manage, coach, and grow a team of research scientists and engineers working on biologi