About this role
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About the Role
We're looking for a Capabilities Researcher to join the team building Claude Security. In this role, you'll identify which security capabilities in frontier models are ready to build on, measure how well they perform, and work out how to make them useful to customers who are not security experts.
Frontier models have become substantially more capable at security work over recent generations, and they continue to improve with each release. That progress creates a set of practical questions: which capabilities are reliable enough to depend on, how they perform in realistic conditions, how they behave when an adversary is involved, and where their limits are. You'll be responsible for answering those questions through fast prototyping and rigorous evaluation, and your findings will shape what the team decides to build.
You'll also work on making those capabilities usable. Together with engineers on the team, you'll design the scaffolding, tooling, and defaults that let a strong model capability do useful work for a non-expert, and you'll stay involved as it becomes a product.
This is a research role on a product team, with a broad charge and real latitude in what you investigate. It suits someone who already has ideas about what AI should be able to do for security teams and wants the models, the time, and the engineering support to pursue them.
Responsibilities
• Prototype rapidly to find define the AI frontier for cybersecurity work
• Design evaluations that measure model performance on the work security teams actually do
• Build the datasets, harnesses, and scoring those evaluations depend on
• Engage with the cybersecurity community to help define where AI can make the most impact
• Work with engineers and researchers to operationalize promising capabilities into something customers can rely on
• Track how model capabilities for security are changing, and what that means for what we build next
• Share findings that inform product direction, and partner with product leadership on priorities
You may be a good fit if you:
• Have deep expertise in one or more security domains, such as vulnerability research, exploit development, reverse engineering, malware analysis, incident response, or offensive security
• Have built AI-powered tools or capabilities for security work
• Can get from an idea to a working prototype quickly, and abandon the ones that don't hold up
• Are comfortable designing rigorous evaluations and interpreting the results honestly
• Can write and communicate clearly about technical findings
• Have 7+ years of experience in security research, security engineerin