Currently tracking 127 active AI roles, down 32% versus the prior 4 weeks. Primary focus: Agent · Engineering. Salary range $46k–$850k (avg $406k).
Anthropic currently has 151 active AI-related job listings, with a significant focus on roles related to agents, which constitute 32% of their openings. Engineering is the most frequent function, followed by Research. The majority of their hiring is concentrated in the United States. Frequent technical tags include evals, model_serving, and agent_orchestration, suggesting a focus on the practical deployment and management of AI systems.
Anthropic currently has 149 active AI-related roles in our index. The most common open titles are: Regional Research Economist, Economic Research (2), Research Engineer, Machine Learning (RL Velocity) (2), Research Engineer, Production Model Post-Training (2), Staff Software Engineer, AI Reliability Engineering (2), Product Manager, Safeguards Rare Harms. Most positions are in Engineering and Research.
Anthropic's active AI hiring is concentrated in: agents (31%), serving infrastructure (15%), post-training (15%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
Anthropic is hiring AI talent in: United States (130 roles), United Kingdom (18 roles), Canada (6 roles), Ireland (3 roles).
Job postings at Anthropic most frequently mention: Machine Learning, AI Safety, Production ML Systems, System Design, Large Language Models (LLMs).
In the past 30 days, Anthropic has posted 32 new AI-related roles.
| Title | Stage | AI score |
|---|---|---|
| Researcher, Cybersecurity Products Researcher on a product team focused on AI for cybersecurity, identifying, measuring, and operationalizing security capabilities in frontier models for non-expert customers. Involves rapid prototyping, rigorous evaluation, and collaboration with engineers to build usable tools. | Eval GateAgent | 9 |
| Research Scientist, Takeoff Intel Research Scientist focused on measuring and understanding recursive-self-improvement in AI systems. This role involves designing evaluations, building quantitative models of capability growth, running experiments, and assessing AI R&D acceleration. The output is focused on graded assessments and system-card sections rather than traditional publications. | Eval Gate |
| 9 |
| Evals Infrastructure Tech Lead / Manager Lead the team building and scaling the distributed systems that orchestrate, schedule, and execute evals for frontier models, ensuring measurement quality, reproducibility, and that eval signal reaches decision-makers. This role involves managing engineers and contributing directly as an engineer, focusing on inference, research, and infrastructure engineering. | Eval GateServe | 9 |
| Research Engineer, Safeguards Labs Research Engineer focused on AI safety, investigating novel methods for detecting misuse, strengthening model safeguards, and building evaluation methodologies for AI systems, particularly in agentic workflows. The role involves leading research projects, designing offline analyses, developing prototypes, and collaborating with production teams. | Eval GatePost-train | 9 |
| Model Quality Software Engineer, Claude Code Staff Software Engineer to set technical direction at the intersection of engineering and research on the Claude Code team. Architect systems, tooling, and evaluation infrastructure to measure, understand, and improve Claude's coding capabilities. Drive architecture, mentor engineers, and influence the direction of Claude Code. | Eval GateAgent | 9 |
| Product Manager, Safeguards (Child Safety) Product Manager for Safeguards (Child Safety) at Anthropic, focusing on building and deploying safety systems for AI models. This role involves ideation, design, development, and deployment of Safeguards systems and product UX to ensure safe advancement of frontier models. Responsibilities include defining safety by design, writing safety evals, prioritizing problems, collaborating with cross-functional teams, understanding AI landscape risks, and developing metrics for risk assessment. Requires 5+ years of product management experience with a focus on technical tradeoffs, user understanding, strategy development, and metric design in a rapidly changing environment. | Eval GatePost-train | 8 |
| Product Manager, Safeguards Rare Harms Product Manager for Anthropic's Safeguards team, focusing on building and deploying systems to ensure AI safety and prevent misuse. This role involves ideation, design, development, and UX for safeguards, working closely with research and product teams to mitigate risks associated with frontier models across various platforms. | Eval GateAgent | 8 |
| Product Manager, Safeguards (Verticals) Product Manager for Anthropic's Safeguards team, focusing on building and deploying systems to ensure AI safety and prevent misuse. This role involves ideation, design, development, and deployment of safeguards, working closely with research and product teams to create detections, evals, interventions, and tools for risk mitigation. The PM will drive strategy, prioritize features, collaborate with cross-functional teams, and develop metrics to assess performance and inform future planning. | Eval GateAgent | 8 |
| Biological Safety Research Scientist Research Scientist focused on biological safety for AI systems, applying technical skills to design and develop safety systems that detect harmful behaviors and prevent misuse. This role involves designing and executing capability evaluations, collaborating on training data and safety system training, analyzing performance, and stress-testing safeguards. The goal is to ensure biological safety is embedded throughout the model development lifecycle, balancing AI's potential in life sciences with preventing misuse. | Eval GatePost-train | 8 |
| Safeguards Enforcement Analyst, Violence & Extremism This role focuses on building and executing operational workflows to assess AI model behavior, drive enforcement decisions, and develop evaluations for violence and extremism policy areas. It involves designing automated enforcement systems, creating evals to measure model performance, and partnering with engineering and data science teams to optimize detection and enforcement. The role also requires reviewing flagged content, identifying misuse patterns, and staying updated on emerging threats and AI policy enforcement best practices. | Eval GateAgent | 7 |
| Data Scientist, Safeguards This role focuses on building and scaling a data-driven culture within an AI company, specifically for safeguards. The Data Scientist will analyze user behavior, define key metrics, identify opportunities for product improvement, design and analyze experiments, and establish data best practices to inform product and commercial strategy for safe, frontier AI deployment. | Eval Gate | 7 |
| Safeguards Enforcement Analyst, User Well-being This role supports the design and deployment of mental health guardrails for AI systems, focusing on detection systems, review queues, and intervention evaluation. It involves translating clinical guidance and data analysis into actionable changes for AI response and monitoring. The role partners with Engineering and Data Science to build and tune detection models, monitors system performance, reviews flagged content, and supports the development of in-product features. It requires experience in trust & safety or policy, with a focus on mental health harms, and proficiency in data analysis tools. Experience with generative AI products and LLM-based classification systems is preferred. | Eval GateAgent | 5 |
| Safeguards Enforcement Analyst, Bio Harms This role focuses on enforcing AI usage policies, specifically for biological harms, by analyzing model interactions, investigating violations, and improving automated enforcement systems. It involves working with AI tools for content review and data analysis, and partnering with Engineering and Data Science teams to optimize detection models. | Eval Gate | 5 |
| Safeguards Enforcement Analyst, Chem & Explosives Harms This role focuses on enforcing AI usage policies related to chemical and explosives harms. The analyst will monitor platform activity, investigate violations, and work with data science and engineering teams to improve detection models and automated enforcement systems. The role requires expertise in chemistry and experience in trust & safety or policy enforcement, with a focus on using AI tools to enhance review workflows. | Eval Gate | 5 |
| Safeguards Enforcement Analyst, Age-Appropriate Design This role focuses on designing and executing enforcement workflows for AI products, specifically concerning age-appropriateness and detecting/mitigating potential harm. It involves partnering with engineering and data science teams to optimize detection models and enforcement systems, reviewing flagged content, and working with legal and policy stakeholders to ensure compliance with evolving regulations. The role is responsible for Anthropic's layered age assurance approach and may expand to broader user well-being enforcement. | Eval Gate | 5 |
| Safeguards Enforcement Analyst, Safety Evaluations This role focuses on evaluating AI models against safety and policy standards, running and monitoring evaluations, driving mitigations, and coordinating the creation of new evaluation frameworks. It involves cross-functional collaboration with policy experts and engineering teams to ensure model behavior meets required standards and to build scalable processes for evaluation. | Eval Gate | 5 |
| Technical Program Manager, Safeguards (Infrastructure & Evals) Technical Program Manager for Safeguards Infrastructure and Evals at Anthropic. This role focuses on owning the operational health, reliability, and forward momentum of AI safety infrastructure, including classifiers, detection pipelines, evaluation platforms, and monitoring systems. Responsibilities include driving incident response, post-mortem execution, establishing and maintaining SLOs with partner teams, maintaining runbook quality, managing platform migrations, and coordinating improvements to the evals platform. Requires technical depth in production ML systems and strong program management skills in operational and infrastructure-heavy environments. | Eval GateServe | 5 |