AI Frontier · AI lab
Currently tracking 127 active AI roles, down 32% versus the prior 4 weeks. Primary focus: Agent · Engineering. Salary range $46k–$850k (avg $406k).
Anthropic currently has 151 active AI-related job listings, with a significant focus on roles related to agents, which constitute 32% of their openings. Engineering is the most frequent function, followed by Research. The majority of their hiring is concentrated in the United States. Frequent technical tags include evals, model_serving, and agent_orchestration, suggesting a focus on the practical deployment and management of AI systems.
Anthropic currently has 149 active AI-related roles in our index. The most common open titles are: Regional Research Economist, Economic Research (2), Research Engineer, Machine Learning (RL Velocity) (2), Research Engineer, Production Model Post-Training (2), Staff Software Engineer, AI Reliability Engineering (2), Product Manager, Safeguards Rare Harms. Most positions are in Engineering and Research.
Anthropic's active AI hiring is concentrated in: agents (31%), serving infrastructure (15%), post-training (15%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
Anthropic is hiring AI talent in: United States (130 roles), United Kingdom (18 roles), Canada (6 roles), Ireland (3 roles).
Job postings at Anthropic most frequently mention: Machine Learning, AI Safety, Production ML Systems, System Design, Large Language Models (LLMs).
In the past 30 days, Anthropic has posted 32 new AI-related roles.
| Title | Stage | AI score |
|---|---|---|
| Staff Software Engineer, Code RL Staff Software Engineer focused on the engineering aspects of reinforcement learning for AI coding capabilities, specifically creating and scaling agentic coding environments. The role involves designing frameworks and APIs for researchers, managing production RL runs, and improving the reliability and structure of research codebases. It emphasizes Python expertise, API design, and anticipating system failures. | AgentData | 9 |
| Staff+ Software Engineer, Enterprise AI Products Staff+ Software Engineer for Anthropic's Enterprise AI Products team, focusing on building organizational context and workflows (plugins, skills, connectors, webhook-triggered agents) to make Claude a daily-use tool for enterprise customers. This role involves technical leadership, end-to-end product delivery, customer interaction, and close collaboration with research to integrate model capabilities into production. |
| AgentShip |
| 9 |
| Lead, Frontier Red Team (Cyber) Lead a new team focused on researching the impacts of advanced AI models on cybersecurity, developing defenses, and steering the world through the next generation of cybersecurity. This role involves leading a research team, publishing frontier research, and building/deploying defenses to give defenders a permanent advantage. | AgentPost-train | 9 |
| Research Engineer, Chip Design RL (Reinforcement Learning) Research Engineer role focused on applying Reinforcement Learning to chip design, specifically for agentic RTL generation, design verification, and physical design optimization. The role involves inventing RL environments, optimizing EDA tools, conducting experiments, and delivering work into research and production training runs. Requires expertise in ASIC/FPGA design and familiarity with EDA tools, with strong candidates having RL experience. | AgentData | 9 |
| [Pipeline] Staff+ Software Engineer, Developer Acceleration Staff+ Software Engineer to join Developer Acceleration team, responsible for the infrastructure enabling thousands of employees to be productive via agents. The role will define the future of agentic productivity at scale, own technical strategy, build and ship the runtime platform and tooling, write evals to benchmark agent behaviors, and ensure infrastructure scalability and reliability. | Agent | 9 |
| Research Engineer, Computer Use Research Engineer focused on teaching AI models (Claude) to perceive, use, and understand computer interfaces, enabling them to reliably and safely operate real software. This involves designing experiments, developing evaluation frameworks, building RL training environments, and collaborating with training and product teams to integrate research advances into production. | AgentPost-train | 9 |
| Manager, Applied AI Engineering, Life Sciences (Beneficial Deployments) Manager for Applied AI Engineering focused on Life Sciences, leading a team to build and deploy AI solutions (agents, integrations, tools) for scientific organizations. The role involves deep customer engagement, understanding scientific workflows, and ensuring reliable, reproducible access to biological data for AI agents, with a strong emphasis on safety and responsible deployment in a regulated-like environment. | Agent | 9 |
| Software Engineer, Safeguards Evals Software Engineer role focused on building and owning the evaluation infrastructure for an agentic investigation system. This involves designing experiments, constructing high-quality eval datasets, measuring agent performance, analyzing coverage gaps, and productionizing research into release pipelines. The role also involves building tooling for policy experts and constructing RL environments to improve safety investigation capabilities. | AgentEval Gate | 9 |
| Product Manager, Claude Code Model Performance Product Manager for Anthropic's Claude Code Model Performance team, responsible for driving model launches, building agentic evals, and translating research improvements into developer-facing outcomes. Requires experience building agentic evals, a systems thinking approach, and comfort with both research and engineering. | AgentEval Gate | 9 |
| Research Engineer / Scientist, Frontier Red Team (Cyber) Research Engineer/Scientist focused on AI-enabled cybersecurity, developing tools and frameworks for autonomous vulnerability discovery, remediation, malware detection, and pentesting. Designs and runs experiments to evaluate AI cyber capabilities and builds infrastructure for AI systems operating in security environments. Translates findings into demonstrations for policymakers and collaborates with external experts. Senior candidates will set research strategy and own the technical roadmap. | AgentEval Gate | 9 |
| Research Engineer, Universes Research Engineer role focused on building next-generation agentic environments for training AI models. This role involves implementing novel approaches, contributing to research direction, designing training environments and methodologies, and building evaluations for capable and safe agentic AI. It blends research and engineering, with a focus on reinforcement learning and complex, long-horizon agentic tasks. | AgentPost-train | 9 |
| ML/Research Engineer, Safeguards ML/Research Engineer focused on detecting and mitigating misuse of AI systems, building classifiers, monitoring for harms, evaluating agentic product safety, and conducting research on red-teaming and adversarial robustness. | AgentData | 9 |
| Research Engineer, Knowledge Team Research Engineer focused on redesigning how LLMs interact with external data sources by designing new information architectures and training models to use them. Responsibilities include implementing information architecture strategies, performing finetuning and RL, building knowledge base eval sets, and designing agentic search capabilities. Requires strong Python, ML research experience, and experience with LLMs. Experience with complex agentic systems, RAG, and distributed information retrieval is a plus. | AgentPost-train | 9 |
| Staff Software Engineer, Claude Code Software Engineer to build and maintain new agentic coding tools for developers, leveraging advanced LLM features like tool-use, chaining, and orchestration. Requires expertise in React, full-stack development, and hands-on experience with LLMs and prompt engineering. Experience with safety, security, or compliance requirements is a plus. | Agent | 8 |
| Product Manager, Claude Code Product Manager for Claude Tag, an agentic system integrated into collaboration tools, focusing on expanding its reach to new platforms and managing external partnerships. The role involves defining core interactions, owning strategy and roadmap, managing adoption, and working with engineers and partners. | Agent | 8 |
| Red Team Engineer, Safeguards This role focuses on adversarial testing and red teaming of AI systems and products to uncover vulnerabilities and ensure safety. It involves simulating sophisticated threat actors, researching novel testing approaches for capabilities like agent systems and tool use, and developing automated testing frameworks. The goal is to translate findings into concrete improvements and establish metrics for detection effectiveness. | Agent | 8 |
| Engineering Manager, Agent Runtime Platform Engineering Manager for the Agent Runtime Platform, responsible for provisioning and running compute and runtimes for internal agents. This role involves owning technical strategy, managing execution, prioritizing work, building and shipping the platform, defining secure agent execution, building reusable primitives, and ensuring infrastructure scalability and reliability. The role also requires collaboration with security teams and supporting internal teams in agent development. | AgentServe | 8 |
| Engineering Manager, Cybersecurity Products Engineering Manager for AI-powered cybersecurity products, leading a team to prototype and ship products using frontier models. The role involves setting technical direction, partnering with research, and staying close to customers. It requires hands-on technical involvement, product instincts, and scaling the team. | AgentShip | 8 |
| Manager of Applied AI Architecture, Startups Manager of Applied AI Architecture for startups, leading a team of technical architects to drive adoption of frontier AI and help startups build AI-native products using the Claude Developer Platform. Focuses on technical partnership, building teams, defining playbooks, and influencing product roadmap. | Agent | 8 |
| Staff Software Engineer, People Products Staff Software Engineer focused on building AI-native workflows and LLM-native features for internal people products at Anthropic. The role involves full-stack development, designing and implementing AI tools, evals, and prompts, and working directly with internal stakeholders to iterate quickly. Emphasis on autonomy, shipping products rapidly, and making independent product and architecture decisions in a low-structure environment. | Agent | 8 |
| Staff+ Software Engineer, Claude App Infrastructure Staff+ Software Engineer role focused on building the agentic layer and infrastructure for Claude App, enabling task execution, tool use, and safe interaction with external services. This involves designing and building sandboxed compute environments, state management for agent tasks, authentication/authorization, and observability tools for agent execution at scale. | AgentServe | 8 |
| Staff Software Engineer, Environments Infrastructure Staff Software Engineer focused on building and maintaining the infrastructure for reinforcement learning environments at Anthropic. This role involves designing APIs, frameworks, and tooling to support researchers in building and operating AI agents, with a focus on stateful distributed systems, agent runtimes, and productionizing research for AI capabilities improvement. | AgentData | 7 |
| Recruiting Solutions Engineer This role focuses on enabling recruiters to effectively use AI (Claude) in their hiring workflows. It involves educating users, providing technical support, and building prototypes for AI-assisted hiring processes. The engineer will translate recruiter-built skills into reusable assets and design evaluation approaches for these workflows. | Agent | 7 |
| Staff+ Software Engineer, Safeguards Review Tooling Staff+ Software Engineer for Anthropic's Safeguards Review Tooling team. This role focuses on building internal tools and platforms for safety investigators to identify and act on harmful behavior. Responsibilities include developing investigation tooling, a platform layer for reusable APIs and data storage, scaling review through automation with Claude, and partnering with various teams to ensure effective enforcement systems. The role emphasizes building guardrails for sensitive tools, instrumenting shipped tools, and ensuring tooling evolves with privacy and data retention commitments. Experience with full-stack or platform engineering, shipping internal tools, and cross-functional collaboration is required. Preferred qualifications include experience with trust and safety tooling, privacy/compliance constraints, integrating LLMs/agentic systems, and building developer platforms. | Agent | 7 |
| Technical Enablement Lead, Claude Platform This role focuses on enabling Anthropic's go-to-market teams by creating technical content, demos, and hands-on labs for the Claude Developer Platform. The Technical Enablement Lead will translate new platform capabilities into field-ready materials, deliver training, coach technical sellers, and provide feedback to the product team. The role requires hands-on experience with LLM APIs and agents, including tool use and agent frameworks, and experience in a customer-facing technical role. | Agent | 7 |
| Engineering Manager, Safeguards Review Tooling Engineering Manager for Anthropic's Safeguards Review Tooling team, focusing on building and scaling systems for AI safety investigation and enforcement. This role involves leading a team to develop tooling that supports human reviewers and integrates AI (Claude) for automation, with a strong emphasis on privacy, analytics, and a sandbox environment for rapid iteration. | AgentServe | 7 |
| Applied AI Architect, Commercial This role is a Pre-Sales architect focused on becoming a trusted technical advisor helping customers understand the value of Claude and how they can successfully integrate and deploy Claude into their technology stack. The role involves being a hands-on builder, creating reusable blueprints, demos, and enablement. Responsibilities include partnering with account executives, serving as the primary technical advisor, supporting customers building with the Claude API, shipping working code, building prototypes and proof-of-concepts, developing eval frameworks, writing near-production examples, building reusable blueprints and demos, guiding technical architecture decisions, helping customers integrate Claude, and helping customers develop evaluation frameworks. | Agent | 7 |
| Global Applied AI Architecture Lead, Beneficial Deployments Lead a global team of Applied AI Architects who partner with mission-driven organizations to deploy Claude, focusing on responsible and effective adoption to accelerate their missions. This role involves setting strategy, scaling the team, influencing product roadmaps, and representing Anthropic as a senior technical leader in high-impact partnerships. | Agent | 7 |
| Staff+ Software Engineer, Full-stack Staff+ Software Engineer, Full-stack at Anthropic, focusing on building and scaling AI products for enterprise customers. This role involves end-to-end ownership of products like Claude.ai, the Anthropic API, enterprise deployments, and specialized industry applications. Responsibilities include developing developer tools, enhancing enterprise workflows with features like plugins and retrieval, ensuring security and compliance, and driving user growth and monetization. The role requires a product-oriented mindset and technical leadership across the full stack. | Agent | 7 |
| Applied AI Architect, Government Technology Pre-Sales architect for Anthropic's Claude LLM, focusing on becoming a trusted technical advisor to GovTech companies. The role involves architecting LLM solutions, guiding customers from discovery to deployment, helping them develop evaluation frameworks, and integrating Claude into their technology stack. | Agent | 7 |
| Applied AI Architect, State and Local Government Pre-sales architect for state and local government, acting as a technical advisor to help agencies integrate and deploy Claude LLM into their technology stack. Focuses on architecting solutions, developing evaluations, and designing scalable architectures. | Agent | 7 |
| Software Engineer, Cybersecurity Products Software Engineer focused on building AI-powered cybersecurity products. This role involves prototyping, iterating based on customer feedback, collaborating with research teams on model capabilities, and engaging with customers to inform product direction. It sits at the intersection of research, product, and go-to-market. | Agent | 7 |
| Software Engineer, Safeguards Software Engineer focused on building safety and oversight mechanisms for AI systems, including monitoring, misuse prevention, and abuse detection. The role involves developing systems to detect unwanted model behaviors, enforce policies, and build scalable defenses. | Agent | 7 |