AI Frontier · AI lab
Anthropic currently has 151 active AI-related job listings, with a significant focus on roles related to agents, which constitute 32% of their openings. Engineering is the most frequent function, followed by Research. The majority of their hiring is concentrated in the United States. Frequent technical tags include evals, model_serving, and agent_orchestration, suggesting a focus on the practical deployment and management of AI systems.
Currently tracking 127 active AI roles, down 32% versus the prior 4 weeks. Primary focus: Agent · Engineering. Salary range $46k–$850k (avg $406k).
Anthropic currently has 149 active AI-related roles in our index. The most common open titles are: Regional Research Economist, Economic Research (2), Research Engineer, Machine Learning (RL Velocity) (2), Research Engineer, Production Model Post-Training (2), Staff Software Engineer, AI Reliability Engineering (2), Product Manager, Safeguards Rare Harms. Most positions are in Engineering and Research.
Anthropic's active AI hiring is concentrated in: agents (31%), serving infrastructure (15%), post-training (15%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
Anthropic is hiring AI talent in: United States (130 roles), United Kingdom (18 roles), Canada (6 roles), Ireland (3 roles).
Job postings at Anthropic most frequently mention: Machine Learning, AI Safety, Production ML Systems, System Design, Large Language Models (LLMs).
In the past 30 days, Anthropic has posted 24 new AI-related roles. That is a -27% change versus the prior 30 days (33 → 24).
| Title | Stage | AI score |
|---|---|---|
| Research Engineer, Machine Learning (Reinforcement Learning) Research Engineer focused on Reinforcement Learning to advance capabilities and safety of large language models. This role involves implementing novel approaches, contributing to research direction, and creating agentic models for tasks like computer use and autonomous software generation, while also improving reasoning abilities and developing prototypes. Key responsibilities include architecting RL infrastructure, designing training environments and methodologies, driving performance improvements, and collaborating across teams. | AgentPost-train | 10 |
| Research Engineer, Frontier Red Team (Autonomy) Research Engineer focused on building and evaluating autonomous AI systems and defensive agents to counter adversarial AI, with a focus on cyberphysical risks and AI safety. This role involves creating model organisms, developing defensive agents, and translating findings into policy-relevant demonstrations. |
| AgentEval Gate |
| 10 |
| Staff Software Engineer, Code RL Staff Software Engineer focused on the engineering aspects of reinforcement learning for AI coding capabilities, specifically creating and scaling agentic coding environments. The role involves designing frameworks and APIs for researchers, managing production RL runs, and improving the reliability and structure of research codebases. It emphasizes Python expertise, API design, and anticipating system failures. | AgentData | 9 |
| Staff+ Software Engineer, Enterprise AI Products Staff+ Software Engineer for Anthropic's Enterprise AI Products team, focusing on building organizational context and workflows (plugins, skills, connectors, webhook-triggered agents) to make Claude a daily-use tool for enterprise customers. This role involves technical leadership, end-to-end product delivery, customer interaction, and close collaboration with research to integrate model capabilities into production. | AgentShip | 9 |
| Lead, Frontier Red Team (Cyber) Lead a new team focused on researching the impacts of advanced AI models on cybersecurity, developing defenses, and steering the world through the next generation of cybersecurity. This role involves leading a research team, publishing frontier research, and building/deploying defenses to give defenders a permanent advantage. | AgentPost-train | 9 |
| Research Engineer, Chip Design RL (Reinforcement Learning) Research Engineer role focused on applying Reinforcement Learning to chip design, specifically for agentic RTL generation, design verification, and physical design optimization. The role involves inventing RL environments, optimizing EDA tools, conducting experiments, and delivering work into research and production training runs. Requires expertise in ASIC/FPGA design and familiarity with EDA tools, with strong candidates having RL experience. | AgentData | 9 |
| [Pipeline] Staff+ Software Engineer, Developer Acceleration Staff+ Software Engineer to join Developer Acceleration team, responsible for the infrastructure enabling thousands of employees to be productive via agents. The role will define the future of agentic productivity at scale, own technical strategy, build and ship the runtime platform and tooling, write evals to benchmark agent behaviors, and ensure infrastructure scalability and reliability. | Agent | 9 |
| Research Engineer, Computer Use Research Engineer focused on teaching AI models (Claude) to perceive, use, and understand computer interfaces, enabling them to reliably and safely operate real software. This involves designing experiments, developing evaluation frameworks, building RL training environments, and collaborating with training and product teams to integrate research advances into production. | AgentPost-train | 9 |
| Manager, Applied AI Engineering, Life Sciences (Beneficial Deployments) Manager for Applied AI Engineering focused on Life Sciences, leading a team to build and deploy AI solutions (agents, integrations, tools) for scientific organizations. The role involves deep customer engagement, understanding scientific workflows, and ensuring reliable, reproducible access to biological data for AI agents, with a strong emphasis on safety and responsible deployment in a regulated-like environment. | Agent | 9 |
| Software Engineer, Safeguards Evals Software Engineer role focused on building and owning the evaluation infrastructure for an agentic investigation system. This involves designing experiments, constructing high-quality eval datasets, measuring agent performance, analyzing coverage gaps, and productionizing research into release pipelines. The role also involves building tooling for policy experts and constructing RL environments to improve safety investigation capabilities. | AgentEval Gate | 9 |
| Product Manager, Claude Code Model Performance Product Manager for Anthropic's Claude Code Model Performance team, responsible for driving model launches, building agentic evals, and translating research improvements into developer-facing outcomes. Requires experience building agentic evals, a systems thinking approach, and comfort with both research and engineering. | AgentEval Gate | 9 |
| Manager of Forward Deployed Engineering Manager of Forward Deployed Engineering at Anthropic, responsible for leading a team that embeds with strategic customers to ship production AI applications and agent deployments built on Claude. This player-coach role involves hiring, developing, and mentoring engineers, overseeing customer engagements, reviewing technical architectures, and collaborating with cross-functional teams to drive AI transformation for enterprise clients. | Agent | 9 |
| Forward Deployed Engineer The Forward Deployed Engineer (FDE) role at Anthropic focuses on embedding with strategic customers to drive the adoption of advanced AI applications. This role involves building production applications using Claude models within customer systems, delivering technical artifacts like sub-agents and agent skills, and providing deployment support. The FDE will work closely with post-sales, product, and engineering teams, combining engineering expertise with customer-facing skills to solve complex business challenges and represent Anthropic's mission. | Agent | 9 |
| Prompt Engineer, Agent Prompts & Evals This role focuses on prompt engineering and evaluation development for AI-first products and features, bridging model capabilities with user experience. It involves designing, testing, and optimizing prompts, building evaluation suites, supporting model launches, and contributing to prompt development frameworks. The role requires strong software engineering skills, LLM and prompt engineering experience, and understanding of evaluation methodologies. | AgentEval Gate | 9 |
| Research Engineer / Scientist, Frontier Red Team (Cyber) Research Engineer/Scientist focused on AI-enabled cybersecurity, developing tools and frameworks for autonomous vulnerability discovery, remediation, malware detection, and pentesting. Designs and runs experiments to evaluate AI cyber capabilities and builds infrastructure for AI systems operating in security environments. Translates findings into demonstrations for policymakers and collaborates with external experts. Senior candidates will set research strategy and own the technical roadmap. | AgentEval Gate | 9 |
| Research Engineer, Frontier Red Team (Hardware Lead) Research Engineer focused on leading hardware research for frontier AI safety, specifically interfacing LLMs with robotics and cyberphysical systems. The role involves designing and building systems, developing evaluations, creating training environments, and demonstrating capabilities to inform policy and build defenses against advanced AI risks. | AgentEval Gate | 9 |
| Applied AI Engineer, Startups Applied AI Engineer role focused on advising and partnering with AI-native startups to build on the Claude Developer Platform. Responsibilities include technical guidance, developing evaluation frameworks, designing scalable architectures, and creating technical resources to help startups succeed with Claude. Requires production experience with LLM-powered applications, agent architectures, and evaluation frameworks. | Agent | 9 |
| Research Engineer, Universes Research Engineer role focused on building next-generation agentic environments for training AI models. This role involves implementing novel approaches, contributing to research direction, designing training environments and methodologies, and building evaluations for capable and safe agentic AI. It blends research and engineering, with a focus on reinforcement learning and complex, long-horizon agentic tasks. | AgentPost-train | 9 |
| Forward Deployed Engineer Forward Deployed Engineer (FDE) embeds with strategic customers to drive AI adoption by shipping advanced AI applications built on Claude models. Collaborates with customer teams, Post-Sales, Product, and Engineering to solve business challenges using frontier AI, focusing on safety and reliability. Operates autonomously, builds customer relationships, and identifies new AI deployment opportunities. | Agent | 9 |
| ML/Research Engineer, Safeguards ML/Research Engineer focused on detecting and mitigating misuse of AI systems, building classifiers, monitoring for harms, evaluating agentic product safety, and conducting research on red-teaming and adversarial robustness. | AgentData | 9 |
| Research Engineer / Scientist, Tool Use Safety Research Engineer/Scientist focused on advancing the frontier of safe tool use in AI models, specifically addressing prompt injection, data exfiltration, adversarial attacks, and autonomous agent behavior with large tool sets. The role involves designing and implementing RL methodologies, building evaluations, and shipping research advances into production models, with a strong emphasis on safety and reliability. | AgentPost-train | 9 |
| Research Engineer / Scientist, Tool Use Research Engineer/Scientist focused on advancing the frontier of tool use for AI agents, aiming to improve accuracy, reliability, safety, and efficiency in complex workflows. The role involves defining research agendas, designing RL methodologies, building evaluations, and shipping research advances into production models, with a strong emphasis on safety and collaboration. | AgentPost-train | 9 |
| Machine Learning Systems Engineer - Data & Evaluation, Horizons Machine Learning Systems Engineer on the Horizons team, focusing on building software infrastructure for AI models to use tools effectively and measure performance. This involves extending the agent framework, creating evaluations, managing training data pipelines, and applying data science techniques to improve model capabilities. The role combines software development with empirical analysis to advance model performance and capabilities, working closely with research and production teams. | AgentEval Gate | 9 |
| Research Engineer, Knowledge Team Research Engineer focused on redesigning how Claude interacts with external data sources by designing new information architectures and training language models to use them. This includes performing finetuning and RL, building knowledge base eval sets, and designing/evaluating agentic search capabilities. | AgentPost-train | 9 |
| Research Engineer, Knowledge Team Research Engineer focused on redesigning how LLMs interact with external data sources by designing new information architectures and training models to use them. Responsibilities include implementing information architecture strategies, performing finetuning and RL, building knowledge base eval sets, and designing agentic search capabilities. Requires strong Python, ML research experience, and experience with LLMs. Experience with complex agentic systems, RAG, and distributed information retrieval is a plus. | AgentPost-train | 9 |
| Research Engineer, Agents Research Engineer focused on advancing agentic AI systems, involving finetuning Claude for agentic tasks, developing tools for agents (memory, communication), prompt engineering, automated evaluation, and optimizing data mixes for model training. The role also involves creating and maintaining infrastructure for prompt iteration and testing. | AgentPost-train | 9 |
| Staff Software Engineer, Claude Code Software Engineer to build and maintain new agentic coding tools for developers, leveraging advanced LLM features like tool-use, chaining, and orchestration. Requires expertise in React, full-stack development, and hands-on experience with LLMs and prompt engineering. Experience with safety, security, or compliance requirements is a plus. | Agent | 8 |
| Product Manager, Claude Code Product Manager for Claude Tag, an agentic system integrated into collaboration tools, focusing on expanding its reach to new platforms and managing external partnerships. The role involves defining core interactions, owning strategy and roadmap, managing adoption, and working with engineers and partners. | Agent | 8 |
| Applied AI Engineer, Beneficial Deployments Applied AI Engineer focused on deploying AI for social impact partners, advising on AI systems, building ecosystem tooling, and prototyping agents. Requires production experience with LLM applications and a builder mindset. | Agent | 8 |
| Red Team Engineer, Safeguards This role focuses on adversarial testing and red teaming of AI systems and products to uncover vulnerabilities and ensure safety. It involves simulating sophisticated threat actors, researching novel testing approaches for capabilities like agent systems and tool use, and developing automated testing frameworks. The goal is to translate findings into concrete improvements and establish metrics for detection effectiveness. | Agent | 8 |
| Engineering Manager, Agent Runtime Platform Engineering Manager for the Agent Runtime Platform, responsible for provisioning and running compute and runtimes for internal agents. This role involves owning technical strategy, managing execution, prioritizing work, building and shipping the platform, defining secure agent execution, building reusable primitives, and ensuring infrastructure scalability and reliability. The role also requires collaboration with security teams and supporting internal teams in agent development. | AgentServe | 8 |
| Engineering Manager, Agent Runtime Platform Engineering Manager for the Agent Runtime Platform, responsible for provisioning and running compute and runtimes for internal agents. The role involves owning technical strategy, managing execution, prioritizing work, building and shipping the runtime platform and tooling, defining secure agent execution, building reusable primitives, and ensuring infrastructure scalability and reliability. Requires strong experience in building and operating large-scale platforms/distributed systems and technical management, with a focus on agent platforms and security-sensitive infrastructure. | AgentServe | 8 |
| Manager, Applied AI Engineering Manager of Applied AI Engineers leading a team that advises enterprise customers on adopting and deploying LLM APIs (Claude). Responsibilities include hiring, coaching, setting technical direction, developing evaluation frameworks, and channeling field insights back into product development. Focuses on advanced implementation patterns for LLMs, including prompting and agentic systems. | Agent | 8 |
| Applied AI Engineer Applied AI Engineer role focused on being a technical advisor to customers deploying Claude (LLM). Responsibilities include guiding architecture, developing evaluation frameworks, and implementing cutting-edge LLM patterns via API. Requires strong Python skills and production experience with LLMs, including agent development and retrieval frameworks. | Agent | 8 |
| Product Engineer, Computer Use Product Engineer role focused on building and shipping AI-powered computer-use and browser-control product surfaces. This involves full-stack development, agent harness, and working with LLM APIs and agent frameworks. The role requires end-to-end ownership and iteration based on user feedback, with a focus on reliability and robustness of the agent harness. | Agent | 8 |
| Applied AI Architect (Startups) This role focuses on partnering with startups to help them build and scale AI solutions using Anthropic's Claude Developer Platform. The architect will guide technical decisions, win evaluations, and provide feedback to product and engineering teams. Requires strong technical expertise in LLM application development and deployment, with a customer-facing background. | Agent | 8 |
| Engineering Manager, Cybersecurity Products Engineering Manager for AI-powered cybersecurity products, leading a team to prototype and ship products using frontier models. The role involves setting technical direction, partnering with research, and staying close to customers. It requires hands-on technical involvement, product instincts, and scaling the team. | AgentShip | 8 |
| Manager of Applied AI Architecture, Startups Manager of Applied AI Architecture for startups, leading a team of technical architects to drive adoption of frontier AI and help startups build AI-native products using the Claude Developer Platform. Focuses on technical partnership, building teams, defining playbooks, and influencing product roadmap. | Agent | 8 |
| Technical Specialist, Claude Code This role focuses on driving adoption of Anthropic's AI products (Claude Code, Claude Enterprise) within enterprise customers. It involves technical enablement, delivering workshops, building demo apps, and supporting strategic pilots, acting as a trusted technical voice to developers and leaders. The role bridges pre-sales and post-sale engagement to ensure deep integration and usage of AI capabilities. | Agent | 8 |
| Applied AI Architect, Startups This role is for an Applied AI Architect on the Startups Applied AI team. The primary responsibility is to act as a technical partner for startups, helping them build on Anthropic's Claude Developer Platform. This involves architecting LLM solutions, winning technical evaluations, and guiding customers from discovery through deployment. The role requires deep technical expertise in building and deploying LLM-powered applications, understanding AI engineering best practices, and communicating complex AI concepts to technical founders and engineering teams. | Agent | 8 |
| Technical Deployment Lead Technical Deployment Lead responsible for delivering customized AI agent solutions to enterprise clients in highly regulated industries. This role involves managing end-to-end engagements, from scoping and technical discovery to production launch, and requires strong client-facing and technical leadership skills to guide the deployment of AI agents within critical business processes. | Agent | 8 |
| Applied AI Engineer, Beneficial Deployments Applied AI Engineer role focused on deploying AI to mission-driven organizations, advising on evals and agent architectures, building ecosystem tooling, and prototyping new agents. Requires production experience with LLM applications and a builder mindset. | AgentEval Gate | 8 |
| Applied AI Engineer Applied AI Engineer role focused on being a technical advisor to customers adopting Claude LLMs. Responsibilities include guiding architecture design, developing evaluation frameworks, and advising on implementation patterns for LLMs via API. Requires production experience with LLMs, strong Python skills, and expertise in common LLM implementation patterns. | Agent | 8 |
| Applied AI Engineer Applied AI Engineer to serve as a technical advisor for companies building on the Claude Developer Platform, focusing on implementation, agent design, and building AI applications. Responsibilities include technical engagements, developing evaluation frameworks, designing architectures, and creating technical resources. | Agent | 8 |
| Staff Software Engineer, People Products Staff Software Engineer focused on building AI-native workflows and LLM-native features for internal people products at Anthropic. The role involves full-stack development, designing and implementing AI tools, evals, and prompts, and working directly with internal stakeholders to iterate quickly. Emphasis on autonomy, shipping products rapidly, and making independent product and architecture decisions in a low-structure environment. | Agent | 8 |
| Applied AI Engineer Applied AI Engineer role focused on being a technical advisor to customers adopting Anthropic's Claude LLM. Responsibilities include guiding customers through architecture design, developing evaluation frameworks, and implementing cutting-edge LLM patterns via API. Requires strong programming skills (Python) and production experience with LLMs, including agent development and evaluation. | Agent | 8 |
| Applied AI Engineer, Life Sciences (Beneficial Deployments) Applied AI Engineer role focused on deploying Claude in life sciences to accelerate scientific progress. The role involves partnering with research institutions, building agents integrated into scientific workflows, and developing ecosystem infrastructure like MCP servers, benchmarks, and agent skills. The goal is to make Claude a go-to tool for the life sciences ecosystem, from discovery to pharma pipelines. | Agent | 8 |
| Forward Deployed Engineer, Federal Civilian Forward Deployed Engineer for Anthropic's Applied AI team, embedding with federal civilian customers to drive AI adoption and ship advanced AI applications built on Claude models. Responsibilities include building production applications, delivering technical artifacts like sub-agents, providing deployment support, identifying repeatable patterns, and building customer relationships. Requires strong programming skills, production LLM experience (prompt engineering, agent development, evaluation, scaling), and experience with government agencies. | Agent | 8 |
| Applied AI Engineer, Beneficial Deployments Applied AI Engineer focused on deploying AI to mission-driven organizations, advising on AI applications like evals and agent architectures, and building infrastructure to scale impact. Requires production experience with LLM applications and a builder mindset. | Agent | 8 |
| Staff+ Software Engineer, Claude App Infrastructure Staff+ Software Engineer role focused on building the agentic layer and infrastructure for Claude App, enabling task execution, tool use, and safe interaction with external services. This involves designing and building sandboxed compute environments, state management for agent tasks, authentication/authorization, and observability tools for agent execution at scale. | AgentServe | 8 |