Data AI · ML experiment tracking
Currently tracking 35 active AI roles, down 14% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $92k–$341k (avg $211k).
Weights & Biases currently has 47 active AI-related job listings. The majority of these roles, 51%, are focused on serving infrastructure, with an additional 34% dedicated to agents. Engineering is the primary function being hired for, with the United States being the dominant hiring country. The company is frequently seeking candidates with experience in model serving, inference infrastructure, and LLM observability. Over the last 30 days, there has been a 75% decrease in new AI roles posted, with 4 new positions compared to 16 in the preceding 30-day period.
Weights & Biases currently has 45 active AI-related roles in our index. The most common open titles are: Account Solution Architect (5), Account Solutions Architect - Greenfield (2), Account Solution Architect - Financial Services, Applied AI Engineer, Inference, Forward Deployed Engineer, AI Agents. Most positions are in Engineering and Product.
Weights & Biases's active AI hiring is concentrated in: serving infrastructure (51%), agents (38%), application (4%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
Weights & Biases is hiring AI talent in: United States (43 roles), United Kingdom (2 roles), Canada (1 role).
Job postings at Weights & Biases most frequently mention: GPU Computing, Kubernetes, Production ML Systems, ML Ops, Storage Systems.
In the past 30 days, Weights & Biases has posted 5 new AI-related roles. That is a -50% change versus the prior 30 days (10 → 5).
| Title | Stage | AI score |
|---|---|---|
| Senior Software Engineer - GPU Kernel Authoring & Optimization Senior Software Engineer focused on authoring and optimizing GPU kernels for large-scale LLM inference serving. The role involves deep understanding of GPU architecture, CUDA programming, and performance benchmarking to achieve maximum throughput and minimum latency. Responsibilities include kernel development, optimization, benchmarking, and contributing to the inference stack's performance and reliability. | Serve | 9 |
| Senior Applied Research Engineer The OpenPipe team at CoreWeave is building tools to help agents learn from experience, focusing on solving bottlenecks for self-improving agents. This role involves applied research to enable continuous learning in production, working with LLM post-training techniques and agent development. | Post-trainAgent | 9 |
| Staff Applied Research Engineer CoreWeave's OpenPipe team is building tools for agents to learn from experience, aiming to make them reliable for autonomous long tasks. The role involves applied research to solve obstacles in continuous learning for production, focusing on LLM post-training, reinforcement learning, and agent development. The team has released tools like ART, RULER, and Serverless RL, and is seeking someone with deep expertise in LLM training and a strong research background to contribute to self-improving agents. | Post-trainAgent | 9 |
| VP of Product, Research and Training Infrastructure VP of Product for Research and Training Infrastructure at an AI cloud provider. This role owns the product strategy and engineering execution for services powering AI research labs, focusing on specialized orchestration, evaluation, and iteration tools for massive-scale pre-training and post-training. Key responsibilities include evolving orchestration tools (SUNK), developing automated training-based evaluation frameworks, and building infrastructure for RL/RLHF pipelines. Requires deep knowledge of HPC, distributed training, and supporting frontier model research. | PretrainPost-train | 9 |
| Forward Deployed Engineer, AI Agents This role involves embedding within customer AI teams to help them build and deploy AI agents using the W&B Weave developer toolkit. It's a hands-on, customer-facing engineering position focused on designing, prototyping, and productionizing AI agents, acting as a technical advisor, developing reference implementations, and gathering customer feedback to influence the product roadmap. | Agent | 8 |
| Account Solutions Architect - Greenfield CoreWeave is seeking an Account Solutions Architect to be the technical partner for prospective AI customers, helping them design, build, and scale AI workloads on CoreWeave's platform. This role involves leading PoCs, acting as a trusted advisor on AI best practices, and translating customer needs into technical solutions. The ideal candidate has hands-on experience with training, fine-tuning, evaluating, and deploying deep learning models, particularly LLMs, and designing production LLM-powered applications. | AgentPost-train | 8 |
| Staff Specialist Field Engineer, Autonomous Vehicles Staff Specialist Field Engineer for Autonomous Vehicles at CoreWeave, focusing on deploying AI solutions in enterprise engineering organizations. The role involves establishing and leading the AV vertical, defining engagement strategies, setting technical standards, and leading customer engagements from initial contact to deployment. Requires deep ML expertise to build, validate, and deploy solutions independently using CoreWeave's AI platform, and translating field observations into product signals. | ShipAgent | 8 |
| Staff Specialist Field Engineer, Robotics Staff Specialist Field Engineer for Robotics at CoreWeave, focusing on deploying AI/ML solutions for robotics customers. The role involves establishing the robotics vertical, defining engagement strategies, leading customer engagements from scoping to deployment, and building/iterating on customer-facing applications using CoreWeave's AI platform. Requires deep understanding of robotics systems, ML capabilities, and the ability to translate field observations into product signals. | ShipAgent | 8 |
| Account Solution Architect This role is for an Account Solutions Architect at CoreWeave, an AI-focused cloud provider. The architect will be the technical partner for prospective customers, designing demos, leading PoCs, and advising on best practices for AI workloads, including training, fine-tuning, evaluating, and deploying deep learning models and LLM-powered applications. The role requires strong Python skills, experience with major cloud platforms, and the ability to solve complex technical problems. | AgentPost-train | 8 |
| Principal Solution Specialist, Infrastructure This role focuses on bringing CoreWeave's AI developer services, such as MLOps platforms and LLM observability tools, to market. It involves defining commercial and technical strategies, driving adoption with early customers, and translating field insights into product roadmap requirements. The role requires deep expertise in the ML development lifecycle, LLM application patterns, and MLOps ecosystem, with a focus on experiment tracking, model lifecycle governance, and observability. | AgentEval Gate | 8 |
| Applied AI Engineer, Inference Applied AI Engineer focused on optimizing and benchmarking the inference serving platform for AI models, working on latency, throughput, and reliability. | Serve | 8 |
| Principal Engineer - Perf and Benchmarking Principal Engineer role focused on leading the Benchmarking & Performance team at CoreWeave, a cloud provider for AI. The role involves defining strategy, leading end-to-end MLPerf submissions (Training & Inference), designing and implementing a Kubernetes-native benchmarking service for latency and throughput, and building CI/CD pipelines for scale. It requires deep expertise in distributed systems, GPU performance, model-serving stacks, and Kubernetes, with a focus on achieving industry-leading performance data and publications. | ServeEval Gate | 8 |
| Account Solutions Architect - Greenfield This role is a customer-facing technical partner for existing AI customers, focusing on deepening platform adoption, identifying expansion opportunities, and serving as a trusted advisor for scaling AI workloads in production. It involves hands-on experience with training, fine-tuning, evaluating, and deploying deep learning models and LLM-powered applications. | AgentServe | 7 |
| Senior Software and AI Engineer This role focuses on developing AI agents and first-party applications for CoreWeave's internal ecosystem, integrating with enterprise systems and using frameworks like LangChain and LangGraph. The engineer will work across the full stack, from agentic frameworks to backend services and frontend interfaces, with a strong emphasis on building production-grade AI/LLM-based applications. | Agent | 7 |
| Solutions Architect Solutions Architect for CoreWeave Federal, focusing on designing and deploying AI workloads for U.S. government agencies within regulated environments. This role requires expertise in AI/ML model deployment, agentic applications, and compliance with federal standards like FedRAMP and DoD Impact Levels. | AgentServe | 7 |
| Technical Program Manager, Inference CoreWeave is seeking a Technical Program Manager (TPM) focused on inference to join their AI/ML Platform Services team. This role will drive end-to-end program management for inference platform initiatives, including reliability, customer onboarding, launch readiness, and runtime optimization. The TPM will lead cross-functional programs to ensure the successful delivery of scalable, reliable, and high-performance inference services, working closely with engineering, product, and infrastructure teams to improve how these services are launched, onboarded, operated, and optimized. The role requires strong technical fluency in distributed inference systems, GPU compute, and cloud-native architectures, with a focus on measurable improvements in reliability, performance, and customer delivery. | Serve | 7 |
| Technical Program Manager - Performance & Benchmarking This role is for a Technical Program Manager focused on Performance & Benchmarking within an AI/ML Platform Services organization. The TPM will drive end-to-end program execution for initiatives related to infrastructure validation, performance testing, benchmark execution, and observability, ensuring CoreWeave's infrastructure is performant and stable for AI workloads. The role involves partnering with engineering, infrastructure, product, and go-to-market teams to improve workload performance, validate new environments, and create visibility into system performance. | Serve | 7 |
| Solution Specialist, AI Runtime Services This role focuses on bringing new AI runtime services, such as model serving and sandboxes, to market. It involves driving initial customer adoption, gathering feedback for the product roadmap, and enabling sales teams to position these services. The role requires deep expertise in AI runtime infrastructure, including serving frameworks, inference optimization, and execution isolation. | Serve | 7 |
| Staff Software Engineer, Cluster Orch (SUNK) Staff Software Engineer role focused on advancing CoreWeave's orchestration platform (SUNK - Slurm on Kubernetes) for AI training and inference at scale. The role involves technical leadership, architectural direction, and ensuring efficient, reliable workload execution across large GPU clusters. | Serve | 7 |
| Account Solution Architect Account Solutions Architect for a financial services customer portfolio, focusing on AI/ML and LLM workloads on CoreWeave's cloud platform. Responsibilities include deepening platform adoption, designing end-to-end solutions, guiding customers on AI lifecycle, and resolving technical challenges. Requires Python proficiency, experience with deep learning models, LLM applications, and financial services customer engagement. | AgentPost-train | 7 |
| Account Solution Architect - Financial Services Account Solutions Architect for Financial Services at CoreWeave, a cloud provider focused on AI workloads. This role involves being a technical partner to existing financial services customers, helping them deepen platform adoption, identify expansion opportunities, and scale their AI workloads in production. The role requires understanding customer needs around model development, deployment, governance, and infrastructure efficiency, and working with sales, product, and engineering teams to ensure customer success. | AgentServe | 7 |
| Account Solution Architect Account Solutions Architect for financial services customers, focusing on scaling AI/ML workloads on CoreWeave's cloud platform. Responsibilities include technical partnership, platform adoption, identifying expansion opportunities, and advising on model development, deployment, and infrastructure efficiency for AI/ML-intensive organizations. | AgentServe | 7 |
| Account Solution Architect Account Solutions Architect for financial services customers, focusing on deepening platform adoption for AI/ML workloads, identifying expansion opportunities, and serving as a trusted advisor for production AI scaling. Requires hands-on experience with training, fine-tuning, evaluating, and deploying deep learning models and LLM-powered applications. | AgentServe | 7 |
| Account Solution Architect Account Solution Architect for CoreWeave, focusing on AI/ML infrastructure and MLOps solutions for customers in Northern EMEA. The role involves technical discovery, solution design, proof-of-concept engagements, and acting as a customer advocate to internal teams. Requires strong knowledge of ML training/inference, MLOps platforms, and underlying infrastructure like GPUs, networking, and Kubernetes. | ServeData | 7 |
| Staff Software Engineer, Inference Staff Software Engineer on the Inference Platform Team at CoreWeave, focusing on building and operating a Kubernetes-native inference platform for AI workloads. The role involves technical leadership in architecture, performance optimization (latency, throughput, GPU utilization), and system reliability for low-latency, high-throughput systems at massive scale, with deep work in distributed systems and Kubernetes infrastructure. | Serve | 7 |
| Staff Technical Program Manager - Cluster Orchestration & Applied Training Staff Technical Program Manager to lead cross-functional programs for AI/ML Platform Services, focusing on Cluster Orchestration (scheduling, launching, managing AI workloads) and Applied Training (enabling researchers to use infrastructure for pre-training, fine-tuning, RL, evaluations). The role involves partnering with engineering, product, and research teams to improve workload execution and user interaction with training platforms, driving delivery across various AI training workflows and ensuring successful launches and operational ownership. | ServePost-train | 7 |
| Principal Engineer, Cluster Orchestration CoreWeave is seeking a Principal Engineer to lead the design and evolution of their AI infrastructure's cluster orchestration systems, including Slurm, Kubernetes, and SUNK. This role involves defining long-term architecture, solving scaling problems, and ensuring the reliability and efficiency of GPU resource utilization for AI training and inference workloads. | Serve | 7 |
| Senior Software Engineer, Observability Insights Senior Software Engineer to lead development of agentic interfaces and product experiences for AI system observability, focusing on multi-tenant APIs, Grafana, and tool servers. Requires experience in backend systems, distributed APIs, reliability engineering, and agentic applications/LLM features. | AgentServe | 7 |
| Senior Software Engineer I, Inference CoreWeave is seeking a Senior Software Engineer to own and improve their Kubernetes-native inference platform, focusing on latency, throughput, and reliability. The role involves leading design, implementing optimizations, strengthening incident posture, and mentoring junior engineers. Requires experience with distributed systems, Kubernetes, and inference internals. | Serve | 7 |
| Sr. Software Engineer - Perf and Benchmarking Senior Software Engineer focused on performance and benchmarking of AI infrastructure, including Kubernetes-native services, MLPerf runs, and model-serving stacks. The role involves building and improving services to measure latency, throughput, and cost, and ensuring reproducible benchmarking processes. | ServeEval Gate | 7 |
| Sr. Engineering Manager, Inference Senior Engineering Manager for AI/ML Platform team at CoreWeave, focusing on productizing and operating their inference offering. The role involves leading a team to ensure service reliability, observability, operational excellence, and developer experience for AI workloads. Responsibilities include roadmap execution, engineering processes, cross-functional partnerships, and building a culture of technical excellence. | Serve | 7 |
| Software Engineer, Inference AI/ML Software Engineer focused on improving the latency, reliability, and cost of model serving on a GPU platform, working with services like Triton, vLLM, and TensorRT-LLM. | Serve | 7 |
| Senior Software Engineer II, Inference Senior Software Engineer II focused on owning and optimizing CoreWeave's Kubernetes-native inference platform to meet strict P99 SLAs at scale. Responsibilities include leading design reviews, implementing advanced optimizations for latency and throughput, strengthening incident posture, and mentoring junior engineers. Requires strong experience in distributed systems, Python/Go, networked systems performance, Kubernetes, and ML inference internals. | Serve | 7 |
| Senior Systems Engineer, OS Automation Senior Systems Engineer focused on automating and scaling Linux OS and Kernel build pipelines, with a strong emphasis on integrating AI/ML technologies like LLMs, RAG, and predictive modeling to create AI-native infrastructure, smart CI/CD, auto-remediation, and predictive regression detection. | ServeAgent | 7 |
| Staff Product Manager - Container Services Staff Product Manager for Container Services at CoreWeave, focusing on their Kubernetes Service and other container-based solutions for AI workloads like agents and RL. The role involves driving product strategy, roadmap, and execution, market research, and go-to-market strategies. | — | 5 |
| Senior Analyst, RevOps - Field Planning and Performance This role is for a Senior Analyst, RevOps - Field Planning and Performance at CoreWeave, a cloud provider for AI. The role focuses on analyzing sales performance, identifying improvement opportunities, and supporting strategic initiatives. It involves working with various teams like Sales Leadership, Marketing, RevOps, and Finance. A key aspect is leveraging AI tools to enhance RevOps functions, reporting, and forecasting, aiming to build an AI-driven RevOps team. The role requires strong analytical skills, experience with CRM and BI tools, and a proactive, AI-first mindset. | — | 5 |
| Field Business Development Manager This role is for a Field Business Development Manager focused on managing strategic partner relationships for CoreWeave, an AI cloud provider. The role involves driving joint pipeline, co-sell motions, and customer success with partners like NVIDIA, AWS, Google Cloud, and Microsoft Azure. It requires experience in partner management, business development, and go-to-market execution within the cloud or enterprise software space, with a preference for experience in GPU computing, AI/ML infrastructure, Model Ecosystem, and MLOps. | — | 5 |
| Director, Partner Marketing Director of Partner Marketing to build and lead CoreWeave’s joint go-to-market strategy across its technology and ecosystem partners, focusing on AI infrastructure and developer platforms. | — | 5 |
| Business Development Manager- Physical AI This role focuses on business development and strategic partnerships within the Physical AI (robotics, autonomous systems, industrial automation, embodied AI) compute vertical. The individual will lead partnerships with key players in this space, including simulation ISVs, robotics platform companies, OEMs, and the NVIDIA Physical AI ecosystem. Responsibilities include defining joint strategy, product integrations, co-innovation, and go-to-market motions, as well as identifying and closing compute opportunities across simulation, training, and inference workloads. The role requires deep domain knowledge of Physical AI and familiarity with the NVIDIA Physical AI toolchain. | — | 5 |
| Senior Director of Compute Services Senior Director of Compute Services responsible for the infrastructure supporting internal and external customer platforms, delivering compute capacity to AI labs. This role defines and drives the engineering roadmap for compute products, leads engineering teams, owns service availability and reliability, partners with Product Management, and oversees secure, scalable systems. Requires extensive experience in software/infrastructure engineering, leadership, and delivering at-scale compute products, preferably in an AI/ML environment. | — | 5 |
| Operations Enablement Analyst, Data Center Operations This role focuses on analyzing and improving data center operations by creating standard operating procedures, runbooks, and training content. While it leverages AI/LLM tools for analysis, the core function is operational enablement and documentation, not direct AI/ML model development or deployment. | — | 5 |
| Senior Software Engineer Senior Software Engineer role focused on the reliability, performance, and scalability of a Kubernetes-based data platform that supports internal AI workloads. The role involves architecting and operating highly available, multi-region systems, optimizing CI/CD pipelines, and ensuring robust observability and security. While the platform supports AI workloads, the core craft is infrastructure and platform engineering, not direct AI/ML model development. | — | 5 |
| Director of Legal Operations The Director of Legal Operations will support the technology and processes that power the Legal and Government Affairs (LGA) team's day-to-day work, from the legal tech stack and vendor management to AI adoption and operational metrics. This role will lead LGA's continued transformation into an innovative, leading-edge AI and technology-enabled organization, managing a Legal Operations team to serve as the connective tissue between LGA and global business partners. Responsibilities include managing legal tools and platforms, professional services and vendor management, LGA budget, metrics and reporting, defining strategy and operating models, cross-vertical support, team leadership, and cross-functional partnership. | — | 5 |
| Senior Product Manager, Employee Experience This role focuses on building and scaling an AI-driven intranet and internal digital experience platforms for a rapidly growing company. The goal is to create a unified, intelligent front door for company knowledge, tools, and workflows, integrating elements like company events, Slack, and enterprise search. The role involves defining strategy, governance, and ensuring AI is a core design principle for internal tools. | — | 5 |
| Product Marketing Manager Product Marketing Manager for CoreWeave software products (W&B Models, W&B Weave) focused on AI researchers and developers. The role involves defining narratives, product launches, and go-to-market strategy for AI model and agent development, evaluation, and operations. It sits at the intersection of AI research, model/agent development, and production AI operations, translating technical capabilities into compelling stories. Responsibilities include positioning, messaging, GTM strategy, customer advocacy, market insights, sales enablement, and metrics tracking for products spanning experiment tracking, model evaluation, AI observability, agent development, and MLOps. The role emphasizes accelerating the agent loop and shaping the future of agent workflows. | — | 5 |
| Employee Lifecycle Partner This role is for an Employee Lifecycle Partner at an AI infrastructure company. The role focuses on end-to-end lifecycle operations, data integrity in Workday, employee inquiry handling, and process improvement. It involves partnering with Payroll, Legal, and Total Rewards, and contributing to operational readiness for expansion and federal contracts. The ideal candidate will have experience in People Operations, multi-state US employment law, Workday, and process improvement, with a preference for federal contractor experience and AI tool usage. | — | 5 |
| Senior Product Manager, Supply Chain Product Manager for Supply Chain Systems at CoreWeave, focusing on defining strategy and roadmap for enterprise applications supporting GPU cloud infrastructure and hyperscale data center operations. The role involves leading the product lifecycle, driving automation, and leveraging AI/GenAI technologies for measurable business outcomes in planning, procurement, inventory, logistics, and warehouse operations. | — | 5 |
| Senior Executive Talent Sourcer (Contract) – Business, Operations & IT This role is for a Senior Talent Sourcer (Contract) at CoreWeave, an AI cloud company. The position focuses on high-impact hiring for executive and senior-level roles across Operations, IT, Revenue, GTM, G&A, and Marketing functions. The ideal candidate will act as a hybrid sourcer, researcher, and talent partner, building executive pipelines, generating market intelligence, and utilizing modern AI sourcing tools. The role requires strong research capabilities, executive-level messaging, and the ability to manage multiple searches in a fast-paced environment. | — | 5 |
| Learning Partner- Technical Development This role is for a senior Learning Partner Lead responsible for designing and delivering technical learning programs for employees in an AI infrastructure environment. The role requires strong instructional design skills, the ability to partner with technical stakeholders, and a genuine curiosity about AI infrastructure to translate complex concepts into accessible learning experiences. The individual will lead program design, manage stakeholder partnerships, and define evaluation frameworks for learning initiatives. | — | 5 |
| Senior Manager, Production Engineering Senior Manager of Production Engineering to lead and expand the SRE team and practices, focusing on the reliability, performance, and operational efficiency of a large-scale, distributed cloud infrastructure, with experience in AI infrastructure environments and GPU-accelerated workloads. | — | 5 |