Senior Lead AI Engineer (gen AI Platform Services, Agentic Ai)

Capital One Capital One · Banking · San Jose, CA +4

Senior Lead AI Engineer role focused on building and deploying AI-powered products and foundational AI systems, including LLM inference, agentic AI, similarity search, guardrails, and model evaluation. The role involves optimizing performance, scalability, cost, and latency of large-scale production AI systems.

What you'd actually do

  1. Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One.
  2. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
  3. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
  4. Invent and introduce state-of-the-art LLM optimization techniques to improve the performance — scalability, cost, latency, throughput — of large scale production AI systems.
  5. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.

Skills

Required

  • Python
  • Go
  • Scala
  • Java

Nice to have

  • AWS
  • Google Cloud
  • Azure
  • Huggingface
  • VectorDBs
  • Nemo Guardrails
  • PyTorch
  • C++
  • C#
  • LLM Inference
  • Similarity Search
  • Guardrails
  • Memory
  • optimization techniques
  • training software
  • inference software

What the JD emphasized

  • responsible and reliable AI systems
  • responsible and scalable ways
  • responsible and scalable AI solutions

Other signals

  • building and deploying proprietary solutions
  • empower teams across Capital One to enhance their products with the transformative power of AI
  • design, develop, test, deploy, and support AI software components
  • invent and introduce state-of-the-art LLM optimization techniques