Staff Software Engineer, AI Engines 3p Tpu Inference

Google Google · Big Tech · Mountain View, CA +1

Staff Software Engineer focused on AI Engines 3P TPU Inference, developing and optimizing ML infrastructure, compilers, and runtimes for Google's AI models, including Gemini. The role involves migrating existing frameworks, ensuring high performance, and collaborating with research and product teams to deliver impactful ML software infrastructure.

What you'd actually do

  1. Exercise judgment to guide sustainable engineering choices for ML systems at scale.
  2. Innovate next directions for infrastructure over a 12-month time horizon given a rapidly changing technology landscape.
  3. Deliver impactful capability and optimization impact to Cloud and partner product areas.
  4. Migrate existing frameworks (TensorFlow, JAX, PyTorch) runtimes (TF Executor, TFRT, PJRT) and product areas custom workflows (AdBrain) onto ML Runtime, minimizing any user disruption.
  5. Partner with GDM to transfer key innovations into products developed by your team and partner teams.

Skills

Required

  • software development
  • software design and architecture
  • ML infrastructure
  • ML design
  • model deployment
  • model evaluation
  • data processing
  • debugging
  • fine tuning
  • Speech/audio
  • reinforcement learning
  • ML compilers
  • runtimes

Nice to have

  • Master’s degree or PhD
  • data structures and algorithms
  • technical leadership
  • complex, matrixed organization
  • TPUs
  • TPU system design
  • GPUs

What the JD emphasized

  • ML compilers and runtimes
  • TPU inference
  • ML infrastructure
  • ML design and ML infrastructure

Other signals

  • ML infrastructure
  • TPU inference
  • ML compilers
  • runtime optimization