Research Engineer, Conversational Agentic Ai, Deepmind

Google Google · Big Tech · Mountain View, CA +2

Research Engineer role focused on designing, developing, and deploying novel multimodal conversational agents, including audio-first models for complex dialogs with tool use and real-time capabilities, leveraging both real and synthetic data, and optimizing for low-latency streaming bi-directional dialog.

What you'd actually do

  1. Partner with the Gemini or DeepMind teams to design, develop, and deploy novel multimodal conversational agents.
  2. Develop audio-first models capable of orchestrating and planning complex dialogs, including leveraging external tools like search when necessary.
  3. Leverage new sources of data (real and synthetic) to empower new real-time dialog capabilities.
  4. Work with infra teams to design models suitable for streaming bi-directional dialog, so the user experience is always fluid and low-latency.
  5. Prototype and evaluate new technologies.

Skills

Required

  • data preparation
  • training ML models
  • evaluation of ML models
  • AI/ML-driven features
  • AI/ML-driven infrastructure
  • Large Language Models (LLMs)
  • NLP
  • data pipelines
  • Machine Learning
  • Artificial Intelligence
  • AI algorithms
  • data analysis

Nice to have

  • publications in conferences or journals (e.g., NeurIPS, ICML, ICLR, AAAI, CVPR)
  • Research background in NLP/Generative AI

What the JD emphasized

  • 8 years of experience in data preparation, training, and evaluation of ML models
  • Experience building or implementing AI/ML-driven features or infrastructure
  • Experience in Machine Learning, Artificial Intelligence, AI algorithms, and data analysis

Other signals

  • novel multimodal conversational agents
  • audio-first models capable of orchestrating and planning complex dialogs
  • leveraging external tools like search
  • real and synthetic data
  • streaming bi-directional dialog
  • low-latency