Research Scientist, AI Safety and Security

Google Google · Big Tech · Singapore

Research Scientist focused on AI Safety and Security, developing techniques for model robustness, interpretability, and adversarial defense. The role involves creating evaluation benchmarks and collaborating on transitioning research into production solutions.

What you'd actually do

  1. Drive foundational machine learning research in model robustness, continual learning, interpretability, and multiobjective optimization to advance trustworthy AI.
  2. Design and develop rigorous evaluation protocols, scenario-based benchmarks, and stress-testing methodologies to assess frontier AI capabilities and multi-agent consensus.
  3. Curate advanced datasets and conduct fine-tuning or optimization experiments to enhance model resilience against emerging threats and ensure adherence to safety constraints.
  4. Collaborate extensively with regional engineering hubs, core product teams, and academic partners to transition theoretical proofs-of-concept into robust production solutions.
  5. Publish groundbreaking research in machine learning venues and actively participate in academic and industry research communities

Skills

Required

  • Machine Learning
  • Adversarial Machine Learning
  • Evaluating Frontier AI Systems
  • Supervised Learning
  • Unsupervised Learning
  • Reinforcement Learning
  • ML Interpretability
  • Adversarial Robustness
  • ML Safety
  • Generative Models
  • Agentic AI
  • Multi-object Optimization
  • Scientific Publication Submission

Nice to have

  • General purpose programming languages (e.g., Python)
  • Investigating emerging technical threats (e.g., automated scams, deepfake generation, or rogue agent vulnerabilities)
  • Designing robust, proactive defense mechanisms
  • AI agent security
  • Data poisoning
  • Prompt injection
  • Model backdoor detection
  • Applying a security mindset to artificial intelligence
  • Debugging complex ML failure modes
  • Reverse engineering model behaviors
  • Red-teaming frontier AI systems
  • First-authored publications in top machine learning, safety/security tracks in machine learning or AI conferences, or HCI conferences

What the JD emphasized

  • PhD degree in Computer Science, a related field, or equivalent practical experience.
  • Experience in machine learning, adversarial machine learning or evaluating frontier AI systems, which includes but not limited to supervised learning, unsupervised learning and reinforcement learning, ML interpretability, adversarial robustness, ML safety, generative models, agentic AI, multi-object optimization.
  • One of more scientific publication submission(s) for conferences, journals, or public repositories (such as CVPR, ICCV, NeurIPS, ICML, ICLR, etc.).
  • First-authored publications in top machine learning, safety/security tracks in machine learning or AI conferences, or HCI conferences.

Other signals

  • AI Safety
  • Adversarial ML
  • Interpretability
  • Evaluation Benchmarks