Opensource Al Workload Software Engineer

AMD AMD · Semiconductors · San Jose, CA · Engineering

Software Engineer to optimize AI workloads (LLM frameworks like vLLM, SGLang) on AMD GPUs, focusing on kernel-level performance analysis and integration with AMD software stacks (ROCm, ATen) and open-source frameworks (PyTorch, JAX, Triton).

What you'd actually do

  1. Optimize AMD GPU performance for open-source LLM frameworks such as vLLM and SGLang.
  2. Analyze and resolve performance bottlenecks in AI workloads at the kernel level.
  3. Integrate AMD software stacks (ROCm, ATen) into open-source frameworks including PyTorch, JAX, and Triton.
  4. Collaborate cross-functionally with software and hardware engineering teams to identify gaps and deliver performance improvements.
  5. Build strong technical relationships with internal teams and external partners.

Skills

Required

  • Optimize AMD GPU performance for open-source LLM frameworks such as vLLM and SGLang.
  • Analyze and resolve performance bottlenecks in AI workloads at the kernel level.
  • Integrate AMD software stacks (ROCm, ATen) into open-source frameworks including PyTorch, JAX, and Triton.

Nice to have

  • Deep experience in AI infrastructure and open-source ecosystems (e.g., vLLM, SGLang, JAX, XLA, PyTorch, Triton).
  • Strong kernel optimization expertise using DSLs and HIP (or CUDA), including familiarity with PTX/SASS or equivalent.
  • Solid understanding of modern GPU architectures.
  • Proven track record of open-source contributions (e.g., GitHub).
  • Strong leadership presence with excellent communication and interpersonal skills.

What the JD emphasized

  • state-of-the-art LLM inference and training
  • kernel level
  • open-source LLM frameworks
  • kernel optimization expertise

Other signals

  • Optimize AI workloads on AMD GPUs
  • Founding member of a highly skilled team
  • Drive performance at scale