Senior / Staff Infrastructure Engineer (on-premise)

Waabi Waabi · Robotics · Dallas, TX · Software Engineering

This role focuses on building and maintaining the physical infrastructure (servers, networks, compute) that supports ML efforts, CI systems, and software delivery for an autonomous transportation company. It involves collaboration with hardware and software teams on vehicle compute, network design, OS delivery, and data pipelines, with a strong emphasis on on-premise systems and performance tuning.

What you'd actually do

  1. Work alongside a team of multidisciplinary Engineers and Research Scientists using an AI-first approach to enable safe self-driving at scale.
  2. Collaborate with hardware team on vehicle compute and network designs and implementation.
  3. Work with software and hardware teams on OS and software delivery into vehicles.
  4. Assist with vehicle compute and network performance tuning and diagnostics as needed.
  5. Work with software and hardware teams to implement and maintain high volume data pipelines.

Skills

Required

  • BS, MS/PhD in Computer Science or similar technical field of study or equivalent practical experience.
  • 5+ years of relevant industry experience.
  • Deep understanding of computer and network hardware and how detailed selection affects performance and reliability.
  • Experience with on-premise servers, network equipment and storage systems.
  • Experience in Linux and Linux kernel (packaging, performance tuning, low level debugging, hardening).
  • Experience with configuration management (Ansible preferred) and system administration.
  • Familiarity with security concepts, access management, and data security.
  • Experience with containers and container orchestration (i.e., Docker, Kubernetes).
  • Experience with high performance network design and debugging.
  • Experience with CI/CD pipelines and release management.
  • Experience with Python and Bash.

Nice to have

  • Experience with rugged wireless technologies (ie. LTE/5G modems).
  • Experience with scale-out storage systems.
  • Experience working with public cloud platforms (AWS preferred).
  • Experience with infrastructure as code systems (Terraform preferred).
  • Experience with automated building of OS images for deployment onto physical computers.
  • Familiarity with self-driving vehicle sensors and hardware.
  • Some experience in reading and developing production quality software.
  • Familiarity with GO, Rust or C++ ecosystems.
  • Experience with building platform services that enable other teams to do their best work.
  • Experience in common ML tools, workflows and frameworks (i.e. systems like Kubeflow or MLFlow).

What the JD emphasized

  • physical computers and networks
  • physical infrastructure
  • vehicle compute and network designs and implementation
  • high volume data pipelines