Senior Platform Engineer (kubernetes, Application Infrastructure)

Whoop Whoop · Consumer · Boston, MA · Software

Senior Platform Engineer responsible for designing, developing, and operating Kubernetes clusters on AWS, focusing on scalability, resilience, security, and developer productivity. The role involves building systems to improve deployment safety and release velocity, advancing CI/CD capabilities, and mentoring other engineers.

What you'd actually do

  1. Design, develop, and operate WHOOP’s Kubernetes clusters running on AWS infrastructure
  2. Drive architectural decisions to improve scalability, resiliency, performance, and security across the build and deployment platform
  3. Build systems and tooling that increase deployment safety and accelerate release velocity to Kubernetes
  4. Advance CI/CD capabilities to support frequent, reliable production deployments
  5. Lead developer productivity improvements through tooling, automation, and platform integrations

Skills

Required

  • 5+ years of experience in DevOps, Platform, Site Reliability, CloudEngineering, or Backend Software Engineering roles
  • Deep understanding of Kubernetes architecture and core components
  • Strong knowledge of container networking concepts, including overlay networking, service meshes, and network policies
  • Experience with multi-cluster Kubernetes environments and inter-cluster communication patterns
  • Hands-on experience operating cloud infrastructure, preferably in AWS (e.g., IAM, VPC, EC2, S3, RDS, CloudTrail, Organizations)
  • Hands-on experience with Infrastructure as Code tools (e.g. Terraform)
  • Experience developing backend or infrastructure-adjacent services using Java, C#, or Python
  • Proven ability to evaluate system performance, identify bottlenecks, and use data to drive improvements
  • Experience collaborating with multiple stakeholders and prioritizing work for maximum business impact

Nice to have

  • Experience operating Kafka or other large-scale distributed systems
  • Experience with Kubernetes security best practices, including RBAC, secrets management, and pod security standards
  • Exposure to service reliability practices such as SLOs, SLIs, and error budgets
  • Prior experience supporting compliance or security-focused infrastructure initiatives

What the JD emphasized

  • security
  • scalability
  • resiliency
  • performance
  • secure-by-default infrastructure practices
  • Kubernetes
  • AWS
  • CI/CD
  • deployment safety
  • release velocity