Senior Devops Sre Engineer

Axon Axon · Enterprise · Tel-Aviv Yafo, Israel · 2041 Carbyne Development

Senior DevOps SRE Engineer responsible for owning and evolving AWS infrastructure using Infrastructure-as-Code (Terraform/Terragrunt), architecting and scaling AWS environments, deploying and managing containerized workloads with Kubernetes and Docker, leading deployment and release processes, defining and enforcing SLOs/SLIs/error budgets, driving toil reduction, and leading observability efforts with Datadog. The role also involves building self-service internal developer platforms, taking end-to-end ownership of infrastructure projects, partnering with engineering teams on technical planning, and incorporating AI-assisted engineering practices into daily workflows. Experience with AI/ML infrastructure is a plus.

What you'd actually do

  1. Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
  2. Architect and scale AWS environments
  3. Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
  4. Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
  5. Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform

Skills

Required

  • DevOps/SRE experience
  • AWS
  • Kubernetes
  • Terraform/Terragrunt
  • Bash scripting
  • SRE principles (SLOs, SLIs, error budgets, toil reduction, blameless post-mortems)
  • Incident management
  • APIs
  • Microservices
  • Distributed systems
  • Project leadership
  • Communication

Nice to have

  • AI-assisted engineering tools (e.g., Claude, Cursor)
  • MCP-style integrations
  • Building AI/ML infrastructure (model deployment, inference pipelines)

What the JD emphasized

  • At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
  • Hands-on, production-level AWS experience.
  • Hands-on production experience with Kubernetes and containerization
  • Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
  • Strong incident management / on-call experience