Senior Site Reliability Engineer

Fivetran Fivetran · Data AI · Oakland, CA · Engineering Department

Fivetran is seeking a Senior Site Reliability Engineer to ensure the performance, reliability, and scalability of their data platform infrastructure. This role involves monitoring, incident response, and collaborating with engineering teams to evolve systems and maintain high availability.

What you'd actually do

  1. Responsible for ongoing reliability and robustness of Fivetran’s production infrastructure by monitoring availability, capacity, and throughput.
  2. Evolve systems by adding reliability into our product roadmap
  3. Coordinate the re-prioritize or fix critical bugs for support or sales requirements as needed
  4. Make recommendations to production infrastructure by interfacing with engineering to ensure 100% availability
  5. Ensure scalable artifacts deployment to all environments by automation scripts
  6. Constantly monitor infrastructure vulnerabilities and remedy them by working with the security team

Skills

Required

  • 5+ years of experience working with SaaS products at scale
  • Working knowledge of managed Kubernetes (EKS, AKS and GKE)
  • Knowledge of Cloud Platforms and related tooling: AWS, Azure, GCP, Terraform, Ansible, Buildkite, Pulumi and ArgoCD
  • Experience in Python/Shell scripting
  • Experience with Linux operating systems internals and administration
  • Experience with cloud networking like VPNs, Privatelinks, and Private Service connect (GCP)
  • Experience with databases such as PostgreSQL

Nice to have

  • Java, GoLang Programming skills