Data Center Operations Coordinator

Together AI Together AI · Data AI · San Francisco, CA · Engineering

This role manages and tracks break/fix activities across multiple data center locations, acting as a central point of coordination for hardware incidents, vendor dispatches, ticket management, asset tracking, and operational reporting to ensure maximum uptime and fast issue resolution. Responsibilities include monitoring ticket queues, coordinating with on-site technicians and vendors, maintaining hardware records, escalating outages, scheduling maintenance, providing status reports, and identifying trends in hardware failures.

What you'd actually do

  1. Track and manage all break/fix incidents across multiple data centers
  2. Monitor ticket queues and ensure SLA compliance for incident response and resolution
  3. Coordinate with on-site technicians, remote hands teams, vendors, and engineering groups
  4. Maintain accurate records of failed hardware, replacements, RMAs, and repair status
  5. Escalate critical outages and recurring infrastructure issues to leadership and engineering teams

Skills

Required

  • Experience working in data center operations, IT infrastructure, or hardware support
  • Strong understanding of server, storage, and networking hardware
  • Experience with ticketing systems such as ServiceNow, Jira, or Remedy
  • Ability to manage multiple priorities across several sites simultaneously
  • Excellent communication and organizational skills
  • Familiarity with SLA management and incident escalation processes
  • Proficiency with Excel, reporting dashboards, and inventory tracking tools

Nice to have

  • Experience supporting enterprise or hyperscale data centers
  • Knowledge of remote hands operations and vendor management
  • Understanding of ITIL processes and change management
  • CompTIA Server+, Network+, or similar certifications