Senior Software Engineer -alerts

New Relic New Relic · Enterprise · OR · AIOps

Senior Software Engineer role focused on backend services for an observability platform's alerting system. The role involves designing, developing, and deploying Java/Kotlin services that process high-volume telemetry and alerting workloads, with a strong emphasis on reliability, customer impact, and working with distributed data systems at scale. While not directly building core AI models, the role supports AI agents by exposing telemetry data and involves integrating with LLM APIs and RAG workflows as a bonus.

What you'd actually do

  1. Design, develop, and deploy backend services in Java/Kotlin that process high-volume telemetry and alerting workloads, with reliability and customer impact top of mind
  2. Collaborate with product managers and engineers who specialize in high-throughput data streaming systems, computing infrastructure, design, UIs, and customer-facing APIs
  3. Implement exciting new Alerting features that affect our entire pipeline, and also help reduce tech debt and retire old architecture
  4. Advocate for architecture improvements, provide future direction, and clearly articulate reasons why while assessing tradeoffs
  5. Develop and deploy your code to customers multiple times per day

Skills

Required

  • 5+ years of professional backend software engineering experience
  • Strong proficiency in Java (Kotlin preferred)
  • Solid grasp of OOP principles
  • Experience with RESTful APIs
  • Experience with multi-threaded programming
  • Experience building multi-threaded Java services
  • Experience shipping reliable high-throughput services to customers in a production environment
  • Experience with relational databases: complex SQL, optimization, pagination, partitioning, and scaling
  • Experience working with distributed systems
  • Understanding of how to write code and queries that perform at scale
  • Experience delivering APIs consumed by internal and/or external customers
  • Demonstrated empathy for the end user
  • Experience working in an agile environment characterized by rapid change
  • Strong interpersonal skills
  • Ability to seek consensus
  • Ability to lead by example
  • Persistence and tenacity

Nice to have

  • Hands-on experience building with LLMs and AI agents
  • Designing prompts
  • Integrating LLM APIs
  • Building retrieval-augmented workflows
  • Evaluating model output quality
  • Developing/maintaining MCP (Model Context Protocol) servers
  • Familiarity with message queuing systems
  • Familiarity with streaming patterns like Kafka
  • Familiarity with Flink
  • Familiarity with Spark Streaming
  • Familiarity with AMQP (RabbitMQ)
  • Familiarity with gRPC
  • Familiarity with Kubernetes
  • Familiarity with Docker
  • Familiarity with Terraform
  • Cloud computing experience (AWS, GCP, or Azure)
  • Frontend awareness or working knowledge (React, TypeScript, GraphQL, CSS)

What the JD emphasized

  • backend services
  • high-volume telemetry
  • alerting workloads
  • customer-facing product
  • distributed data systems
  • high-throughput pipelines
  • backend data logistics, persistence, and retrieval at scale directly affect what customers experience

Other signals

  • backend services
  • high-volume telemetry
  • alerting workloads
  • customer-facing product
  • distributed data systems
  • high-throughput pipelines