Software Engineer

Cognite Cognite · Industrial · India · Engineering

Software Engineer role focused on building high-performance distributed data infrastructure for industrial digitalization and AI solutions, processing terabyte-scale industrial datasets and supporting AI agents.

What you'd actually do

  1. Build high-performance data pipelines using Spark, Flink, and Kafka to process terabyte-scale industrial datasets.
  2. Develop low-latency APIs and services supporting thousands of concurrent users with sub-second response times.
  3. Optimize time-series, sensor, and operational data storage and retrieval for massive scale.
  4. Engineer distributed processing solutions, including real-time streaming that handles millions of events per second.
  5. Design and evolve cost-efficient data lake architectures (S3/GCS) using modern formats like Parquet/ORC.

Skills

Required

  • Scala, Java, or Python
  • distributed data systems
  • backend engineering
  • platform engineering
  • cloud platforms (AWS/GCP/Azure)
  • data lake/object storage
  • Spark
  • Flink
  • Kafka
  • JVM performance
  • GC tuning
  • memory management
  • high-throughput REST/gRPC services
  • caching with Redis
  • monitoring and observability (Prometheus, Grafana, OpenTelemetry)

Nice to have

  • large-scale data
  • OLAP systems
  • industrial/IoT data
  • contributions to open-source
  • industrial data/AI platforms
  • ClickHouse
  • Pinot
  • Druid
  • Parquet/ORC
  • distributed tracing

What the JD emphasized

  • high-performance distributed systems
  • low-latency APIs
  • sub-second response times
  • millions of events per second
  • high-throughput REST/gRPC services
  • strong hands-on experience with Flink/Kafka
  • scale systems to 10K+ QPS

Other signals

  • industrial digitalization
  • AI agents
  • data solutions
  • AI-ready data platforms