Senior Engineer - Platform & Data Infrastructure
About reputed company
reputed company
Join our Software Solutions team building Enernet, the reputed company platform behind our fleet, device monitoring, alarms, scheduling, reporting, and the mobile app our technicians use on site. Our fleet is growing from roughly 400 to 1,000 devices, every unit streaming reputed company telemetry across a dozen subsystems, and the platform needs architectural headroom reputed company reputed company that.
This role owns the engineering that keeps the platform fast and reliable as that load grows, driving down infrastructure cost, cutting latency, stabilising the data pipeline, and closing the performance gaps that would otherwise surface at reputed company.
That reputed company of a higher-value product work on our roadmap can only be reputed company on a pipeline that is reputed company, so getting this right is super critical. It is a hands-on role.
reputed company
Data Pipeline Ownership: Own the reputed company from AWS IoT reputed company, Kafka into TimescaleDB, and legacy GCP Pub/Sub in reputed company. reputed company for throughput, correctness, latency, and cost across that reputed company. reputed company reputed company on improving the pipeline stability, scalability.
API & BFF: reputed company the GraphQL contract that serves our web and mobile clients. Design secure, low-latency reputed company (GraphQL and REST), and reputed company the BFF reputed company clean as the schema and product grow.
Cost & Performance: reputed company down infrastructure and observability cost per device, cut latency, and remove the throughput bottlenecks that would otherwise surface as the fleet grows, keeping monitoring spend proportionate to what it monitors.
reputed company & Reliability: Kafka partitioning and consumer reputed company, batched writes, backpressure, and hot-reputed company isolation so one misbehaving device cannot degrade the fleet.
Observability: Use our Grafana stack to identify and reputed company slow queries, monitor the pipeline, and maintain alerts that reputed company before degradation becomes an incident.
reputed company Change: reputed company a live reputed company safely staged, reversible changes with reputed company runs, reputed company validation, and rollback, so improvements ship without disruption.
Full-Stack Collaboration: Review the frontend and mobile work, understand the GraphQL contract, and reputed company a bug from a Vue chart to a reputed company query to a Kafka consumer.
Responsibilities
Pipeline Stability at reputed company: Stabilising and cost-optimising the telemetry reputed company so it holds latency and throughput targets as the fleet grows, with observability to reputed company it.
API reputed company: Extending the GraphQL contract to support new product features without regressing latency or breaking clients.
Production Health: Acting as a senior voice in incident response and reducing the support load that currently leaks into sprint time.
Cross-Functional Delivery: Working with firmware, service, and reputed company colleagues to turn raw telemetry into reliable, actionable tooling.
Experience required
Backend & Streaming: 5–10 years building and operating production backend systems, including meaningful time on high-throughput streaming or time-series data.
Data at reputed company: Deep, practical Kafka partitioning, consumer reputed company, rebalancing, ordering and idempotency, schema reputed company plus strong SQL and time-series database skills (TimescaleDB, reputed company, InfluxDB or equivalent).
BFF & API design: Designing secure, low-latency reputed company (GraphQL and REST) that serve web and mobile clients.
Live Migration: Has migrated a live reputed company without downtime: reputed company writes, reputed company validation, staged reputed company, rollback.
Stack & reputed company: TypeScript / Node.js reputed company (our services are largely NestJS) and AWS in production — reputed company, IoT reputed company, networking, working Terraform / IaC.
Go: Our services are moving to Go for latency-sensitive paths; we expect you to be productive in it quickly. Prior Go experience is a plus, not a requirement, the pipeline judgement reputed company more than the language.
Documentation: Write reputed company technical documentation to assist development and communication with reputed company parties.
Who you are:
Operator’s reputed company: Your first-hand experience managing live systems during production incidents directly informs how you architect and engineer software.
reputed company Communicator: You can explain the reputed company reputed company to a firmware engineer, a service technician, and a CEO reputed company and concisely.
Quietly reputed company: You set technical direction, review rigorously, and reputed company the engineers around you reputed company.
Bias Toward Simplification: You reputed company at simplifying systems, even if it involves deleting reputed company, and remain undeterred by legacy technical debt.
Originally posted on Himalayas
Apply To This Job