
Senior Engineer – Platform & Data Infrastructure
Posted Aug 5

Posted Aug 5
This is a fully remote position, open to applicants in India.
• Take ownership of the telemetry data flow from AWS IoT Core and Kafka into TimescaleDB, including legacy GCP Pub/Sub, ensuring accountability for factors such as throughput, accuracy, latency, cost, stability, and scalability.
• Advance the GraphQL contract and design secure, low-latency GraphQL and REST APIs tailored for web and mobile applications.
• Lower infrastructure and observability costs per device, reduce latency, and eliminate throughput bottlenecks as the fleet expands.
• Implement Kafka partitioning, consumer groups, batched writes, backpressure strategies, and hot-path isolation techniques.
• Utilize Grafana to identify slow queries, monitor the pipeline, and maintain proactive alerting mechanisms.
• Safely evolve the live system through staged reversible changes, parallel runs, shadow validation, and rollback procedures.
• Review frontend and mobile implementations, tracing issues across Vue charts, Timescale queries, and Kafka consumers.
• Stabilize and optimize the telemetry pathway for cost while adhering to latency and throughput targets.
• Extend APIs for new product features without increasing latency or disrupting existing clients.
• Serve as a senior voice during incident response efforts to alleviate support load.
• Collaborate with firmware, service, and commercial teams to transform telemetry into reliable tools.
• 5–10 years of experience in building and operating production backend systems, with significant experience in high-throughput streaming or time-series data.
• In-depth practical knowledge of Kafka partitioning, consumer groups, rebalancing, ordering, idempotency, and schema evolution.
• Strong skills in SQL and time-series databases, including TimescaleDB, ClickHouse, InfluxDB, or similar.
• Proven experience in designing secure, low-latency GraphQL and REST APIs for web and mobile platforms.
• Experience in migrating live systems without downtime using parallel writes, shadow validation, staged cutover, and rollback strategies.
• Proficient in TypeScript and Node.js, with experience in NestJS services.
• Production-level experience with AWS, including ECS, IoT Core, networking, and Terraform/IaC.
• Capability to quickly become productive in Go; prior experience with Go is advantageous but not mandatory.
• Ability to produce clear technical documentation.
• Direct experience managing live systems during production incidents.
• Skill in articulating systems clearly and concisely to both technical and non-technical stakeholders.
• Competence in setting technical direction, conducting thorough reviews, and enhancing engineering practices.
• Ability to streamline systems and navigate legacy technical debt.
• Competitive benefits package.
• Strong opportunities for professional development.
• Collaborative, multinational work environment.
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.