
Senior Platform Engineer
Posted 20 hours ago

Posted 20 hours ago
This is a fully remote position, open to applicants in United States, +1 more country.
β’ Develop, construct, and manage GitLab Orbit backend services, primarily utilizing Rust, within a cloud-native, distributed environment.
β’ Enhance deployment, monitoring, and operational processes across GitLab.com, Dedicated, and Self-Managed environments using technologies like Kubernetes, Helm, Terraform, and either AWS or GCP.
β’ Automate operational tasks and create tools for safer deployments, upgrades, recovery, capacity management, and service maintenance.
β’ Enhance observability through metrics, logs, traces, dashboards, alerts, and service-level indicators.
β’ Work in collaboration with SRE teams on incident response, on-call preparedness, runbooks, and troubleshooting workflows.
β’ Investigate production issues and address root causes related to concurrency, partial failures, retries, consistency, idempotency, performance, and multi-tenant isolation.
β’ Develop and optimize the graph query engine, SDLC and code indexing pipelines, cloud storage integrations, APIs, and MCP interfaces.
β’ Design dependable, scalable, and cost-effective data workflows utilizing S3, ClickHouse, NATS, and Siphon.
β’ Take ownership of changes from technical design through implementation and iteration, documenting constraints and trade-offs.
β’ Collaborate asynchronously with product, data, infrastructure, security, delivery, AI, and SRE teams.
β’ Proven experience in designing, building, and operating production backend services with strong Rust expertise or demonstrable ability to adapt to a Rust-first, performance-sensitive codebase.
β’ Background in distributed-system design, encompassing concurrency, failure management, consistency, messaging, data partitioning, scalability, and multi-tenant isolation.
β’ Practical knowledge of AWS, GCP, or both, covering cloud networking, identity and access management, compute, storage, and object storage like Amazon S3.
β’ Proficient in deploying and troubleshooting applications on Kubernetes using Helm.
β’ Experience in contributing to repeatable, reviewable infrastructure changes through Terraform or similar infrastructure-as-code tools.
β’ Proven track record of improving the reliability, observability, maintainability, and on-call preparedness of backend services.
β’ Skillful in diagnosing issues across application, data, orchestration, and infrastructure layers.
β’ Strong system design capabilities, including architectural decisions, documentation of constraints, and alignment of trade-offs.
β’ Ability to work independently in ambiguous situations by identifying challenges, driving solutions, and taking ownership.
β’ Capacity to learn and implement new languages, tools, and frameworks such as Ruby, Go, TypeScript, and Vue.
β’ Excellent written communication and asynchronous collaboration abilities.
β’ GitLab anticipates all team members to integrate AI into their daily workflows.
β’ Comprehensive benefits to support your health, finances, and overall well-being.
β’ Flexible Paid Time Off.
β’ Team Member Resource Groups.
β’ Equity Compensation & Employee Stock Purchase Plan.
β’ Growth and Development Fund.
β’ Parental Leave.
Supabase
Sysdig
Satellite Office
Bedrock Ocean Exploration
Get handpicked remote jobs straight to your inbox weekly.