
Site Reliability Engineer
Posted Jul 15

Posted Jul 15
This is a fully remote position, open to applicants in California, +1 more state.
• At Offchain, we are not merely creating products; we are spearheading a movement.
• As leaders in blockchain scalability and security, we are at the cutting edge of redefining how the world engages with decentralized applications.
• We are establishing the groundwork that will shape the future of digital commerce, governance, and interpersonal interaction.
• We address the real-world challenges associated with scaling blockchain technology, ensuring adherence to its fundamental principles: decentralization, security, and transparency.
• Central to this vision is our team. We consist of innovative thinkers and proactive doers who welcome new challenges and pursue solutions that expand existing limits.
• If you are motivated by tackling unprecedented challenges and believe in the potential of decentralized systems to foster a more equitable digital future, we want to connect with you.
• Offchain is setting the standard for the entire Ethereum ecosystem.
• We developed the Arbitrum stack, which powers Arbitrum One, the most widely adopted Ethereum scaling solution available today.
• Over 100 diverse teams have leveraged Offchain technology to create their own Arbitrum chains.
• Prominent players in the industry, such as Robinhood, BlackRock, Ethena Labs, Securitize, Aave, and Apechain, are utilizing the Arbitrum stack.
• We are backed by $124 million in funding and have consistently demonstrated execution excellence with billions in secured value, thousands of supported projects, and infrastructure capable of processing millions of transactions seamlessly.
• Eager to explore blockchain technology, even if it is uncharted territory.
• Enjoy tackling infrastructure challenges in innovative ways and thinking beyond conventional approaches.
• Utilize tools like k9s or ArgoCD for efficiency and abstraction, while being comfortable delving into YAML, logs, or low-level debugging when issues arise.
• Experienced with GitOps-style systems, treating both infrastructure and application delivery as code.
• Have scaled deployment automation using patterns like ArgoCD ApplicationSets or comparable tools.
• Inquisitive about the inner workings of systems, not content with superficial fixes.
• Proficient in Linux, fluent in shell scripting, and productive in programming languages such as Python or Go.
• Comfortable working within a cloud platform (e.g., AWS, GCP, Azure), with a solid understanding of the underlying components to facilitate adaptation or migration across providers.
• Have participated in an on-call rotation, addressing incidents, troubleshooting under pressure, and conducting postmortems to enhance system reliability over time.
• Design systems with security as a priority, applying principles such as least privilege and threat modeling.
• Bring a robust technical foundation, exceptional problem-solving abilities, and a sincere commitment to high-quality work.
• Take initiative, collaborate transparently, and contribute to a culture of clarity, curiosity, and continuous growth.
• Have operated production Kubernetes clusters and constructed scalable, declarative infrastructure using Terraform or similar tools.
• Deployed and maintained Kubernetes environments, managed system components, and troubleshot applications running on the platform.
• Created CI/CD workflows with ArgoCD, GitHub Actions, CodeBuild, or similar tools, encompassing both infrastructure and application deployments.
• Designed and managed observability systems using time-series metrics, logs, and dashboards with tools like Prometheus, Loki, Mimir, Grafana, and CloudWatch.
• Diagnosed complex networking and storage challenges across intricate, distributed systems.
• Implemented secure-by-default infrastructure and contributed to architecture reviews and threat models.
• Automated operational workflows using scripting or programming in Python, Go, or Bash.
• SREs come from diverse backgrounds. If you possess strong problem-solving skills, curiosity, and a passion for building reliable systems, we would love to hear from you, even if your experience does not align perfectly with every bullet point.
• Remote-first global workforce with a New York office.
• Annual company offsite and team onsite events.
• Professional reimbursement program (supports attendance at industry conferences, certifications, and more).
• Comprehensive medical, dental, and vision coverage (available in the US and some other countries).
• 401k retirement plan with company match (US only).
• Wellness stipend.
• Home office setup and ergonomic equipment program.
The Codest
IRIUM
Sólides
Resilinc
Get handpicked remote jobs straight to your inbox weekly.