
Senior Site Reliability Engineer
Posted 5 days ago

Posted 5 days ago
This is a fully remote position, open to applicants in California.
• Develop and manage scalable microservices, platform capabilities, and shared engineering libraries.
• Design, deploy, and sustain secure and reliable cloud infrastructure along with Kubernetes-based services.
• Enhance CI/CD pipelines, developer workflows, and automate software delivery across engineering teams.
• Engineer and maintain Circle’s autonomous coding-agent orchestration platform and AI-driven developer tools.
• Implement observability, monitoring, incident response, and operational best practices for production systems and AI services.
• Collaborate with Product and Engineering teams to create resilient architectures and enhance platform reliability.
• Continually improve security measures, access controls, auditability, and cost efficiency across cloud infrastructure and AI-driven workflows.
• Diagnose production issues and document operational standards.
• Assist teams in the safe adoption and operation of AI-enabled tools and workflows.
• 3–6 years of professional experience in software development.
• Proficiency in Golang, Java, JavaScript/TypeScript, Python, or Rust.
• Practical experience in building and managing cloud-native applications on AWS or GCP.
• Familiarity with Kubernetes-based infrastructure.
• Background in engineering and operating production workflow orchestration systems, developer platforms, or autonomous agent solutions.
• Experience in integrating AI capabilities via APIs, SDKs, or workflow tools.
• Knowledge of applying observability, access controls, auditability, evaluation, monitoring, and human approval operational guardrails.
• Strong grasp of distributed systems.
• Solid understanding of RESTful API design.
• Comprehensive knowledge of SQL database design, schema modeling, and query optimization.
• Commitment to writing clean, maintainable, and well-tested code.
• Bachelor’s degree in Computer Science or a related technical discipline, or equivalent practical experience.
• Ability to work collaboratively across distributed engineering teams and convey technical concepts effectively.
• Self-driven, with a growth mindset, eagerness to learn new technologies, and capability to excel in a fast-paced environment.
• Flexible work environment.
• Inclusive workplace culture.
• Equal opportunity employment.
• Interview accommodations or assistance for disabilities.
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.