Remotery

Senior Site Reliability Engineer

Posted Jul 23

This is a fully remote position, open to applicants in Brazil.

📋 Description

• Design, upgrade, architect, and construct scalable infrastructure solutions utilizing Kubernetes, AWS, RDS (MySQL/Postgres), and contemporary distributed patterns.

• Steer the infrastructure team's roadmap, guiding us towards enhanced reliability, recoverability, and scalability.

• Lead capacity planning, benchmarking, and collaborate with the team to conduct stress tests on our systems, identify bottlenecks, and prepare for future business growth.

• Define, uphold, and enforce SLAs and alerts throughout our infrastructure.

• Guide teams towards improved anomaly detection, developing more adaptable alerting mechanisms.

• Assist in spearheading Sezzle’s AI enablement initiatives by identifying opportunities to leverage AI and automation to improve infrastructure reliability, developer efficiency, and internal tools.

• Ensure consistency and scalability in a distributed microservices architecture while preserving performance and reliability.

• Establish and refine engineering best practices for observability, security, and CI/CD across various teams.

• Mentor engineers and promote a culture of learning, innovation, and operational excellence.

• Collaborate across functions to translate business objectives into technical roadmaps and deliver impactful results.


⛳️ Requirements

• Over 12 years of professional experience in software engineering or infrastructure engineering, with substantial SRE and backend expertise.

• Implemented significant modifications to a production application or infrastructure setup within the last 30 days.

• Strong command of Golang, coupled with experience in building and maintaining RESTful APIs.

• Proficiency with SQL-based RDBMS (MySQL, PostgreSQL) and experience enhancing schema and queries for performance at scale.

• Familiarity with observability tools such as Prometheus, Grafana, Datadog, and New Relic.

• Comprehensive understanding of distributed systems design patterns (e.g., transactional outbox, event-driven architecture, stream processing, and queues).

• Proven ability to introduce new ideas, influence decision-making, and lead complex technical projects.

• Bachelor’s degree in Computer Science or equivalent practical experience.


🏝️ Benefits

• Competitive salary and performance-based bonuses.

• Comprehensive health, dental, and vision insurance.

• Flexible work hours and remote work options.

• Professional development opportunities and support for continuous learning.

People also viewed

CWILL13 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3714 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT15 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group15 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo15 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch16 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers