Remotery

Lead Site Reliability Engineer – Cloud

Posted Aug 2

This is a fully remote position, open to applicants in France.

📋 Description

• Ensure the reliability, availability, and resilience of production systems.

• Predict potential failures and design efficient incident response procedures.

• Industrialize and automate platform operations for enhanced efficiency.

• Maintain a high standard of service quality for our clients while adhering to contractual obligations (SLAs).

• Analyze performance metrics, identify bottlenecks, and suggest enhancements to optimize resource utilization and scalability.

• Define, implement, and enhance observability tools (monitoring, metrics, logs, alerting) with a proactive mindset.

• Provide level-3 customer support in collaboration with support teams, in accordance with SLAs.

• Lead and facilitate incident retrospectives (post-mortems), identify root causes, and establish sustainable corrective measures.


⛳️ Requirements

• Strong expertise in cloud environments and distributed infrastructure.

• Proficiency in observability practices (logs, metrics, alerting) with a structured approach to diagnosing complex incidents.

• Solid understanding of containerized environments and their operational challenges.

• Proven experience with production databases: reliability, backups, restorations, replication, and scaling.

• Hands-on experience with Infrastructure as Code and automation of environments.

• Awareness of operational security considerations.

• Comfortable utilizing AI tools to enhance daily efficiency.

• Ability to operate effectively in complex, changing, or uncertain environments with rigor and reliability.

• Capable of prioritizing tasks, even in incident situations.

• Clear and structured communication style, with a preference for cross-functional collaboration and knowledge sharing.

• A blameless mindset, technical curiosity, composure, and a focus on user impact.

• Ability to provide technical leadership, mentor others, and promote collective practices.


🏝️ Benefits

• Fully remote position with one trip per quarter (Strasbourg or another city).

• Company events: one annual offsite and regular afterworks/social gatherings.

• Telework allowance (€57.60).

• Meal vouchers (Ticket Restaurant) (€11.52 per voucher) and Swile card with associated benefits.

• Flexible working hours under a fixed-hours agreement (including RTT).

• Linux laptop provided.

• Budget contribution for additional equipment.

People also viewed

DATAGROUP1 day ago

DevOps Engineer

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Ambush1 day ago

DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
DuoKey1 day ago

DevOps Engineer

MU flagMauritius OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
TEKsystems2 days ago

SRE – CloudOps, Practice Architect II

US flagIllinois OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
TEKsystems2 days ago

SRE CloudOps Practice Architect II

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$148.2k – $222.4k/year
ApplyView job
Level Data2 days ago

Senior DevOps Engineer

US flagMassachusetts OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$120k – $135k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers