Remotery

Customer Reliability Engineer

Posted Jul 24

This is a fully remote position, open to applicants in Romania.

📋 Description

• Take full responsibility for the complete deployment of the product on Kubernetes, which includes rollout, rolling updates, rollback, scaling, and the pause/resume functionality of releases utilizing Kubernetes Deployment primitives that offer declarative updates for Pods and ReplicaSets.

• Ensure the health and reliability of the product through the use of monitoring, logs, traces, dashboards, and SLOs/SLIs along with error budgets that inform risk management and prioritization decisions.

• Deliver customer-facing technical operations — assist clients in managing on-premises deployments; provide technical support and health assurance for hosted deployments — serving as the technical lead for customer escalations.

• Maintain comprehensive knowledge of each customer's environment — including architecture, configurations, requirements, and customizations.

• Ensure configuration consistency across customers and environments by utilizing configuration management and a GitOps-style single-source-of-truth to minimize drift and enforce compliance automatically.

• Foster continuous operational enhancements by documenting resolutions, root causes, and best practices into a knowledge base, automation, and runbooks.

• Collaborate across functions with Engineering, Product, and Operations teams to ensure production readiness, plan releases, conduct user acceptance testing (UAT), and secure production go-live approval.


⛳️ Requirements

• Over 5 years of experience in infrastructure, DevOps, or Site Reliability Engineering roles, preferably in fintech, financial services, or other regulated sectors.

• Proven hands-on experience with Kubernetes and Helm in a production environment.

• Proficiency in infrastructure as code and configuration management tools such as Terraform, Ansible, or their equivalents.

• Experience in developing and maintaining CI/CD pipelines.

• Understanding of Linux and cloud-native security principles.

• Strong coding and scripting abilities with an emphasis on automation and efficiency.

• Experience in customer technical operations, including enterprise customer engagement and escalation management.

• Must possess legal authorization to work in the country where the position is located without requiring current or future visa sponsorship.

• **Nice to Have**

• Familiarity with modern software development practices is advantageous.

• Experience in a globally distributed start-up or high-growth environment.


🏝️ Benefits

• Meaningful Impact – Contribute significantly to the future of reliable cross-border payments and fraud prevention, addressing crucial challenges faced by financial institutions and businesses globally.

• Learn from Experienced Industry Leaders – Become part of a team with extensive knowledge in payments, fintech, and technology, while gaining exposure to international customers and markets.

• Ownership & Growth – Join a rapidly expanding global fintech where your contributions are recognized and valued, with the chance to participate in our Employee Stock Option Plan (ESOP) and share in our long-term success.

People also viewed

CWILL21 hours ago

DevOps/SRE Engineer, Bilingual Mandarin

US flagCalifornia, +4 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$100k – $130k/year
ApplyView job
a3722 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT22 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group22 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo23 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch23 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers