Senior DevOps Engineer, Observability

Posted 9 hours ago

This is a fully remote position, open to applicants in Latin America.

📋 Description

• Design, construct, and maintain infrastructure systems that cater to NetBox Labs’ SaaS and On-Premise engineering requirements.

• Manage and enhance AWS and other cloud infrastructures with an emphasis on cost-effectiveness, security, and performance optimization.

• Contribute to the development of internal platform tools, CI/CD automation utilizing GitHub Actions, and capabilities for developer self-service.

• Improve observability and incident response systems, encompassing monitoring, alerting, and SLOs.

• Work collaboratively with product teams aligned with streams to grasp their requirements and continuously enhance internal platforms.

• Assist in enforcing and refining security and compliance standards, including SOC 2 controls.

• Contribute to the creation of documentation, onboarding resources, and internal support procedures.

• Take part in the on-call rotation.


⛳️ Requirements

• Over 5 years of experience in DevOps, SRE, or platform engineering positions.

• More than 2 years of experience in a B2B software startup environment.

• Extensive experience with AWS (EC2, VPC, IAM, RDS, etc.), particularly with EKS/Kubernetes and infrastructure-as-code tools such as Terraform and Helm.

• Familiarity with CI/CD pipelines and automation tools, preferably GitHub Actions.

• Knowledge of observability tools like Prometheus, Grafana, Mimir, Loki, or similar technologies.

• Proficiency in programming languages such as Python, Go, or shell scripting.

• Ability to thrive in a dynamic, fast-paced startup atmosphere.

• Excellent communication and documentation abilities.

• Experience with Change Data Capture (CDC) and event streaming systems, or experience in scaling large, multi-tenant observability systems that include ingestion, analysis, and alerting.

• Familiarity with messaging technologies such as MQTT, AMQP, and others.

• Understanding of the NetBox ecosystem or network automation tools.

• Experience with or contributions to open-source projects.

• Familiarity with AI tools like Copilot, ChatGPT, or Cursor.


🏝️ Benefits

• Equity offerings.

• Bonus opportunities.

• Commitment to being an equal opportunity employer.

• Provision for accommodations during the hiring process, if required.

People also viewed

Fairsource9 hours ago

DevOps, Kubernetes Consultant

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€110k – €140k/year
ApplyView job
VELZI.AI LIMITED9 hours ago

DevOps Engineer

ID flagIndonesia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
First Due9 hours ago

Platform Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$165k/year
ApplyView job
CoDev9 hours ago

Senior DevOps Engineer

PH flagPhilippines OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Fusable9 hours ago

Senior Dev Ops Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $155k/year
ApplyView job
NoaNet12 hours ago

Senior Full Stack DevOps Engineer

US flagIdaho, +2 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$130k – $170k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers