
Staff DevOps Engineer
Posted Jul 29

Posted Jul 29
This is a fully remote position, open to applicants in Mexico, +1 more state.
• Spearhead the design, implementation, and continuous enhancement of dependable, scalable, high-performing, and secure production platforms and services.
• Collaborate closely with cross-disciplinary teams to establish and uphold resilient infrastructure and deployment strategies.
• Offer technical guidance and mentorship to engineers throughout the organization, fostering robust engineering standards and operational best practices.
• Engage in a 24x7 on-call rotation to support essential services and guarantee platform availability.
• Promote standardization, automation, and documentation to enhance consistency, minimize operational overhead, and facilitate knowledge sharing.
• Contribute throughout the entire lifecycle of platform and service delivery, from design and construction to operation and optimization.
• Over 5 years of experience in DevOps, Site Reliability Engineering (SRE), platform engineering, or software engineering roles.
• Extensive experience with Kubernetes at scale, possessing a deep understanding of containers and container orchestration.
• Practical experience with infrastructure as code tools such as Terraform, Ansible, or Puppet.
• Strong programming expertise in at least one object-oriented language, coupled with effective scripting and automation skills.
• In-depth knowledge of security principles and best practices across infrastructure, platforms, and services.
• Substantial hands-on experience with at least one major cloud platform, with broad exposure to AWS, GCP, or OCI.
• Proficient in monitoring, alerting, and observability using tools like Prometheus, Grafana, or similar platforms.
• Solid grasp of networking fundamentals and distributed systems.
• Strong experience in Linux and/or Windows systems administration.
• Familiarity with software delivery automation, CI/CD pipelines, and secure Software Development Life Cycle (SDLC) practices, including exposure to static and dynamic security testing.
• Good understanding of SRE concepts such as SLIs, SLOs, SLAs, toil reduction, availability, and observability.
• Experience in managing and scaling Elasticsearch in production is highly preferred.
• Comprehensive health benefits package.
• Opportunities for professional development and continuous learning.
• Flexible work arrangements to promote work-life balance.
• Supportive and inclusive team culture.
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.