
Senior DevOps Engineer – Contract
Posted 12 hours ago

Posted 12 hours ago
This is a fully remote position, open to applicants in India.
• Take ownership and continuously enhance cloud infrastructure and platform capabilities that support applications, data systems, and engineering teams.
• Identify and implement enhancements to reliability, scalability, security, performance, cost efficiency, developer experience, and operational effectiveness.
• Design, construct, and maintain secure and scalable cloud infrastructure utilizing infrastructure-as-code and cloud-native methodologies.
• Create automation, self-service capabilities, reusable infrastructure patterns, and platform tooling to reduce repetitive operational tasks.
• Enhance Kubernetes and container-based infrastructure for standardized and reliable deployment, scaling, recovery, and operational procedures.
• Take charge of and advance CI/CD and GitOps workflows.
• Collaborate with Software Engineers, Product Engineers, Analytics Engineers, and Technology teams to develop platform solutions.
• Enhance observability through monitoring, logging, tracing, alerting, dashboards, and actionable service health indicators.
• Lead technical investigations during complex production incidents and drive systemic enhancements.
• Improve resilience, disaster recovery, backup strategies, security controls, access management, and infrastructure risk management.
• Evaluate infrastructure and architecture for cost efficiency and minimize unnecessary expenditures.
• Utilize AI-assisted engineering in infrastructure development, troubleshooting, root-cause analysis, automation, documentation, and operations.
• Contribute to architectural decisions and long-term platform strategy development.
• Define and advocate for engineering standards, operational practices, reusable patterns, and platform capabilities.
• Mentor engineers and elevate standards for infrastructure ownership, automation, reliability, and operational excellence.
• Proven ability to independently identify infrastructure and operational issues, determine root causes, design solutions, and implement improvements yielding measurable outcomes.
• Established experience managing business-critical production infrastructure and complex platform initiatives with minimal supervision.
• Strong judgment in balancing reliability, security, engineering velocity, cost, complexity, and business requirements.
• Ability to operate effectively during production incidents and maintain ownership until resolution and follow-up.
• Demonstrated ability to utilize AI-assisted engineering tools and critically assess AI-generated infrastructure changes before implementation in production.
• Extensive experience in designing, operating, troubleshooting, and enhancing highly available production systems in cloud environments.
• Solid experience with Kubernetes, containerization, and modern cloud-native infrastructure.
• Proficient with Infrastructure as Code using Terraform, OpenTofu, or similar technologies.
• Strong understanding of CI/CD, GitOps, automated deployment practices, and contemporary software delivery workflows.
• Familiarity with monitoring, logging, tracing, alerting, and incident management practices.
• Comprehensive understanding of reliability engineering, failure modes, resilience, capacity, recovery, and operational risk.
• Experience with secure infrastructure patterns, access controls, secrets management, system hardening, and cloud security best practices.
• Strong Linux and systems troubleshooting capabilities across applications, containers, networks, infrastructure, and cloud services.
• Significant experience managing production workloads in AWS or a comparable cloud platform.
• Experience with GitHub Actions, ArgoCD, or similar platforms.
• Familiarity with Datadog or analogous observability tools.
• Strong scripting or programming skills in Python, Bash, or another relevant language.
• Basic understanding of networking, DNS, CDN/WAF technologies, load balancing, TLS, and application delivery architecture.
• Experience supporting MySQL, Redis, Redshift, or similar stateful systems.
• Excellent communication skills with both technical and non-technical stakeholders.
• Ability to work collaboratively with Software Engineering, Analytics Engineering, Product, Security, and other teams.
• Self-motivated, highly autonomous, and comfortable working in a remote-first environment.
• Ability to constructively challenge existing practices and influence technical direction.
• Proven ability to mentor engineers and raise engineering standards.
• Must be currently authorized to work in India at the time of hire and maintain authorization throughout employment.
• Employer work visa sponsorship and support are not available.
• Pre-employment screening is required.
• No perks or benefits are offered with this contract role.
• Remote work.
• Opportunities for global team collaboration through company all-hands, team events, Slack communities, learning sessions, and in-person gatherings.
OnePay
Arista Networks
Octus
Tandem Diabetes Care
Get handpicked remote jobs straight to your inbox weekly.