Cloud / DevOps Engineer, Infra & IaC

Posted 1 day ago

This is a fully remote position, open to applicants in United States.

📋 Description

• Collaborate with research and engineering teams to pinpoint knowledge gaps and enhance AI model performance within cloud infrastructure, DevOps, Kubernetes, and Infrastructure-as-Code domains.

• Create realistic and technically challenging assignments that encompass Kubernetes troubleshooting, AWS service integration, infrastructure automation, and production operations.

• Develop precise and comprehensive reference solutions for intricate infrastructure engineering scenarios.

• Assess and analyze AI-generated technical solutions for accuracy, reliability, scalability, security, and compliance with production best practices.

• Offer clear, structured written feedback that highlights technical deficiencies, erroneous assumptions, and areas for improvement.

• Formulate detailed evaluation criteria, rubrics, and benchmarks for assessing Kubernetes troubleshooting, IaC architecture, AWS integrations, and CI/CD reasoning.

• Design scenarios that involve cluster failures, infrastructure automation, deployment workflows, service integrations, and operational reliability.

• Collaborate closely with other technical subject matter experts to ensure consistency, accuracy, and quality across evaluation datasets.

• Convert practical production experience into structured guidance to enhance AI-generated infrastructure solutions.

• Set high-quality standards for AI systems dealing with complex Cloud, DevOps, Kubernetes, AWS, and IaC challenges.

• Transform practical engineering insights into structured tasks, reference solutions, evaluation frameworks, and technical feedback to advance next-generation AI models.


⛳️ Requirements

• 4+ years of professional experience in Cloud Infrastructure, DevOps, Site Reliability Engineering, Platform Engineering, or a closely related domain.

• Strong hands-on experience managing Kubernetes in production settings, particularly in diagnosing, troubleshooting, and resolving cluster failures and operational challenges.

• Experience with Kubernetes that extends beyond merely writing manifests or utilizing managed Kubernetes control planes.

• Demonstrated production experience with Infrastructure-as-Code, specifically with Terraform and/or AWS CDK.

• Strong practical understanding of AWS cloud services, including production integration with AWS Lambda, API Gateway, and DynamoDB.

• Experience in designing, implementing, and maintaining CI/CD pipelines for production workloads.

• Comprehensive understanding of cloud architecture, infrastructure automation, deployment strategies, observability, reliability, and operational best practices.

• Demonstrated career growth with increased ownership and responsibility in infrastructure, DevOps, or platform engineering roles.

• Ability to reliably commit to 40 hours per week during standard business hours.

• Exceptional written and verbal communication skills, with the ability to elucidate complex technical concepts and engineering decisions clearly.

• Strong analytical and troubleshooting skills, particularly when addressing distributed systems and infrastructure failures.

• Experience working with large-scale cloud infrastructure or highly distributed systems.

• Familiarity with Kubernetes networking, security, storage, scaling, and cluster lifecycle management.

• Experience in implementing infrastructure security and reliability best practices.

• Knowledge of AWS architecture patterns and cloud-native application design.

• Experience with GitOps, containerization, monitoring, logging, and observability platforms.

• Familiarity with contemporary DevOps and platform engineering methodologies.

• Experience in reviewing or evaluating technical documentation, engineering solutions, or AI-generated outputs.


🏝️ Benefits

• Opportunities for professional development and growth.

• Collaborative and innovative work environment.

• Competitive salary and benefits package.

• Flexible working hours and remote work options.

People also viewed

TEKsystems2 hours ago

Cloud Deployment Engineer, Secret Clearance Required – 30% Travel

US flagDistrict of Columbia, +1 more stateFreelanceDevOps & Site Reliability Engineer (SRE)$80 – $110/hour
ApplyView job
Arctiq14 hours ago

Site Reliability Engineer – Vulnerability Remediation Consultant

US flagUnited States OnlyFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
GE Vernova15 hours ago

Senior Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$152.4k – $254k/year
ApplyView job
Ecosistemas15 hours ago

Senior DevOps Engineer, Bilingüe Inglés

AR flagArgentina OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Intetics1 day ago

Senior DevOps Engineer

PH flagPhilippines OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
OZmap1 day ago

Mid-Level SRE

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers