
Senior Platform Architect
Posted Sep 10

Posted Sep 10
This is a fully remote position, open to applicants in Brazil, +4 more countries.
β’ Develop and manage scalable backend and AI infrastructure for both real-time and batch processing workloads.
β’ Create and sustain deployment workflows that incorporate versioning, staged rollouts, automated releases, monitoring, and rollback strategies.
β’ Construct and maintain production-level LLM and agentic systems, integrating model providers, APIs, gateways, tools, and external services.
β’ Design reusable services, APIs, automation, and data pipelines to support AI-powered products and internal platform functionalities.
β’ Enhance infrastructure-as-code by utilizing Terraform and reusable provisioning patterns.
β’ Oversee GitOps deployment workflows with tools such as ArgoCD.
β’ Execute distributed workloads on Kubernetes (GKE), managing scaling, workload placement, tenant isolation, service reliability, and infrastructure capacity.
β’ Enhance observability and reliability through metrics, logging, tracing, SLOs, alerting, incident response, and operational tools.
β’ Identify opportunities for performance and infrastructure cost optimization across cloud services, compute, APIs, and AI workloads.
β’ Employ agentic coding and AI development tools for scaffolding, code generation and review, debugging, and workflow automation.
β’ Influence the platform's evolution to support increased scale, new architectural domains, and an expanding ML/AI presence.
β’ A minimum of 5 years in platform engineering, SRE, or infrastructure, including significant experience with operating production systems at scale.
β’ Robust foundation in SRE/DevOps, including ownership of production reliability, SLOs, post-mortems, and quantifiable improvements.
β’ Extensive expertise in Terraform, including complex state management, reusable modules, multi-project configurations, and CI-driven plan/apply workflows.
β’ Strong background in GitOps with experience using ArgoCD or Flux in a production environment.
β’ In-depth knowledge of Kubernetes, including production clusters, failure modes, and control-plane-level troubleshooting.
β’ Comprehensive experience with cloud infrastructure across AWS, Azure, or GCP, covering networking, compute, IAM, storage, and multi-account or multi-project design.
β’ Practical experience in building and managing CI/CD pipelines using GitHub Actions, Cloud Build, GitLab CI, or similar tools.
β’ An automation-first approach at a senior level.
β’ Active engagement with agentic coding tools.
β’ Excellent written and verbal communication abilities.
β’ Bachelor's Degree or equivalent experience (application question).
β’ Degree in IT, Computer Science, or equivalent experience (application question).
β’ Availability to work within the EST time zone.
β’ Preferred: Experience with MLOps, GCP, BigQuery, Dataflow, Pub/Sub, Dataproc, GPU scheduling, LLM inference, ML pipelines, model registries, drift monitoring, FinOps, data infrastructure, multi-tenant infrastructure, and startup scaling.
β’ Compensation in USD.
β’ Opportunity for remote work within LATAM.
β’ Flexible location options throughout Argentina, Brazil, Colombia, Mexico, and Peru.
RR Donnelley
plotdesk
CmdScale GmbH
Colsubsidio
Get handpicked remote jobs straight to your inbox weekly.