
Mid Cloud Observability Engineer
Posted Jul 3

Posted Jul 3
This is a fully remote position, open to applicants in Brazil.
• A cloud observability engineer focuses on transforming complex systems into comprehensible formats, enhancing signal quality, and facilitating quicker, more intelligent debugging across teams.
• Monitor system health and assess alerts/incidents.
• Triage alerts effectively.
• Conduct thorough investigations of issues.
• Enhance observability instrumentation.
• Create and refine dashboards.
• Optimize alert processes.
• Collaborate with development teams and other engineering partners.
• Pursue continuous improvement initiatives.
• Provide release support.
• Design and establish observability frameworks utilizing metrics, logs, and distributed tracing.
• Develop dashboards, alerts, and visualizations for system health monitoring.
• Standardize observability practices across engineering teams (logging, telemetry, tracing).
• Implement and manage native monitoring tools.
• Create alerting systems to minimize alert fatigue.
• Participate in on-call rotations.
• Develop intelligent alerting to enhance system reliability and proactively mitigate risks.
• Identify reliability risks to fortify systems against failures.
• Achieve a reduction in alert noise and false positives.
• Increase observability coverage (% of services instrumented).
• Enhance SLO compliance.
• Integrate observability into applications.
• Incorporate tracing and metrics into the code.
• Standardize logging formats.
• Ensure all services are observable from end to end.
• Experience in SRE, DevOps, or Cloud Engineering.
• Proficiency in cloud platforms, specifically AWS.
• Familiarity with AWS services such as CloudWatch, X-Ray, Lambda, ECS/EKS, API Gateway, RDS, DynamoDB, and S3.
• Hands-on experience with observability tools like Dynatrace, Splunk, Datadog, OpenTelemetry, Prometheus, Grafana, and AWS Distro for OpenTelemetry.
• Strong understanding of containers and orchestration technologies (Docker, Kubernetes).
• Knowledge of CI/CD pipelines.
• Experience with Infrastructure as Code (Terraform, CloudFormation).
• Ideally, experience in building observability platforms at scale within AWS.
• Familiarity with multi-account AWS environments.
• Experience in optimizing costs for observability (logging/metrics ingestion).
• Background in high-scale distributed systems.
• Health insurance.
• Opportunities for professional development.
Get handpicked remote jobs straight to your inbox weekly.