
Senior Site Reliability Engineer, SRE
Posted Sep 10

Posted Sep 10
This is a fully remote position, open to applicants in Mexico.
• Oversee technical investigations and establish comprehensive application traceability across intricate systems.
• Identify and rectify network, proxy, load balancer, and TLS-related issues.
• Evaluate API latency, timeouts, retries, and failures in dependencies.
• Implement transaction tracing from the browser to the backend.
• Conduct Real User Monitoring and analyze frontend errors.
• Partner with global teams on significant automation and cloud initiatives.
• Extensive experience in Java application performance engineering.
• Proficient diagnostic skills in WebSphere and WebLogic.
• Expertise in JVM profiling, garbage collection analysis, heap dumps, and thread dumps.
• Implementation of structured logging at the application level and correlation ID.
• Ability to perform distributed tracing across applications, APIs, and their downstream dependencies.
• Practical experience with APM tools such as Datadog, New Relic, or similar.
• Proficient in Oracle performance diagnostics, including SQL execution plans, AWR, and ASH.
• Knowledge of JDBC connection pools, application thread pools, and session management.
• Experience in tracing transactions in legacy or partially instrumented systems.
• Strong resilience, emotional intelligence, and commitment to agile delivery.
• Familiarity with cloud-native principles or AI coding assistants.
• Advanced proficiency in Oral English.
• Advanced proficiency in Spanish.
• Flexible working arrangements.
• Opportunities for continuous learning and professional development.
• Engagement in high-impact projects and complex engineering challenges.
• Collaboration with global teams across 39 delivery centers.
• Involvement with AI-driven automation and cloud solutions.
• A culture of radical ownership and an agile delivery environment.
FourEnergy GmbH
ICF
Mastercam
C&S Informática
Get handpicked remote jobs straight to your inbox weekly.