
Platform Operations Engineer
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in Europe.
• Handle and manage requests through Telegram, Slack, and email
• Clarify the nature of issues and collect necessary details from users
• Determine the specific component or service associated with an incident
• Take ownership of incidents from detection all the way to resolution
• Keep stakeholders informed throughout the incident resolution process
• Conduct basic infrastructure troubleshooting involving logs, service status, and configurations
• Perform deterministic runbook actions, such as restarting failed services, testing recovery processes, and confirming resolutions
• Escalate issues to L2/L3 DevOps and developers with well-prepared context
• Create tickets and bug reports in the tracking system
• Collaborate with the team to develop and maintain runbooks and knowledge base entries for recurring issues
• Experience with monitoring and logging systems
• Proficiency in reading and analyzing logs using Grafana, Kibana, and Loki
• Basic skills in issue localization across network, DNS, and service connectivity
• Fundamental understanding of Kubernetes: kubectl logs, kubectl describe, kubectl get
• Knowledge of application configuration: Helm values, ConfigMaps, and environment variables
• Infrastructure-level troubleshooting expertise, including service availability, node status, and resource management
• Background in Quality Assurance or technical support
• Willingness to work night shifts to cover European hours
• Experience with Helm is advantageous
• Familiarity with CI/CD pipelines is a nice addition
• Understanding of microservices architecture is a plus
• Ability to ask specific clarifying questions
• Structured problem description when escalating issues
• Independence and ownership in work
• Strong desire to understand issues rather than simply pass them along
• Growth-oriented mindset
• Preference for candidates located near the EST timezone
• 21 vacation days plus public holidays and 5 sick days
• Private English lessons through Preply
• Fully remote work across Europe
• Access to a cutting-edge tech stack including Speech Technologies, NLP, Generative AI (LLMs, diffusion models), and voice-first agentic architecture with a focus on privacy and on-premises deployment
• Rapid career advancement opportunities
• Dynamic startup environment combined with enterprise-level stability, real clients, and tangible revenue
Veeam Software
Acosta
Bright Vision Technologies
Samsara
Get handpicked remote jobs straight to your inbox weekly.