
Cloud Infrastructure Engineer
Posted 12 hours ago

Posted 12 hours ago
This is a fully remote position, open to applicants in Hungary.
• Designing, constructing, and managing AWS infrastructure utilizing Terraform (EC2, RDS, S3, SQS, Lambda, ALB, ElastiCache, Route 53, VPC networking)
• Developing and maintaining Puppet modules to configure and oversee fleets of EC2 instances across various auto-scaling groups
• Sustaining and enhancing Python-based automation and tools that facilitate platform operations
• Operating and refining distributed service discovery and configuration management (etcd)
• Overseeing and optimizing a multi-tier caching strategy (Varnish, Redis/Valkey, PHP OPcache)
• Managing and scaling our observability stack (Prometheus, Grafana, Loki, Fluentd, PagerDuty) and taking part in on-call rotations
• Assessing and implementing distributed storage solutions as the platform progresses
• Enhancing deployment workflows and release procedures
• Collaborating with internal teams on API agreements, integration patterns, and operational tools
• Engaging in incident response, root cause analysis, and improvements to platform reliability
• Extensive experience with AWS services in production — particularly EC2, RDS, S3, SQS, Lambda, ALB, ElastiCache, Route 53, IAM, and VPC networking
• Expertise in creating and maintaining Terraform modules for production infrastructure
• Proficiency in authoring and maintaining Puppet modules (or similar agent-based configuration management) for fleet oversight
• Strong Python programming skills — you will be writing and maintaining production daemons, not just scripts
• Comprehensive Linux systems knowledge (Ubuntu) — comfortable with Apache/Nginx, PHP-FPM, Varnish, systemd, filesystem mounts, and networking fundamentals
• Understanding of distributed systems concepts: consensus, leader election, distributed locking, eventual consistency, and the associated tradeoffs
• Skilled in constructing and maintaining observability pipelines (Prometheus, Grafana, Loki, or similar) in a production environment
• Comfortable operating within a GitLab-based CI/CD workflow
• Effective communicator who can document architectural choices and articulate technical trade-offs to both technical and non-technical stakeholders
• Practical experience with distributed storage systems such as Ceph, GlusterFS, JuiceFS, CubeFS, or AWS EFS — particularly related to migration or assessment
• Familiarity with etcd (or comparable distributed key-value stores like Consul or ZooKeeper), including watch APIs, TTL-based locking, and cluster management
• Experience with Varnish and VCL, particularly in dynamic backend routing or multi-tenant configurations
• Working knowledge of PHP — not for application development, but to understand and maintain integration scripts that connect infrastructure and application layers
• Background in multi-tenant SaaS platform design — especially database-per-tenant models on shared infrastructure
• Familiarity with Moodle LMS or educational technology platforms
• Experience with secrets management solutions (AWS Secrets Manager, HashiCorp Vault, Parameter Store) and automated credential rotation
• Experience in designing zero-downtime deployment strategies for VM-based (non-containerized) environments.
• Open LMS is an equal employment opportunity/affirmative action employer and considers qualified applicants for employment without regard to race, gender, age, color, religion, national origin, marital status, disability, sexual orientation, or any other protected factor.
Zencoder
Autodesk
NVIDIA
Astreya
Get handpicked remote jobs straight to your inbox weekly.