
Data Engineer
Posted Sep 17

Posted Sep 17
This is a fully remote position, open to applicants in Mexico.
• Create a scalable data platform that integrates various sources for streamlined access.
• Design and enhance data tools for orchestration, governance, Data-Lakehouse, BI, and other related functions.
• Ensure the seamless operation of data systems for analysts, scientists, and engineers.
• Optimize data pipelines for ingestion, processing, and output within a microservices architecture.
• Construct, maintain, and monitor ETL/ELT processes, orchestrating workflows using Temporal.
• Diagnose and enhance the performance, scalability, and reliability of data infrastructure, including S3, Apache Iceberg, and ClickHouse.
• Collaborate across functions with data scientists, analysts, and backend engineers to grasp data requirements and deliver effective solutions.
• Implement and advocate for data quality, governance, and security best practices throughout the platform.
• Over 3 years of experience as a Data Engineer or in a comparable data infrastructure role.
• Strong SQL proficiency and practical experience with data modeling.
• Familiarity with data lake/lakehouse architectures (e.g., Apache Iceberg, S3, or similar).
• Experience with analytical/columnar databases (e.g., ClickHouse or similar).
• Proven experience in building and orchestrating ETL/ELT pipelines (e.g., Temporal, Airflow, or similar).
• Strong programming capabilities in Python and/or Scala/Java.
• Experience working within a microservices architecture and cloud environments (AWS preferred).
• Self-motivated with strong multitasking abilities and a proven team player.
• Excellent communication skills and the ability to operate both independently and collaboratively.
• Practical experience with Apache Spark (or similar technologies) for large-scale data processing.
• Professional proficiency in both written and spoken English.
• This role is focused on batch data processing rather than real-time streaming.
• Nice to have: Experience contributing to open-source data platforms and tools.
• Nice to have: Familiarity with BI and visualization tools (e.g., Superset, Looker, Tableau, Metabase, or similar).
• Nice to have: Experience with containerization and orchestration (Docker, Kubernetes).
• Nice to have: Experience with infrastructure-as-code and CI/CD practices.
• Nice to have: Experience with AWS EMR and running Apache Spark workloads in a cloud environment.
• Nice to have: Experience utilizing AI-assisted development tools (e.g., GitHub Copilot, Cursor, or similar).
• Remote work.
• Full-time employment.
INTERSPORT Deutschland eG
INTERSPORT Deutschland eG
INTERSPORT Deutschland eG
INTERSPORT Deutschland eG
Get handpicked remote jobs straight to your inbox weekly.