
Principal Data Platform Engineer β Swedish Ad Platform
Posted Sep 9

Posted Sep 9
This is a fully remote position, open to applicants in Germany, +1 more country.
β’ Oversee the technical design and architecture of a contemporary enterprise-level data platform.
β’ Establish scalable data flows from Iceberg-based event storage into both analytical and operational data products.
β’ Assess and select technologies for orchestration, transformation, querying, serving, and storage solutions.
β’ Create reusable canonical entities and data models spanning advertising, campaigns, inventory, billing, and customer sectors.
β’ Develop scalable engineering patterns for batch processing and near-real-time operations.
β’ Construct dependable and observable data pipelines tailored for high-volume AdTech workloads.
β’ Implement capabilities for data quality, lineage, observability, and reconciliation.
β’ Collaborate with Platform Engineering teams to establish CI-enforced data contracts and governance standards.
β’ Define strategies for tenant isolation, access control, and regional data boundaries.
β’ Set standards for testing, deployment automation, schema evolution, and version control.
β’ Enhance platform performance, scalability, and infrastructure cost-effectiveness.
β’ Mentor engineers and promote engineering excellence within the team.
β’ Work closely with leadership and cross-functional stakeholders.
β’ Deliver the initial production-ready version of the platform.
β’ Create the foundational production data architecture, including documented architectural decisions, enforceable event contracts, verified canonical entities, and observable data quality and lineage.
β’ A minimum of 8 years of experience in designing and managing production-grade data platforms.
β’ Extensive expertise in distributed data systems and high-volume event-driven architectures.
β’ In-depth knowledge of modern lakehouse architectures and Apache Iceberg.
β’ Proficient SQL skills.
β’ Practical experience with Python, Java, Scala, or comparable programming languages.
β’ Familiarity with distributed processing and query technologies such as Spark, Flink, or Trino.
β’ Strong understanding of data modeling, partitioning strategies, and performance tuning.
β’ Proven track record in constructing batch and near-real-time data pipelines.
β’ Hands-on experience with AWS cloud infrastructure.
β’ Comprehensive understanding of CI/CD pipelines, Infrastructure as Code, and production observability.
β’ Experience in implementing data contracts, schema evolution, and data quality frameworks.
β’ Capability to make practical architectural decisions in greenfield environments.
β’ Upper-Intermediate or higher proficiency in English.
β’ Strong communication and technical leadership abilities.
β’ Options for remote work.
β’ Opportunity to engage with large-scale international products.
β’ Collaboration with a team of highly skilled professionals.
β’ Potential for long-term career growth.
β’ Mentorship and contributions toward engineering excellence.
Data Elephant
ICF
General Dynamics Information Technology
Logic20/20, Inc.
Get handpicked remote jobs straight to your inbox weekly.