
Principal Data Platform Engineer β Swedish Ad Platform
Posted Sep 4

Posted Sep 4
This is a fully remote position, open to applicants in Germany, +1 more country.
β’ Lead the technical design and architecture of a modern enterprise-scale data platform.
β’ Define scalable data flows transitioning from Iceberg-based event storage into analytical and operational data products.
β’ Evaluate and select technologies for orchestration, transformation, querying, serving, and storage solutions.
β’ Design reusable canonical entities and data models across advertising, campaigns, inventory, billing, and customer domains.
β’ Establish scalable engineering patterns for both batch and near-real-time processing.
β’ Build reliable and observable data pipelines capable of handling high-volume AdTech workloads.
β’ Implement data quality, lineage, observability, and reconciliation capabilities.
β’ Collaborate with Platform Engineering teams to introduce CI-enforced data contracts and governance standards.
β’ Define strategies for tenant isolation, access control, and regional data boundaries.
β’ Establish standards for testing, deployment automation, schema evolution, and versioning.
β’ Optimize platform performance, scalability, and infrastructure costs.
β’ Mentor engineers and contribute to engineering excellence across the team.
β’ Partner closely with leadership and cross-functional stakeholders.
β’ Deliver the first production-ready version of the platform.
β’ Establish the initial production data foundation, including documented architectural decisions, enforceable event contracts, tested canonical entities, and observable data quality and lineage.
β’ 8+ years of experience in designing and operating production-grade data platforms.
β’ Strong expertise in distributed data systems and high-volume event-driven architectures.
β’ Deep understanding of modern lakehouse architectures and Apache Iceberg.
β’ Advanced SQL skills.
β’ Production experience with Python, Java, Scala, or similar programming languages.
β’ Experience with Spark, Flink, or Trino.
β’ Strong knowledge of data modeling, partitioning strategies, and performance optimization techniques.
β’ Proven experience in building both batch and near-real-time data pipelines.
β’ Hands-on experience with AWS cloud infrastructure.
β’ Strong understanding of CI/CD pipelines, Infrastructure as Code, and production observability.
β’ Experience implementing data contracts, schema evolution, and data quality frameworks.
β’ Ability to make pragmatic architectural decisions in greenfield environments.
β’ Upper-Intermediate or higher level of English proficiency.
β’ Strong communication and technical leadership skills.
β’ Remote work options available.
β’ Long-term growth potential.
β’ Opportunity to work on large-scale international products.
β’ Collaboration with highly skilled professionals.
β’ A technically ambitious environment.
Data Elephant
ICF
General Dynamics Information Technology
Logic20/20, Inc.
Get handpicked remote jobs straight to your inbox weekly.