
Principal Data Platform Engineer β Swedish Ad Platform
Posted Sep 11

Posted Sep 11
This is a fully remote position, open to applicants in Germany, +1 more country.
β’ Oversee the technical design and architecture of a contemporary enterprise-scale data platform.
β’ Create scalable data flows from Iceberg-based event storage into analytical and operational data products.
β’ Assess and choose technologies for orchestration, transformation, querying, serving, and storage.
β’ Design reusable canonical entities and data models across domains such as advertising, campaigns, inventory, billing, and customers.
β’ Develop scalable engineering patterns for both batch and near-real-time processing.
β’ Construct dependable and observable data pipelines for high-volume AdTech workloads.
β’ Implement capabilities for data quality, lineage, observability, and reconciliation.
β’ Work in collaboration with Platform Engineering teams to introduce CI-enforced data contracts and governance standards.
β’ Define strategies for tenant isolation, access control, and regional data boundaries.
β’ Set standards for testing, deployment automation, schema evolution, and versioning.
β’ Enhance platform performance, scalability, and infrastructure costs.
β’ Guide engineers and contribute to engineering excellence within the team.
β’ Collaborate closely with leadership and cross-functional stakeholders.
β’ Deliver the initial production-ready version of the platform.
β’ Within the first six months, establish the foundational production data, document architectural decisions, enforce event contracts, create validated canonical entities, ensure data quality and lineage visibility, and empower other engineers to contribute.
β’ Over 8 years of experience in designing and operating production-grade data platforms.
β’ Extensive expertise in distributed data systems and high-volume event-driven architectures.
β’ Profound understanding of modern lakehouse architectures and Apache Iceberg.
β’ Advanced SQL proficiency with practical experience in Python, Java, Scala, or similar programming languages.
β’ Familiarity with distributed processing and query technologies like Spark, Flink, or Trino.
β’ Strong knowledge of data modeling, partitioning strategies, and performance optimization.
β’ Proven track record in building both batch and near-real-time data pipelines.
β’ Practical experience with AWS cloud infrastructure.
β’ Solid understanding of CI/CD pipelines, Infrastructure as Code, and production observability.
β’ Experience in implementing data contracts, schema evolution, and data quality frameworks.
β’ Capability to make pragmatic architectural decisions in greenfield settings.
β’ Competency in English at an Upper-Intermediate level or higher.
β’ Strong communication skills and technical leadership abilities.
β’ Experience in AdTech or other high-volume event-processing fields is an advantage.
β’ Background in designing multi-tenant SaaS data architectures is a plus.
β’ Familiarity with semantic layer technologies is an additional benefit.
β’ Experience in supporting both analytics and ML/AI workloads is advantageous.
β’ Understanding of privacy regulations and data residency requirements is a plus.
β’ Option for remote work.
β’ Long-term growth potential.
β’ Opportunity to work on large-scale international products.
β’ Collaboration with highly skilled professionals.
β’ A technically ambitious environment.
Data Elephant
ICF
General Dynamics Information Technology
Logic20/20, Inc.
Get handpicked remote jobs straight to your inbox weekly.