
Lead Data Engineer – Identity
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Europe.
• Develop and lead a new multi-source identity graph that encompasses identifier synchronization, translation, clustering in collaboration with Data Science, and opt-out management.
• Standardize the onboarding process for partners and clients across web, CTV, and mobile identifiers, including the onboarding process for cleanrooms.
• Prepare the identity audience data layer for self-service creation, activation, state management, and client reporting.
• Take ownership of and enhance observability and alert responses, which include signals, monitors, alerts, freshness and quality commitments, AI-assisted triage context, and runbooks.
• Guide and cultivate domain data engineers by establishing standards for testability, cost efficiency, and reusable patterns.
• Set the technical direction through hands-on coding, code reviews, knowledge sharing, and documenting decisions in Architectural Decision Records (ADRs).
• Collaborate with Product and Data Partnerships to transform ambiguity into a structured roadmap.
• Proven experience in designing and managing large-scale, interdependent data systems, including at least one system developed from the ground up.
• Capability to convert ambiguity into a structured roadmap in partnership with Product and Data Partnerships.
• Experience in leading engineers, establishing direction, reviewing work, nurturing talent, and maintaining a hands-on approach.
• Proficiency in Python, Airflow, and Spark.
• Ability to create idiomatic, testable, cost- and performance-optimized transformations.
• Strong SQL skills specifically for Snowflake.
• Familiarity with AWS and Kubernetes.
• Competence in diagnosing failures and performance issues by analyzing infrastructure logs.
• Experience with integrating third-party APIs within ingestion pipelines.
• Proficiency with AI tools and designing codebases that are easily interpretable by AI technologies.
• Strongly preferred: experience in identity resolution or graph work within AdTech; knowledge of privacy and consent obligations including opt-outs, deletion, GDPR, and CCPA; familiarity with data cleanrooms; CI/CD experience with GitHub Actions/ArgoCD; monitoring with VictoriaMetrics/Prometheus/Grafana.
• Nice to have: experience with Iceberg or similar table formats at production scale; streaming or near-real-time processing using Kafka, Redpanda, or similar technologies; knowledge of low-latency stores like Aerospike; experience with OLAP databases such as ClickHouse.
• Equal Opportunity Employer dedicated to fostering an inclusive and diverse work environment.
• Consideration of qualified candidates with arrest and conviction records in accordance with applicable fair chance laws.
Pluribus Digital
Get handpicked remote jobs straight to your inbox weekly.