
Data Engineer
Posted Jul 27

Posted Jul 27
This is a fully remote position, open to applicants in Romania.
• Develop and sustain Kafka ingestion pipelines
• Establish topic structures, partition strategies, retention policies, and consumer logic for multi-tenant environments
• Oversee data contracts and manage schema evolution
• Create idempotent ingestion services that transfer data into Iceberg tables
• Design and enhance Iceberg tables (including partitioning, compaction, clustering, and retention rules)
• Collaborate with Nessie branches/tags to oversee multi-environment (dev/test/prod) and multi-MNO deployments
• Construct Spark jobs (both batch and micro-batch when necessary) for cleaning and normalizing events, categorizing senders, and performing identity stitching and grouping
• Develop dbt models on top of Iceberg datasets utilizing Dremio and dbt
• Over 3 years of experience as a Data Engineer or in a similar role, with proven experience in handling large datasets
• Proficient Python programming skills (essential)
• Experience in creating pipelines on Apache Kafka (preferably in KRaft mode)
• Strong SQL skills with experience in Iceberg table design and optimization
• Experience with Spark for processing at scale
• Familiarity with dbt and SQL modeling on lakehouse storage
• Experience working with Dremio, Trino, or comparable query engines
• Experience with Kubernetes, Helm, and Git-based CI/CD practices
• Knowledge of PII handling, encryption, and compliance standards
• Capability to operate in distributed, multi-environment settings (dev/test/prod + multi-deployments)
• Competitive salary and benefits package
• Vacation and time-off offerings
• Stock options
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.