
Staff Data Engineer
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in Canada.
• Design, develop, and manage scalable data pipelines along with ETL/ELT processes.
• Architect and enhance data models and storage solutions for analytical and operational purposes.
• Collaborate with data scientists, analysts, and engineers to provide reliable, high-quality datasets.
• Take ownership of and advance components of the data platform utilizing Databricks and Spark.
• Establish observability, alerting, and data quality monitoring for essential pipelines.
• Promote best practices in data engineering, including documentation, testing, and CI/CD.
• Shape the long-term architectural vision and mentor the team on engineering excellence.
• Work alongside stakeholders to ensure data solutions align with business objectives.
• Assist in the design and progression of the next-generation data lakehouse architecture.
• Over 8 years of experience as a Data Engineer or in a comparable backend engineering position.
• Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
• Expertise in Databricks optimization, including tuning Spark jobs, enhancing joins, and managing Delta Lake architecture.
• Familiarity with AI-assisted development tools and AI/ML technologies.
• Strong programming capabilities in Python, Scala, or Java.
• Practical experience with distributed data systems such as Spark or Kafka.
• Proficient in writing and optimizing complex SQL and NoSQL queries.
• Experience in building and sustaining robust production ETL/ELT pipelines.
• Understanding of data modeling techniques, including star schema and dimensional modeling.
• Knowledge of software supply chains, cybersecurity, or data from large-scale software ecosystems.
• Proven history of enhancing data platform reliability, scalability, performance, and cost-effectiveness.
• Familiarity with workflow orchestration tools such as Airflow or Dagster.
• Hands-on experience with cloud data platforms, especially AWS.
• Understanding of Delta Lake, Apache Iceberg, or Apache Hudi.
• Experience in implementing data observability, lineage, governance, and automated data quality frameworks.
• Experience in designing real-time or streaming data architectures utilizing data lake technologies.
• Parental leave.
• Diversity and inclusion working groups.
• Flexible working practices.
• Paid Volunteer Time Off (VTO).
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.