
Senior Databricks Engineer β Lead
Posted 2 days ago

Posted 2 days ago
This is a fully remote position, open to applicants in India.
β’ Design, develop, and sustain scalable data pipelines utilizing Azure Databricks.
β’ Execute ETL/ELT workflows leveraging PySpark, Spark SQL, and Python.
β’ Enhance Spark jobs for improved performance, cost efficiency, and scalability.
β’ Work with both structured and semi-structured data formats such as Parquet, Delta, JSON, and CSV.
β’ Construct and manage Delta Lake tables with features like ACID transactions, time travel, and schema evolution.
β’ Integrate Databricks with Azure Data Lake Storage Gen2.
β’ Create complex queries and transformations using SQL.
β’ Collaborate with data scientists, analysts, and stakeholders to facilitate analytics and machine learning use cases.
β’ Ensure data quality through validation and monitoring processes.
β’ Adhere to best practices regarding security, access control, and governance within Azure.
β’ Over 6 years of experience in Data Engineering.
β’ Strong practical experience with Azure Databricks.
β’ Proficiency in Python for data processing tasks.
β’ Solid understanding of SQL, including joins, window functions, and performance optimization techniques.
β’ Hands-on experience with Apache Spark and PySpark.
β’ Familiarity with Delta Lake.
β’ Knowledge of Azure Data Lake Storage Gen2.
β’ Comprehension of distributed computing concepts.
β’ Experience utilizing Git for version control.
β’ Proficiency with Azure Data Factory.
β’ Exposure to CI/CD pipelines, including Azure DevOps and GitHub Actions.
β’ Basic understanding of data modeling principles.
β’ Awareness of cloud security and Role-Based Access Control (RBAC) in Azure.
β’ Familiarity with streaming data technologies such as Spark Structured Streaming, Event Hub, and Kafka.
β’ Remote work is permitted.
Sigma Software Group
Collectly
Allata
Get handpicked remote jobs straight to your inbox weekly.