
Databricks Solution Architect
Posted Jul 15

Posted Jul 15
This is a fully remote position, open to applicants in United States.
β’ Design and lead the deployment of an enterprise lakehouse on Databricks (Delta Lake, Unity Catalog, Photon, Workflows) across one or multiple major cloud platforms (AWS, Azure, or GCP).
β’ Create scalable batch and streaming data pipelines utilizing PySpark, Spark SQL, Structured Streaming, and Delta Live Tables; establish best practices for data ingestion from operational systems, event streams, and third-party APIs.
β’ Set and uphold platform standards for data modeling (medallion architecture), CI/CD, code quality, testing, observability, and cost efficiency.
β’ Direct the governance strategy using Unity Catalog β implementing fine-grained access control, data lineage, auditing, and PII management β in collaboration with security and compliance teams.
β’ Enhance Spark workload performance and cost-effectiveness: including cluster sizing, Photon, autoscaling, file layout, Z-ordering, caching, and query tuning.
β’ Collaborate with ML engineers and data scientists to operationalize models using MLflow, feature stores, and model serving on Databricks.
β’ Manage the cloud infrastructure for the platform, encompassing networking, IAM, secrets, encryption, and Terraform/IaC for Databricks workspaces and related services.
β’ Mentor a team of data engineers; facilitate architecture reviews, code reviews, and technical design discussions; elevate engineering standards.
β’ Work with stakeholders across analytics, product, and finance to transform business requirements into a strategic roadmap for the data platform.
β’ Over 8 years of data engineering experience, with a minimum of 4 years developing production workloads on Databricks.
β’ Extensive knowledge of Apache Spark (PySpark and Spark SQL) β including performance optimization, partitioning strategies, and the Catalyst/Photon execution model.
β’ Significant hands-on experience with Delta Lake, Unity Catalog, Databricks Workflows, and Delta Live Tables.
β’ Practical experience on at least one major cloud provider (AWS, Azure, or GCP), including networking, IAM, storage (S3/ADLS/GCS), and compute resources.
β’ Proficient in Python and SQL; familiarity with Scala is advantageous.
β’ Experience in designing medallion (bronze/silver/gold) architectures and dimensional models for analytics.
β’ Strong CI/CD and DevOps practices: including Git, Terraform, Databricks Asset Bundles or dbx, and automated testing of data pipelines.
β’ Proven record of leading technical projects from inception to completion and mentoring engineering teams.
β’ Exceptional written and verbal communication skills; capable of fostering alignment between engineering and business stakeholders.
β’ Bounteous is proud to be an equal opportunity employer.
β’ Bounteous does not discriminate on the basis of race, religion, color, sex, gender identity, sexual orientation, age, physical or mental disability, national origin, veteran status, or any other status protected under federal, state, or local law.
β’ Bounteous is willing to sponsor eligible candidates for employment visas.
Ramp
Park Place Technologies
Valtech
Talkdesk
Get handpicked remote jobs straight to your inbox weekly.