
Data Engineer
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in United States.
• Oversee the execution of sophisticated data ingestion pipelines utilizing PySpark, the Databricks Intelligence Platform, AWS services, and Kafka/streaming frameworks.
• Implement schema evolution and Delta Share while ensuring alignment with enterprise architecture.
• Offer technical guidance and mentorship to engineering teams.
• Represent the data engineering team in cross-functional discussions.
• Adhere to standards for code quality, testing, CI/CD, and Terraform infrastructure-as-code.
• Create and maintain architectural documentation, technical decision records, and platform runbooks.
• Lead optimization efforts across data ingestion, transformation, and storage layers.
• Diagnose complex pipeline failures involving multi-system dependencies.
• Collaborate with data scientists, analysts, DevOps engineers, and client teams to deliver cohesive solutions.
• Execute other relevant responsibilities as qualified and trained.
• Bachelor’s degree.
• At least 4 years of relevant experience in Data Engineering.
• Proven experience working on cross-functional engineering projects.
• Proficient in Python for scalable data engineering tasks.
• Practical experience with the Databricks Intelligence Platform, including Notebooks, Lakeflow Jobs, Unity Catalog, and Delta Share.
• Familiarity with AWS services, including S3 and IAM roles and policies.
• Strong understanding of data pipeline design and implementation, data transformation, data modeling, data storage optimization, and best practices in data security.
• Proficient in Git and version control systems.
• Familiarity with Terraform or other infrastructure as code tools.
• Experience in Linux environments, including remote container access, package installation, and file, service, and process management.
• Advanced SQL skills with experience in relational databases, query writing, and various database systems.
• Experience conducting root cause analysis on internal and external data and processes.
• Capability to obtain and retain a Public Trust security clearance.
• Strong problem-solving abilities and the capacity to work independently as well as collaboratively.
• Excellent communication skills, both verbal and written.
• High attention to detail and a commitment to delivering quality software.
• Applicants must be authorized to work for any employer in the U.S.; Bixal does not sponsor employment visas.
• Preferred: Experience with Apache Hive, Apache Hadoop, Apache Spark, federal consulting, Databricks certifications, QuickSight, or Power BI.
• Flexible work hours.
• 401K plan with matching incentive.
• Parental leave.
• Medical, dental, and vision benefits.
• Flexible Spending Account.
• Company-provided short-term disability and life insurance.
• Commuter benefits.
• Paid Time Off (PTO).
• 11 paid holidays.
• Reasonable accommodations for applicants and employees with disabilities.
• Support for professional growth in an inclusive, purpose-driven culture.
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.