
Data Engineer
Posted Aug 26

Posted Aug 26
This is a fully remote position, open to applicants in United States.
• Design, develop, test, and maintain scalable data pipelines and ETL/ELT processes for cloud-based data environments.
• Facilitate bidirectional data movement between VA enterprise data warehouses, cloud workgroups, and operational systems.
• Implement logic for patient matching and record linkage across multiple identifiers.
• Harmonize datasets across VA facilities and data sources by addressing structural, semantic, and format inconsistencies.
• Generate analysis-ready, high-quality data assets.
• Create and uphold data flow documentation, administrator guides, and technical documentation in compliance with federal contracting and VA standards.
• Offer pipeline support and maintenance, troubleshoot data quality issues, optimize performance, and execute transition planning.
• Collaborate with data analysts, informatics specialists, and program stakeholders to translate business and reporting requirements into scalable data engineering solutions.
• Develop Power BI reports utilizing Databricks SparkSQL.
• Over 8 years of relevant experience in data engineering, including the design and development of ETL/ELT pipelines and cloud-based data environments.
• 13+ years of relevant experience may be considered in lieu of a BA/BS degree.
• A BA/BS degree in Computer Science, Information Systems, Data Science, or a related technical field is required; 5 years of additional relevant experience may substitute for the degree requirement.
• Previous experience supporting federal agency programs, preferably within the Department of Veterans Affairs, Department of Defense, or another federal health or human services agency.
• Capability to obtain and maintain a VA Position of Public Trust clearance.
• Ability to adhere to VA information security, privacy, and system access requirements, including necessary VA training and onboarding.
• Must be a U.S. citizen or authorized to work in the United States in compliance with applicable federal contracting requirements.
• Proficiency in SQL or Spark SQL for data transformation and querying.
• Hands-on experience implementing record-linkage or patient-matching logic across multi-source or multi-facility datasets.
• Desired: hands-on experience with Azure Databricks and PySpark.
• Desired: proficiency in Python.
• Desired: experience in developing Power BI reports and dashboards using Databricks SparkSQL.
• Desired: familiarity with Salesforce or Red Hat products such as Ansible and OpenShift.
• Desired: knowledge of VA or federal healthcare data standards such as HL7, FHIR, and CDW/Corporate Data Warehouse.
• Equal opportunity employer.
• Talent Community notifications about matching job opportunities, company news, and upcoming career events.
Solar Coca-Cola
Get handpicked remote jobs straight to your inbox weekly.