
Data Engineer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Virginia.
• Create, develop, and sustain scalable, production-ready data pipelines utilizing Spark, Python, and SQL within a Databricks framework.
• Integrate and transform various health, financial, and operational datasets into dependable, reusable data products.
• Construct data models and analytics datasets that support enterprise reporting and operational decision-making.
• Develop and maintain API-based integrations for data ingestion and exchange across systems.
• Monitor, troubleshoot, and enhance scheduled workflows and production pipelines to ensure performance, reliability, and scalability.
• Establish data validation, quality checks, lineage, and governance practices to bolster trust in data products.
• Participate in automated testing, code reviews, technical documentation, and the creation of reusable engineering patterns.
• Work collaboratively with engineers, architects, analysts, and customer stakeholders to convert data requirements into effective technical solutions.
• Assist the Defense Health Agency Chief Data and Analytics Office in developing an enterprise data orchestration layer for the Military Health System.
• Active Secret clearance.
• Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related discipline.
• 3 to 10 years of experience in data engineering, software engineering, or a related technical field.
• Strong expertise in Python and SQL.
• Familiarity with Apache Spark or similar distributed data-processing frameworks.
• Experience with Databricks or a comparable modern data platform.
• Proven track record in building and maintaining production data pipelines and ETL/ELT processes.
• Working knowledge of data modeling and converting raw data into reusable analytics datasets or data products.
• Experience in integrating or consuming data via REST APIs.
• Skilled in troubleshooting and optimizing data pipelines for performance and reliability.
• Awareness of data validation, quality, governance, security, privacy, and compliance considerations.
• Proficient with Git-based development workflows and contemporary software engineering methodologies.
• Comfortable working in terminal and command-line environments.
• Experience utilizing AI-assisted development tools to expedite routine engineering tasks, testing, documentation, and debugging.
• Strong communication abilities with experience in collaborating with customers or stakeholders to comprehend and address data needs.
• Capability to navigate ambiguity, learn new systems, and independently drive assigned tasks toward resolution.
• AWS cloud experience (preferred).
• Experience working with exceptionally large datasets, including those containing billions of records (preferred).
• Familiarity with Palantir Foundry (preferred).
• Experience with Advana or similar DoW data environments (preferred).
• Experience with GitLab, CI/CD pipelines, or automated deployment workflows (preferred).
• Background working with federal health, financial, or other regulated sensitive data (preferred).
• Competitive base salary.
• Opportunities for variable compensation.
• Health insurance coverage.
• Flexible spending accounts.
• Health savings accounts.
• Retirement savings plans.
• Life and disability insurance programs.
• Paid and unpaid time off from work.
Cypher Consulting Europe S.L.
Sigma Software Group
GFT Technologies
Coinbase
Get handpicked remote jobs straight to your inbox weekly.