
Data Engineer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Virginia.
• Design, develop, and sustain scalable, production-ready data pipelines utilizing Spark, Python, and SQL within a Databricks environment.
• Integrate and transform various health, financial, and operational datasets into dependable, reusable data products.
• Create data models and analytical datasets that support enterprise reporting and operational decision-making.
• Establish and maintain API-based integrations for data ingestion and exchange across systems.
• Monitor, troubleshoot, and enhance scheduled workflows and production pipelines to ensure performance, reliability, and scalability.
• Implement data validation, quality checks, lineage, and governance practices.
• Contribute to automated testing, code reviews, technical documentation, and reusable engineering patterns.
• Collaborate with engineers, architects, analysts, and customer stakeholders to convert data requirements into practical technical solutions.
• Bachelor's Degree or Master's Degree in Computer Science, Engineering, Information Systems, Data Science, or a related technical discipline.
• 3-10 years of experience in data engineering, software engineering, or a related technical role.
• Strong expertise in Python and SQL.
• Experience with Apache Spark or similar distributed data-processing frameworks.
• Familiarity with Databricks or a comparable modern data platform.
• Proven experience in building and maintaining production data pipelines and ETL/ELT processes.
• Working knowledge of data modeling and converting raw data into reusable analytics datasets or data products.
• Experience with integrating or consuming data through REST APIs.
• Skills in troubleshooting and optimizing data pipelines for performance and reliability.
• Awareness of data validation, quality, governance, security, privacy, and compliance considerations.
• Experience with Git-based development workflows and contemporary software engineering practices.
• Proficiency in terminal and command-line environments.
• Experience utilizing AI-assisted development tools to expedite routine engineering tasks, testing, documentation, and debugging.
• Strong communication abilities and experience collaborating with customers or stakeholders to address and resolve data requirements.
• Capacity to navigate uncertainty, learn unfamiliar systems, and independently drive assigned tasks to completion.
• Ability to obtain and maintain a Secret clearance.
• Experience with AWS cloud services.
• Experience handling very large datasets, including those with billions of records.
• Familiarity with Palantir Foundry.
• Experience with Advana or similar Department of War data environments.
• Experience with GitLab, CI/CD pipelines, or automated deployment workflows.
• Experience working with federal health, financial, or other regulated and sensitive data.
• Competitive base salary.
• Opportunities for variable compensation.
• Health insurance coverage.
• Flexible spending accounts.
• Health savings accounts.
• Retirement savings plans.
• Life and disability insurance programs.
• Paid and unpaid time off.
• Industry-leading 401k contribution.
Cypher Consulting Europe S.L.
Sigma Software Group
GFT Technologies
Coinbase
Get handpicked remote jobs straight to your inbox weekly.