
Senior Data Engineer
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in United States.
• Act as a subject-matter expert in the data ecosystem, encompassing both internal systems and third-party data sources.
• Provide guidance on architectural choices across diverse teams.
• Design, construct, and maintain scalable data pipelines and real-time streaming architectures utilizing frameworks like Spark, Kafka, and dbt.
• Create and promote the adoption of workflow automation and orchestration standards using tools such as Apache Airflow or Matillion.
• Lead the technical design for production-quality ML pipelines and APIs that deliver model predictions.
• Utilize AI-assisted development tools like GitHub Copilot, Claude Code, and Cursor for pipeline development, code review, and testing processes.
• Establish team norms for the effective and responsible use of AI-assisted development tools.
• Manage data quality, observability, lineage, and governance strategies.
• Define best practices for monitoring, alerting, and tracking metadata.
• Spearhead logical and physical data modeling initiatives and schema design decisions.
• Collaborate with DevOps and infrastructure teams on platform architecture, performance optimization, and security/compliance strategies.
• Mentor junior and mid-level data engineers through code reviews, technical advice, and knowledge-sharing sessions.
• Assess emerging data tools and technologies.
• Provide build-vs-buy and adoption recommendations to engineering leadership.
• Exhibit eHealth's values in conduct, practices, and decision-making.
• Bachelor's or Master's degree in Computer Science, Engineering, or a related technical discipline.
• 5+ years of relevant data engineering experience with a Bachelor's degree; 3+ years with a Master's degree; or an equivalent combination of education and relevant experience.
• Proficient in SQL for complex query development, optimization, and performance tuning across extensive datasets.
• Strong programming capabilities in Python or Scala.
• Familiarity with CI/CD, git workflows, testing, code review, and API development.
• Experience in architecting solutions on a cloud-native data platform such as Snowflake, BigQuery, Redshift, or Databricks.
• In-depth knowledge of modern ETL/ELT frameworks like dbt, Spark, Informatica, or Matillion.
• Experience with NoSQL databases such as MongoDB, Cassandra, or Hive.
• Significant experience with cloud platforms, preferably AWS, including S3, Glue, Lambda, Redshift, or EMR.
• Strong understanding of data modeling, data governance, and security principles.
• Proven experience in mentoring engineers and/or leading technical projects.
• Excellent communication abilities.
• Preferred: Hands-on experience with Databricks and Delta Lake.
• Preferred: In-depth knowledge of event-driven architectures and tools like Kafka or Kinesis.
• Preferred: Experience in designing RESTful APIs for data delivery and ML model serving.
• Preferred: Experience with Docker and Kubernetes in a production environment.
• Preferred: Visualization experience with Tableau, Power BI, or Looker.
• Preferred: Familiarity with healthcare or health tech, including EHR, claims data, or call center analytics.
• Preferred: Experience working in regulated environments such as HIPAA or SOC 2.
• Preferred: Proven track record of driving automation, data observability, and proactive monitoring initiatives.
• Medical, dental, and vision coverage starting on your first day of employment.
• 401K plan with matching contributions.
• Tuition reimbursement program.
• Employee stock purchase program.
• 12 company-paid holidays.
• Flexible time off (PTO for non-exempt employees).
• Annual performance bonus.
• Professional and personal wellness support through the total rewards package.
CuraLinc Healthcare
VSP Vision Care
Adoreal
Get handpicked remote jobs straight to your inbox weekly.