
Senior Data Engineer
Posted Jul 17

Posted Jul 17
This is a fully remote position, open to applicants in United States.
• Design, construct, and sustain scalable data pipelines along with ELT/ETL workflows that facilitate analytics, operational reporting, and business intelligence applications.
• Develop programmatic data pipelines (primarily utilizing Python) that extract data from applications and third-party systems, transform it into usable formats, and deliver it to downstream data platforms and consumers.
• Take ownership of and enhance core data models and transformations to ensure data accuracy, structure, and ease of use for stakeholders.
• Collaborate with Product, Engineering, and Analytics teams to comprehend data requirements and convert them into dependable data solutions.
• Create and manage systems that facilitate data movement across the platform, ensuring it is appropriately shaped, structured, and accessible for downstream analysis and product applications.
• Assist in shaping and upholding the architecture of MDCalc’s contemporary data stack, encompassing warehousing, orchestration, transformation, and monitoring.
• Enhance data quality, observability, and reliability through testing, validation, and proactive monitoring techniques.
• Aid in the ingestion and integration of data from a diverse array of application, product, and third-party sources.
• Establish and promote best practices surrounding data governance, documentation, naming conventions, and maintainability.
• Identify and pursue opportunities to enhance performance, scalability, and efficiency within our data systems.
• Design effective data workflows that query, transform, and deliver datasets to downstream systems and stakeholders.
• Contribute to technical direction and architectural decisions as a senior team member.
• Act as a thought partner to colleagues and cross-functional stakeholders on optimal ways to leverage data throughout the business.
• Over 5 years of experience in data engineering.
• Strong SQL capabilities and experience in building and optimizing data models for analytical purposes.
• Proven experience in developing and maintaining reliable data pipelines within a modern cloud data environment.
• High proficiency in Python or a similar programming language typically used in data engineering.
• Experience in creating programmatic ETL/ELT pipelines using Python or comparable tools to transfer and transform data across systems.
• Familiarity with data warehouses such as Snowflake.
• Experience with transformation and orchestration tools like dbt, Airflow, Dagster, or similar platforms.
• Comprehensive understanding of data architecture, data modeling, and best practices in pipeline design.
• Capability to work independently, prioritize tasks effectively, and advance work in a fast-paced environment.
• Opportunity to make a significant impact in medicine: MDCalc is the most widely used medical reference utilized by 65% of physicians globally.
• Coverage for Medical, Dental, & Vision, with options to extend to your dependents.
• Company-sponsored short-term insurance.
• Fully-paid 8-week parental leave after 6 months of employment.
• Company-sponsored 401k after 3 months of employment.
• Unlimited vacation for salaried positions - we trust you to take the time you need.
• Tri-annual company offsites to connect, reflect, and strategize together.
• Monthly work-from-home stipend.
• Hybrid work environment with a fantastic team office located in Greenwich Village, NYC.
• A culture filled with fun and motivated team members who are dedicated to a greater mission here at MDCalc.
Heinsohn
Live Nation Entertainment
Get handpicked remote jobs straight to your inbox weekly.