
Data Engineer
Posted Jul 29

Posted Jul 29
This is a fully remote position, open to applicants in Massachusetts.
• Assisting in the creation, upkeep, and enhancement of data pipelines and datasets prepared for analytics.
• Working collaboratively with various teams and stakeholders to address intricate challenges and promote data-driven projects.
• Constructing, maintaining, and refining data pipelines using Azure Data Factory, ensuring reliable ingestion, transformation, and delivery of data to Snowflake for analytics.
• Establishing monitoring, alerts, and testing for data pipeline performance, data quality metrics, and lineage to guarantee dependable data delivery.
• Diagnosing data issues and conducting root cause analysis to proactively address operational challenges.
• Documenting data structures, processes, architectural choices, and best practices for knowledge sharing.
• Developing, maintaining, and optimizing Snowflake objects (schemas, tables, views) and SQL transformations to create curated, analytics-ready datasets.
• Collaborating with analysts, stakeholders, and product owners to convert business requirements into data specifications and stable technical implementations.
• Facilitating data for AI/ML applications by preparing feature-rich datasets, aiding in feature engineering, and ensuring consistency for model training and inference.
• Supporting the deployment and operationalization of machine learning models by integrating pipelines with ML workflows (e.g., batch/real-time scoring).
• Continuously enhancing reporting and analytics, automating or streamlining self-service or manual processes.
• Implementing version control practices for all data engineering code and documentation.
• A Bachelor’s degree in Computer Science, Computer Engineering, Information Technology, or a related discipline; or equivalent experience.
• Over 5 years of experience in data engineering or business intelligence roles, working with ETL, data modeling, data architecture, and creating pipelines and applications for analytics (e.g., BI, reporting, machine learning, deep learning).
• Strong programming abilities in advanced SQL, Python, or other programming languages for data processing and automation.
• Experience in supporting or engaging with AI/ML workflows, including data preparation and feature engineering for machine learning models; integration of data pipelines with ML frameworks (e.g., scikit-learn, TensorFlow, PyTorch, or similar).
• Knowledge of model lifecycle concepts (training, validation, deployment, monitoring).
• Proficiency in working with Snowflake for data warehousing, including schema design, performance tuning, and optimization.
• Familiarity with Git, Azure DevOps, and best practices for collaborative development.
• Experience in designing, developing, and deploying end-to-end pipelines using Azure Data Factory.
• Exceptional benefits for you.
• A company culture that prioritizes inclusivity.
• Enjoyment in the workplace.
• A focus on empathy for our valued employees.
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.