
Data Engineer
Posted Jul 29

Posted Jul 29
This is a fully remote position, open to applicants in India.
• Design and sustain intricate data pipelines utilizing AWS Glue, Step Functions, or Databricks Workflows.
• Create modular data structures employing sophisticated modeling techniques such as Medallion Architecture and Dimensional Modeling.
• Oversee scalable data storage solutions, with AWS S3 serving as the primary landing zone and foundational data lake.
• Enhance storage formats (Delta, Iceberg, Parquet) and computing performance to guarantee high-throughput and cost-efficient processing.
• Construct decoupled, event-driven architectures with AWS SNS and SQS to manage high-throughput messaging between data services.
• Develop and implement real-time ingestion pipelines using AWS Kinesis or Kafka.
• Apply Change Data Capture (CDC) through tools such as Debezium or Fivetran to facilitate low-latency operational analytics.
• Take ownership of end-to-end data validation and QA by integrating automated data quality checks directly into the ETL/ELT pipelines.
• Enforce rigorous data contracts and schema evolution guidelines to uphold high data quality and integrity across domains.
• Establish proactive alerting and observability to detect data drift, pipeline anomalies, and quality declines before they affect downstream users.
• Engineer ML-ready datasets and manage Feature Stores to support the Data Science team.
• Operationalize ML workflows by integrating with services such as Snowflake Cortex, Databricks AI, or AWS Bedrock.
• Follow coding best practices, SQL optimization, and Python development standards.
• Work closely with Product and ML teams to convert architectural designs into functional code.
• 5+ years of experience in data engineering, particularly in large-scale distributed systems.
• Proficient in Python and PySpark with strong SQL capabilities.
• Extensive hands-on experience with Snowflake or Databricks, built primarily within an AWS ecosystem.
• Proven experience in developing streaming applications using Kinesis or Kafka.
• Demonstrated ability to implement automated testing frameworks, data profiling, and pipeline validation (taking ownership of QA for your own pipelines).
• Strong documentation skills (playbooks, technical specifications) along with an ownership mentality.
• Relevant IT professional certifications, such as SnowPro Core, Databricks Certified Data Engineer Professional, or AWS Certified Data Engineer (Nice-to-Have).
• Excellent communication skills with the capability to convey technical concepts clearly to both technical and non-technical stakeholders.
• Collaborative approach with the ability to effectively partner across Product, Engineering, Analytics, ML, and leadership teams.
• High expectations for quality, maintainability, performance, and operational discipline.
• Strong ownership mentality with the ability to act swiftly and resolve issues thoughtfully.
• Competitive base salary: ₹5,750,000 INR/yr
• Participation in Dynatron’s Equity Incentive Plan
• Comprehensive health, dental, and vision insurance
• Employer-paid disability and life insurance
• 401(k) with competitive company match
• Flexible vacation policy and 11 paid holidays
• Remote-first culture
• Ongoing professional development opportunities
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.