
Data Engineer – AI-Driven Data Platforms
Posted Jul 23

Posted Jul 23
This is a fully remote position, open to applicants in United States.
• This position is remote.
• We are looking for a Data Engineer with over 5 years of experience to create and enhance data platforms that support AI-driven analytics utilizing high-volume sensor, satellite, and third-party data.
• The ideal candidate should have a strong background in designing and managing scalable batch and real-time data pipelines using technologies such as Apache Airflow, Apache Spark, Kafka, Python, and SQL.
• Significant expertise with PostgreSQL, TimescaleDB, PostGIS, cloud-based data lake architectures, as well as geospatial and time-series data is crucial.
• The role requires taking ownership of data pipelines from end to end—covering everything from real-time ingestion to analytics—while maintaining data quality, performance, observability, and reliability in a cloud-native setting.
• Strong skills in stakeholder management, problem-solving, and collaboration are essential for delivering impactful data products and analytical solutions.
• Minimum of 5 years of experience in Data Engineering and with production-scale data platforms.
• Experience in designing and managing Apache Airflow pipelines for the ingestion of high-volume sensor, satellite, and third-party data.
• Ability to build and optimize Apache Spark workloads for batch processing, geospatial analytics, and large-scale aggregations.
• Proficient in developing and maintaining data solutions based on PostgreSQL, TimescaleDB, PostGIS, and DuckDB.
• Familiarity with implementing Kafka-based streaming pipelines for real-time data ingestion.
• Experience working with Object Storage (such as AWS S3 or equivalent) for scalable data lake architecture.
• Strong proficiency in Python, SQL, Airflow, Spark, and PostgreSQL.
• Experience with geospatial and time-series data.
• Comfortable working in Ubuntu-based environments.
• Collaborate with customer teams to convert business requirements into data products and analytical solutions.
• Ensure data quality, lineage, observability, performance, and adherence to SLA.
• Ability to troubleshoot production issues and enhance pipeline, database, and query performance.
• Create and maintain technical documentation, data models, and run-books.
• Strong communication, stakeholder management, and problem-solving abilities.
• Comprehensive insurance coverage providing peace of mind, allowing you to focus on delivering your best work.
• Flexible work arrangements designed to promote sustained productivity, personal well-being, and a healthy work-life balance.
• Opportunities for continuous learning and accelerated skill development through hands-on projects and mentorship from seasoned industry leaders.
• Global client exposure across more than 20 countries, offering valuable experience in diverse markets and business environments.
• Chance to work on high-impact, large-scale projects that have collectively generated over $1 billion in measurable business value.
• Competitive, market-aligned compensation packages that acknowledge performance, expertise, and long-term contributions.
• Monthly demo days to celebrate innovation, showcase your work, and provide you a real voice in our development processes.
• Annual recognition programs and performance-driven awards in a truly meritocratic environment.
• Referral bonuses to reward you for helping to grow a strong, like-minded team.
• A strong problem-solving culture with opportunities to address meaningful, real-world challenges.
• A positive, people-first workplace that fosters happiness, balance, and long-term growth.
Railroad19
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.