
Data Engineering Specialist
Posted Sep 9

Posted Sep 9
This is a fully remote position, open to applicants in Brazil.
• Take ownership of data domains from inception to the point where the business utilizes the information daily for decisions regarding purchasing, pricing, inventory, and assortment.
• Develop and enhance comprehensive ingestion and transformation pipelines (ELT/ETL), both in batch and near real time, utilizing data sourced from SAP, Oracle, PostgreSQL, APIs, and files.
• Manage workloads within Cloud Composer (Airflow), ensuring adherence to SLA standards, idempotency, reprocessing, and effective monitoring.
• Design the data warehouse in BigQuery with a layered architecture (Medallion—bronze / silver / gold), using Dataform to deliver dependable, high-performance models for use in Power BI.
• Implement and maintain Change Data Capture (CDC) replication of transactional databases using Datastream and similar technologies.
• Handle substantial volumes of data with PySpark (Dataproc) and Apache Beam (Dataflow) as needed based on the use case.
• Enhance BigQuery performance and manage costs through strategies like partitioning, clustering, query optimization, and data FinOps practices.
• Establish technical standards, review code, and act as a technical resource for the data team.
• Convert business requirements from Commercial, Supply Chain, Finance, and Marketing into effective data solutions.
• A minimum of 5 years of robust experience as a Data Engineer, showcasing senior-level capabilities.
• Proficiency in advanced SQL is essential, including expertise in window functions, CTEs, query optimization, and execution plan analysis.
• Demonstrated experience with Python in data engineering contexts, focusing on modular, tested, and version-controlled coding practices.
• Proven track record in constructing end-to-end ELT/ETL pipelines, with a focus on failure management, reprocessing, and idempotency.
• Familiarity with Airflow (or Cloud Composer) in a production setting is required.
• Experience in managing public cloud data environments in production is essential.
• Knowledge of cloud data warehouses—preferably BigQuery; experience with Snowflake, Redshift, or Synapse is also advantageous.
• Solid background in data modeling and layered architecture (Medallion / dimensional) is required.
• Strong understanding of relational databases such as PostgreSQL, Oracle, SQL Server, or MySQL.
• Experience with various file formats (Parquet, CSV, JSON, Avro) and consuming REST APIs.
• Proficiency in Git, including experience with branching strategies, pull requests, and conducting code reviews.
• Health insurance
• Dental insurance
• Life insurance
• Meal and/or food allowance
• Grocery allowance
• Transportation allowance
• Childcare assistance
• Annual bonus based on individual and company performance
• TotalPass
• Petlove plan
• Birthday day off
• Discounts at our stores
• Career development plan
• Discounts through partnerships for courses and undergraduate programs
CuraLinc Healthcare
VSP Vision Care
Adoreal
Get handpicked remote jobs straight to your inbox weekly.