
Arquiteto de Dados
Posted Aug 14

Posted Aug 14
This is a fully remote position, open to applicants in Brazil.
• Design end-to-end data architectures, from ingestion to consumption, while considering scalability, security, governance, performance, and cost.
• Define solutions for Data Lake, Lakehouse, and Medallion architecture (Bronze, Silver, and Gold).
• Establish strategies for data ingestion, transformation, integration, storage, and availability.
• Create solutions using Microsoft Azure, Microsoft Fabric, and Azure Databricks.
• Set standards for dimensional modeling, including Star Schema, Snowflake, Fact, Dimension, granularity, and SCD.
• Develop strategies for Full, Incremental, CDC, Upsert/MERGE loading, deduplication, and historical control.
• Assess performance, scalability, availability, and costs, including cluster sizing and Spark/PySpark processing.
• Establish standards for security, governance, quality, metadata, and lineage, considering LGPD and compliance.
• Evaluate integrations with SAP/SAP HANA, relational databases, NoSQL, APIs, files, and SaaS systems.
• Guide Data Engineers and teams in Cloud, DevOps, Security, Governance, and BI in the implementation of solutions.
• Produce documentation, diagrams, and architectural decisions, assessing technological alternatives and their technical and financial impacts.
• Bachelor's degree in Computer Science, Engineering, Information Systems, or related fields.
• Strong experience in Data Architecture, Data Lake/Lakehouse, and Medallion architecture.
• Proficiency in dimensional modeling, including Star Schema, Snowflake, Fact, Dimension, granularity, and SCD.
• Knowledge of SQL, Python and/or PySpark, Apache Spark, Delta Lake, and Parquet.
• Experience with ETL/ELT, data pipelines, and strategies for Full, Incremental, CDC, Upsert/MERGE, and distributed processing.
• Solid experience or knowledge in Microsoft Azure, Microsoft Fabric, and Azure Databricks.
• Familiarity with Azure Data Lake Storage Gen2, OneLake, Fabric Lakehouse, and Unity Catalog.
• Experience or strong knowledge in integration with SAP/SAP HANA, relational databases, NoSQL, APIs, and files.
• Understanding of security, governance, quality, data lineage, LGPD, and access control.
• Familiarity with Azure Data Factory, Fabric Data Factory/Data Pipelines, and Databricks Workflows.
• Ability to transform business requirements into scalable, secure, and sustainable data architectures.
• Skill in guiding multidisciplinary teams and making technical decisions on complex projects.
• Analytical capacity to evaluate trade-offs between performance, cost, security, and scalability.
• Proactive in defining standards, automation, and continuous improvement of data solutions.
• Ease in documenting and communicating technical concepts to both technical and non-technical audiences.
• Microsoft certifications related to Azure/Fabric and Databricks certifications are advantageous.
• Experience with Azure DevOps, CI/CD, and Terraform/IaC is a plus.
• Experience with Data Mesh, streaming, and event-driven architectures is a plus.
• Knowledge of Power BI, Semantic Models, FinOps, and hybrid or multi-cloud architectures is a plus.
• Experience in Machine Learning, MLOps, MLflow, Databricks ML, and data architectures for Generative AI is a plus.
• Experience in large-volume corporate projects and critical environments is a plus.
• Comprehensive health and wellness programs.
• Opportunities for professional development and certification.
• Flexible working hours and remote work options.
• Collaborative and innovative work environment.
Data Elephant
ICF
General Dynamics Information Technology
Logic20/20, Inc.
Get handpicked remote jobs straight to your inbox weekly.