Remotery

Staff Engineer – Data Engineer

Posted 3 days ago

This is a fully remote position, open to applicants in Brazil.

📋 Description

• Create logical and physical data models for unstructured and semi-structured content derived from KM pipelines.

• Define the boundaries and ownership of data products within specific domains.

• Establish standards for metadata and tagging taxonomies across various knowledge resources.

• Assign and uphold security and sensitivity classifications in accordance with governance, privacy, and legal/risk standards.

• Register, document, and maintain data products in the Databricks Unity Catalog, including schemas, access permissions, lineage, and catalog-level metadata.

• Collaborate with data engineers to synchronize ingestion, transformation, and storage methods with the modeled domain structures.

• Work alongside stakeholders from Knowledge Products, Research Products, and Architecture/Data/Technology to address downstream consumption requirements.

• Assist in privacy and legal review processes through the classification and documentation of data products.

• Develop and document repeatable modeling standards and playbooks.

• Deliver domain models, metadata taxonomies, registered and discoverable data products, security classifications, and repeatable modeling standards within the initial 6–12 months.


⛳️ Requirements

• Minimum of 5 years of experience in data modeling, data architecture, or information architecture.

• Significant exposure to unstructured or semi-structured data.

• Direct experience in or related to Knowledge Management, content management, or enterprise search.

• Practical experience with a modern data catalog; familiarity with Databricks Unity Catalog is highly preferred.

• Capability to define data domains and product boundaries in a large, multi-stakeholder environment.

• Hands-on knowledge of metadata management, tagging schemas, taxonomies, controlled vocabularies, or ontology design.

• Understanding of data security and sensitivity classification frameworks along with access control in a lakehouse environment.

• Experience collaborating with data engineering teams on ingestion and pipeline design.

• Excellent written and verbal communication skills.

• Preferred experience with enterprise knowledge platforms or AI-driven retrieval systems.

• Familiarity with Databricks Delta Lake, Delta Sharing, or Lakehouse Federation is preferred.

• Previous experience in professional services, consulting, or document/case-intensive knowledge environments is preferred.

• Exposure to Legal/Risk/Privacy review processes is preferred.

• A background in library science, information science, or applied ontology is advantageous but not mandatory.

• Must possess strong Data Modeling and Databricks skills.


🏝️ Benefits

• Opportunity for remote work.

People also viewed

Railroad195 hours ago

Senior Data Engineer – GCP, Python, Iceberg, Delta Lake, Kafka, Snowflake, Databricks

US flagUnited States OnlyFull-timeData Engineer$120k – $180k/year
ApplyView job
Livefront6 hours ago

Data Engineer

PE flagPeru OnlyFull-timeData Engineer
ApplyView job
GFT Technologies6 hours ago

Data Engineer, Mid-level

BR flagBrazil OnlyFull-timeData Engineer
ApplyView job
VIDA7 hours ago

Geospatial Data Engineer – Customer & AI Solutions

DE flagGermany OnlyFull-timeData Engineer
ApplyView job
albo7 hours ago

Data Engineer

MX flagMexico OnlyFull-timeData Engineer
ApplyView job
Leega7 hours ago

Engenheiro de Dados Pleno – AWS

BR flagBrazil OnlyFreelanceData Engineer
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers