Remotery

Lead Data Engineer – Contract, Full-Time

Posted 1 day ago

This is a fully remote position, open to applicants in Pakistan.

📋 Description

• Design and construct scalable data pipelines and infrastructure that support AI and product systems.

• Create and oversee data ingestion, transformation, and storage architectures for both operational and AI workloads.

• Develop and manage both batch and real-time data pipelines.

• Construct and enhance vector search, retrieval, and machine learning data pipeline systems.

• Ensure the reliability, security, and governance of data across the platform.

• Collaborate with AI and backend engineering teams to facilitate training, inference, and product features.

• Implement frameworks for monitoring, observability, and data quality.

• Optimize performance for large-scale datasets and query systems.

• Contribute to decisions regarding technical architecture and long-term data strategy.

• Serve as the initial data hire by establishing culture, standards, and hiring benchmarks for the data function.

• Collaborate with founders and product leadership to translate data capabilities into product decisions.


⛳️ Requirements

• Over 7 years of professional experience, primarily in dedicated data engineering roles.

• Extensive experience in designing and constructing data pipelines and distributed data systems.

• Proficient with relational databases; PostgreSQL is preferred, while MySQL or similar databases are also acceptable.

• Familiarity with NoSQL databases.

• Knowledge of vector databases utilized in modern AI systems.

• Strong programming proficiency in Python.

• Capability to make and justify architectural choices.

• Experience in building scalable backend systems.

• Expertise in designing data models and storage architectures.

• Solid understanding of data processing performance and optimization.

• Highly desirable: Experience with Apache Spark, Apache Airflow, Kafka, Elasticsearch, or OpenSearch.

• Highly desirable: Knowledge of PostgreSQL, MongoDB, Qdrant, Milvus, or pgvector.

• Highly desirable: Familiarity with Pandas or Polars.

• Nice to have: Experience with AI or machine learning platforms.

• Nice to have: Familiarity with stream processing and event-driven architectures.

• Nice to have: Experience with cloud infrastructure such as GCP, AWS, or Azure.

• Nice to have: Background in high-growth startups or early-stage companies.


🏝️ Benefits

• Remote-first work culture.

• A genuine community focused on growth and well-being.

• Full-time, long-term employment opportunities.

• Access to exceptional global teams and products.

• Opportunities for professional and personal development.

People also viewed

Railroad1910 hours ago

Senior Data Engineer – GCP, Python, Iceberg, Delta Lake, Kafka, Snowflake, Databricks

US flagUnited States OnlyFull-timeData Engineer$120k – $180k/year
ApplyView job
Livefront10 hours ago

Data Engineer

PE flagPeru OnlyFull-timeData Engineer
ApplyView job
GFT Technologies10 hours ago

Data Engineer, Mid-level

BR flagBrazil OnlyFull-timeData Engineer
ApplyView job
VIDA11 hours ago

Geospatial Data Engineer – Customer & AI Solutions

DE flagGermany OnlyFull-timeData Engineer
ApplyView job
albo12 hours ago

Data Engineer

MX flagMexico OnlyFull-timeData Engineer
ApplyView job
Leega12 hours ago

Engenheiro de Dados Pleno – AWS

BR flagBrazil OnlyFreelanceData Engineer
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers