
Data Engineer
Posted 4 days ago

Posted 4 days ago
This is a fully remote position, open to applicants in United States, +1 more country.
• Take charge of the reliability, freshness, and scalability of ingestion and orchestration pipelines that bring source data into Snowflake.
• Create monitoring systems, alerting mechanisms, runbooks, and patterns for backfilling and reprocessing data.
• Design, construct, and enhance Bronze, Silver, and Gold dbt models for billing, usage, CRM, product events, and GTM funnel reporting.
• Collaborate with Product, Finance, RevOps, and GTM teams to define metrics and document them in the data warehouse.
• Identify and resolve data quality and freshness issues at their origin.
• Manage Snowflake RBAC, dynamic masking, and PII classification for individuals, agents, and service accounts.
• Evaluate DDL and access requests from Engineering and GTM departments.
• Oversee reverse-ETL systems and runbooks that deliver warehouse data to Salesforce, Slack, and internal agents.
• Enhance semantic views, context, and evaluations for precise agent responses.
• Broaden agent-driven workflows from runbooks and data ingestion to pull requests.
• Develop and improve CI/CD review gates within the data-platform monorepo.
• Manage compute resources, deployments, secrets, access patterns, and environments for the data platform.
• Collaborate with infrastructure engineers to ensure availability and address failure modes.
• Standardize Postgres and SaaS ingestion while consolidating pipeline deployment through Prefect.
• Handle Snowflake RBAC, resource management, and masking policies as code using Terraform.
• Automate masking coverage, Identity Graph matching, transcript aggregation, PII detection, reverse-ETL frameworks, and runbooks.
• Operate self-hosted data services on Kubernetes and manage AWS resources as code with Terraform.
• Over 5 years (or equivalent) of experience in building and managing production data platforms, encompassing the transformation layer.
• Extensive experience with Snowflake, dbt, and an orchestrator such as Prefect, Airflow, or Dagster.
• Familiarity with data ingestion from production databases and SaaS sources, including backfilling and reprocessing methodologies.
• Proficient in monitoring, alerting, debugging, incident response strategies, and practical SLO/SLA considerations.
• Experience with data warehouse access governance: RBAC, masking policies, and PII management.
• Strong judgment in data modeling and an understanding of schema evolution.
• Proficient in SQL and Python.
• Adherence to disciplined engineering practices including testing, documentation, reviews, and CI/CD.
• Proven track record of functioning as the sole engineer on a specific layer.
• Familiarity with GTM, Finance, and Product operations.
• Ability to translate ambiguous business inquiries into technical specifications.
• Comfort in utilizing LLMs and coding agents in daily development tasks.
• Capability to assess AI-generated outputs and integrate them into standardized processes.
• Systems thinking regarding freshness, accuracy, and failure scenarios.
• A pragmatic approach to constructing and scaling solutions.
• Ability to leverage AI and automation to enhance platform operations.
• Healthcare, dental, and vision coverage.
• Parental leave.
• Paid time off.
• Fully remote working options.
• 401k matching.
• Competitive equity opportunities.
• FSA, short-term/long-term disability, and voluntary life insurance.
• Carrot fertility benefits.
• 20 days of paid vacation, plus 10 holidays, and unlimited sick leave.
• 12 weeks of fully paid parental leave.
• Monthly fitness stipend.
• Monthly wellness stipend.
• Commuter benefits for hybrid employees in SF/NYC.
• Unlimited token usage.
CuraLinc Healthcare
VSP Vision Care
Adoreal
Get handpicked remote jobs straight to your inbox weekly.