
Senior Data Engineer
Posted 10 hours ago

Posted 10 hours ago
This is a fully remote position, open to applicants in New York, +3 more states.
• Design, develop, and maintain a near real-time transactional cache platform that supports APIs, applications, and digital experiences.
• Build and oversee CDC pipelines that replicate data from UniVerse to PostgreSQL using vendor-approved replication technologies.
• Create and manage transformation pipelines that convert operational relational data into optimized MongoDB document formats.
• Implement and maintain MongoDB document models, indexing strategies, partitioning methods, and query patterns for high-throughput, low-latency API access.
• Collaborate with API and application development teams to define and deliver domain models that are ready for caching.
• Monitor and enhance data freshness, latency, availability, and reliability of transactional cache workloads.
• Optimize PostgreSQL and MongoDB performance through effective schema design, indexing, query tuning, and capacity planning.
• Diagnose production issues and engage in operational support, incident response, and root cause analysis.
• Develop and maintain operational runbooks, monitoring dashboards, alerting strategies, and support procedures.
• Work alongside architects, senior engineers, and cross-functional teams on scalable data platform solutions.
• Contribute to engineering best practices, code reviews, documentation, and continuous improvement initiatives.
• Mentor junior engineers and share expertise related to CDC, data pipelines, and cloud-native data engineering.
• 5+ years of experience in data engineering, software engineering, or data platform development.
• 3+ years of experience in designing, developing, and supporting enterprise-scale data platforms and distributed systems.
• Proficient in MongoDB Atlas/PostgreSQL, including data modeling, aggregation frameworks, index design, and performance tuning.
• Familiar with Change Data Capture (CDC), replication technologies, and near real-time data movement patterns.
• Experience in designing and supporting operational data stores, transactional caches, or data platforms that serve APIs.
• Strong development skills in Python and SQL.
• Experience with GCP services, including Cloud SQL, Dataproc Serverless, Cloud Composer (Airflow), Monitoring, IAM, and cloud networking fundamentals.
• Proven experience in developing and supporting production ETL/ELT pipelines.
• Experience working in Linux/Unix environments.
• Ability to troubleshoot issues related to distributed data processing and data synchronization.
• Strong problem-solving, communication, and teamwork skills.
• Experience in implementing transactional cache or operational data platform architectures.
• Familiarity with Kafka, Pub/Sub, or event-driven architecture.
• Experience in building data platforms that support APIs and microservices.
• Knowledge of Terraform or Infrastructure as Code.
• Experience with GKE/Kubernetes.
• Proven ability to design highly available and low-latency data platforms.
• Agile/SAFe experience in large enterprises.
• Experience in the healthcare industry is a plus.
• Familiarity with AWS is beneficial.
• Cloud certifications (AWS and/or GCP) are preferred.
• Preference for EST coast working hours.
• Bachelor’s degree or equivalent experience (High School diploma + 4 years of relevant experience).
• CVS Health bonus, commission, or short-term incentive program in addition to base salary.
• Medical insurance coverage.
• Dental insurance coverage.
• Vision insurance coverage.
• Paid time off.
• Retirement savings options.
• Wellness programs.
• Additional resources supporting physical, emotional, and financial well-being, subject to eligibility.
The Home Depot
Genesys
CareMore Health
GFT Technologies
Get handpicked remote jobs straight to your inbox weekly.