
ClickHouse Engineer
Posted Jul 28

Posted Jul 28
This is a fully remote position, open to applicants in United Kingdom.
β’ Oversee cluster operations on a large scale β managing, upgrading, and enhancing a multi-AZ ClickHouse infrastructure on Kubernetes, with established backup and recovery procedures.
β’ Schema and query design β focusing on table and sorting key architecture, partitioning strategies, materialized views, and optimizing queries for multi-terabyte datasets.
β’ Capacity planning and observability β implementing a comprehensive capacity model, storage tiering, data retention strategies, and utilizing Grafana for monitoring and alerting to ensure system health visibility.
β’ Data ingestion and quality assurance β establishing correctness and completeness checks on the data pipelines supplying the system; conducting reconciliations to ensure no data is lost.
β’ Workload segregation β ensuring that ingestion, operations, reporting, and ad-hoc analytics do not interfere with one another as read loads increase.
β’ AI-enhanced operations β leveraging modern tools (including MCP-based AI access to the infrastructure) that enable a small team to efficiently manage a large platform.
β’ Expertise in columnar/OLAP database engineering β ClickHouse is highly preferred; extensive experience with another columnar or large-scale time-series database (such as BigQuery, Redshift, Druid, kdb+, or similar) is also acceptable.
β’ Proficiency in SQL β including execution plans, query optimization, and schema design for very large datasets.
β’ Experience in operations/SRE β having taken on production responsibilities for a data platform, including monitoring, capacity management, incident resolution, and recovery processes.
β’ Familiarity with Linux and Kubernetes β comfortable working with the operational layers underlying the database.
β’ Knowledge of a programming language β preferably Python, but familiarity with Go, R, or MATLAB is also acceptable.
β’ Engineering rigor β adept at writing proposals, strategies, and architectural documents and diagrams, and effectively managing changes.
β’ Proficient in Git β demonstrating excellent version control practices.
β’ Understanding of distributed systems fundamentals β including replication, consistency, and failure modes.
β’ Experience in financial services is beneficial but not mandatory β domain knowledge can be acquired; operational intuition is essential.
β’ A substantial infrastructure β genuine scale and critical importance: serving as the system of record for a live trading network rather than merely a reporting tool.
β’ True ownership β sharing responsibility for a core business-critical platform from the outset, collaborating directly with the platform's lead engineer.
β’ A remarkable team β senior engineers who are passionate about their craft, fostering open collaboration, and maintaining high standards.
β’ Sustainable operations β designed for follow-the-sun support; on-call responsibilities are shared, contracted, and planned rather than relying on heroic efforts.
β’ Competitive compensation package β aligned with your local market, offering competitive pay and benefits.
β’ Access to modern tools β AI-assisted operations and investigative tools that minimize toil rather than adding unnecessary processes.
β’ Opportunities for growth β as the estate scales to multiple times its current volume, this role will expand alongside it.
LiteLLM AI Gateway
Snowflake
RTX
C-MORE
Get handpicked remote jobs straight to your inbox weekly.