
Technical Leader
Posted Sep 8

Posted Sep 8
This is a fully remote position, open to applicants in Canada.
• Oversee the design and development of extensive data plane systems that support Splunk's data ingestion infrastructure.
• Architect and spearhead high-throughput, fault-tolerant distributed systems responsible for ingesting and processing metrics, logs, and traces.
• Define and steer the technical roadmap aimed at enhancing platform reliability, scalability, and operational efficiency.
• Collaborate with product and engineering leadership to convert business needs into practical technical designs.
• Establish best practices and standards for observability and capacity planning.
• Engage in incident response and lead blameless post-incident evaluations.
• Mentor and cultivate senior engineers through design and code reviews.
• Take ownership of systems from design through deployment, monitoring, and post-incident analysis.
• Drive technical decisions across multiple teams and influence architecture from prototype to production.
• A Bachelor’s degree or higher in Computer Science, Electrical Engineering, or a related technical discipline.
• Over 10 years of experience in building and managing large-scale distributed systems in production environments.
• Strong proficiency in programming languages such as Golang.
• Experience with contemporary observability platforms and designing pipelines for logs, metrics, and traces.
• Experience in designing observability for large systems, rather than merely utilizing it.
• Familiarity with AWS, GCP, or Azure at a platform level, including aspects of networking, storage, and cost optimization.
• Background in full-lifecycle agentic development.
• Knowledge of context and harness engineering best practices.
• Experience in technical leadership and cross-functional influence at the Staff or Principal level.
• Proven track record of driving multi-team technical decisions, generating architecture proposals, mentoring senior engineers, and aligning technical strategies with business objectives.
• Demonstrated experience with capacity planning, traffic management, SLI/SLO definitions, and incident response on a large scale.
• Proficiency in C++.
• Practical experience with Docker/OCI containers and Kubernetes at scale.
• Experience in designing systems with security-first principles.
• Excellent written and verbal communication abilities.
• Capability to discern when to prototype swiftly and when to apply thoroughness.
• Medical, dental, and vision insurance.
• 401(k) plan with a Cisco matching contribution.
• Paid parental leave.
• Short- and long-term disability coverage.
• Basic life insurance.
• Cisco restricted stock unit grants may be available, vesting after continued employment.
• 10 paid holidays per full calendar year.
• 1 floating holiday for non-exempt employees.
• Paid employee birthday.
• Paid year-end holiday shutdown.
• 4 paid days off for personal wellness.
• 16 days of paid vacation per full calendar year for non-exempt employees.
• Flexible vacation time off program with no defined limit for eligible exempt employees.
• 80 hours of sick time off provided on hire date and each January 1st thereafter.
• Up to 80 hours of unused sick time carried forward annually.
• Additional paid time off for critical or emergency family issues.
• Optional 10 paid volunteer days per full calendar year.
• Annual bonuses for non-sales roles, subject to Cisco policies.
• Performance-based incentive pay for sales-plan employees.
LMI
Sourcegraph
Verra Mobility
Cint
Get handpicked remote jobs straight to your inbox weekly.