
Technical Leader
Posted Sep 8

Posted Sep 8
This is a fully remote position, open to applicants in Canada.
β’ Design and oversee high-throughput, fault-tolerant distributed systems that ingest and process metrics, logs, and traces at scale.
β’ Define and lead the technical roadmap focused on platform reliability, scalability, and operational efficiency.
β’ Collaborate with product and engineering leadership to convert business requirements into practical technical designs.
β’ Establish and promote best practices and standards for observability and capacity planning.
β’ Engage in incident response and facilitate blameless post-incident reviews.
β’ Mentor and guide senior engineers through design and code reviews.
β’ Maintain ownership of systems from design to deployment, including monitoring and post-incident analysis.
β’ Influence technical direction and architectural decisions from prototype to production.
β’ Bachelorβs degree or higher in Computer Science, Electrical Engineering, or a related technical discipline.
β’ Over 10 years of experience in building and operating large-scale distributed systems in production environments.
β’ Strong proficiency in programming languages such as Golang.
β’ Familiarity with modern observability platforms and designing pipelines for logs, metrics, and traces.
β’ Experience in designing observability solutions for large systems.
β’ Proficiency with tools like Splunk, Datadog, Prometheus, Grafana, or similar technologies.
β’ Experience with AWS, GCP, or Azure at a platform level, including networking, storage, and cost optimization strategies.
β’ Knowledge of full-lifecycle agentic development.
β’ Understanding of context and harness engineering best practices.
β’ Proven technical leadership and cross-functional influence at the Staff or Principal level.
β’ Experience in driving multi-team technical decisions, producing architecture proposals, mentoring senior engineers, and aligning technical strategy with business objectives.
β’ Preferred: demonstrated ownership of high-volume data-stream systems.
β’ Preferred: experience in capacity planning, traffic management, SLI/SLO definitions, and large-scale incident response.
β’ Preferred: proficiency in C++.
β’ Preferred: experience with Docker/OCI containers and Kubernetes at scale.
β’ Preferred: expertise in security-first system design.
β’ Preferred: operational experience, including on-call responsibilities and service ownership.
β’ Medical, dental, and vision insurance.
β’ 401(k) plan with a Cisco matching contribution.
β’ Paid parental leave.
β’ Short- and long-term disability coverage.
β’ Basic life insurance.
β’ Cisco restricted stock unit grants may be available, subject to vesting.
β’ 10 paid holidays per full calendar year.
β’ 1 floating holiday for non-exempt employees.
β’ Paid employee birthday off.
β’ Paid year-end holiday shutdown.
β’ 4 paid days off for personal wellness.
β’ 16 days of paid vacation per full calendar year for non-exempt employees.
β’ Flexible vacation time off program with no defined limit for eligible exempt employees.
β’ 80 hours of sick time off provided on hire date and each January 1st thereafter.
β’ Up to 80 hours of unused sick time carried forward.
β’ Additional paid time away for critical or emergency family issues.
β’ Optional 10 paid volunteer days per full calendar year.
β’ Annual bonuses for non-sales roles, subject to Cisco policies.
LMI
Sourcegraph
Verra Mobility
Cint
Get handpicked remote jobs straight to your inbox weekly.