
Site Reliability Engineer – Networking
Posted Jul 28

Posted Jul 28
This is a fully remote position, open to applicants in California, +2 more states.
• Provide support for a specific, highly available, and secure production environment.
• Evaluate the reliability of the environment and enhance operational practices through networking and coding expertise.
• Diagnose complex events as part of a 24/7 on-call rotation.
• Design network architecture for hybrid cloud implementations.
• Automate the connectivity of endpoint networks.
• Abstract network functions from hardware to develop connectivity solutions.
• Create comprehensive monitoring tools to assess the performance and reliability of network infrastructure.
• Resolve production issues across application, system, and network layers through root-cause analysis.
• Implement preventative strategies to mitigate production issues.
• Design, implement, and manage an overlay network to support thousands of containers.
• Collaborate with other engineering teams to devise complex solutions.
• Bachelor’s degree with 7 years of relevant experience, or a Master’s degree with 4 years, or a PhD with 1 year.
• Proven experience in designing, deploying, and managing mid- to large-scale enterprise or cloud environments.
• Proficiency in scripting or coding with Python, Bash, Ruby, or Go.
• Familiarity with *nix systems.
• Experience in supporting externally facing production environments.
• Knowledge of BGP, OSPF, IPv6, network security, DMVPN, and MACSec.
• Experience with monitoring tools such as Grafana and Splunk.
• Advanced skills in Unix/Linux and system administration.
• Scripting experience in Bash, Ruby, or Python.
• Demonstrated project management experience.
• Experience with AWS or Azure platforms.
• Familiarity with Docker and Kubernetes.
• Exposure to Ansible, REST APIs, and TLS, or experience in cloud/ISP/telco environments.
• Comfort in utilizing AI-assisted engineering tools such as Codex, Claude Code, and Cursor.
• Ability to comprehend agent-based workflows, reusable skills, and MCP integrations.
• Skill in critically evaluating AI-generated outputs for accuracy, security, reliability, and maintainability.
• Cisco networking certifications like CCNP, CCIE, JNCP, or JNCIE are preferred qualifications.
• Medical, dental, and vision insurance.
• 401(k) plan with Cisco matching contributions.
• Paid parental leave.
• Coverage for short- and long-term disability.
• Basic life insurance.
• Availability of Cisco restricted stock unit grants.
• 10 paid holidays per full calendar year.
• 1 floating holiday for non-exempt employees.
• 1 paid day off for the employee’s birthday.
• Paid holiday shutdown at year-end.
• 4 paid days off for personal wellness.
• 16 days of paid vacation per full calendar year for non-exempt employees.
• Flexible vacation time off program with no defined limit for eligible exempt employees.
• 80 hours of sick time off provided upon hire and annually, with the option to carry forward up to 80 unused hours.
• Additional paid time off for critical or emergency family matters.
• Optional 10 paid volunteer days per full calendar year.
• Annual bonuses for non-sales roles, subject to Cisco policies.
• Sales incentive compensation for employees on sales plans.
• Opportunities for growth, experimentation, and learning.
• A global network of thinkers, doers, experts, and creators.
DATAGROUP
Ambush
DuoKey
TEKsystems
Get handpicked remote jobs straight to your inbox weekly.