
Senior Engineer, System Software, SDN Operations
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in California.
• Create and implement next-generation multi-tenant cloud SDN control and data plane software utilizing OVS, OVN, and OpenFlow.
• Develop Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST for BMaaS, VMaaS, and Kubernetes.
• Lead upstream contributions to OVN-Kubernetes and similar open-source initiatives.
• Design and develop network observability tools for monitoring, telemetry, intelligent metering, and performance analysis.
• Manage and support OVS-OVN-based SDN solutions within large-scale NVIDIA AI Cloud environments.
• Take ownership of end-to-end SDN observability, incorporating monitoring, alerting, distributed tracing, and dashboarding.
• Construct, improve, and maintain GitLab CI/CD pipelines across Linux host networking, OVS, OVN, and Kubernetes CNIs.
• Apply GitOps methodologies and securely integrate with cloud infrastructure.
• Enhance reliability through incident management, resource monitoring, and performance optimization.
• Collaborate with SRE, DevOps, and network engineering teams to ensure production readiness and develop operational tools.
• BS/MS in Computer Science or a related technical discipline, or equivalent professional experience.
• Over 8 years of demonstrated software development experience in large-scale distributed systems.
• In-depth expertise in OVN, OVS, OpenFlow, and contemporary network protocols.
• Proficient programming skills in C and Go.
• Advanced scripting capabilities in Bash and Python.
• Comprehensive understanding of Kubernetes and practical experience in deploying and supporting CNIs, particularly OVN-Kubernetes.
• Hands-on experience with Infrastructure-as-Code and deployment tools such as Ansible, Terraform, ArgoCD, and Flux.
• Experience in designing and managing complex, multi-stage CI/CD pipelines.
• Practical knowledge in developing secure, high-performance gRPC and REST services with TLS and robust authentication.
• Strong understanding of datacenter routing, switching, and Linux host/VM networking.
• Contributions to open-source projects, especially OVS, OVN, OVN-Kubernetes, or other Kubernetes networking initiatives.
• Familiarity with hardware acceleration technologies such as GPU or DPU for networking.
• Practical experience with AWS, Azure, GCP, and hybrid/multi-cloud deployments.
• Expertise in SRE/DevOps, including on-call responsibilities, incident management, reliability targets, and production ownership.
• Experience with observability platforms and tools such as Prometheus, Grafana, Jaeger, OpenTelemetry, and ELK.
• Competitive salary package.
• Equity options.
• Comprehensive benefits.
• Commitment to an equal opportunity and inclusive work environment.
OpenText
Akamai Technologies
Professional Physical Therapy
Get handpicked remote jobs straight to your inbox weekly.