
Lead Data Center NOC Engineer, L2
Posted 6 days ago

Posted 6 days ago
This is a fully remote position, open to applicants in California.
• Serve as the senior technical leader during shifts, guiding and assisting L1 NOC technicians
• Act as the main escalation point for incidents until they are resolved or handed off to L3 Engineering
• Oversee shift transitions, prioritize tasks, and make operational decisions
• Provide mentorship and coaching to junior NOC staff and assist with onboarding
• Represent the NOC in cross-functional incident bridge calls
• Function as Incident Commander for medium to high-severity incidents
• Manage the complete incident lifecycle: detection, triage, mitigation, communication, and resolution
• Conduct root cause analysis and implement corrective measures for recurring issues
• Ensure high-quality, executive-level incident communications
• Strive to continuously enhance MTTR and operational reliability
• Monitor and troubleshoot physical infrastructure utilizing PLC, BMS, and DCIM platforms
• Conduct health assessments on UPS, PDUs, generators, and CRAC/CRAH systems
• Address MEP-related challenges and coordinate maintenance and remote hands support
• Assist with modular, containerized, and micro data center deployments
• Enforce physical and logical security protocols
• Monitor and troubleshoot switches, routers, firewalls, and edge connectivity
• Execute L2/L3 troubleshooting including VLANs, IP addressing, MTU, and routing basics
• Validate fiber/copper connectivity, optics, link status, and redundancy paths
• Troubleshoot VPNs and secure remote access solutions
• Escalate architectural or design concerns to Network Engineering (L3)
• Utilize observability and ITSM platforms such as Grafana, Zenduty, ServiceNow, Jira, and SolarWinds
• Maintain operational runbooks, SOPs, and escalation protocols
• Enhance alerting quality, dashboards, and reduce noise
• Support automation, reporting, audits, and compliance inquiries
• Collaborate with Engineering and Product teams
• Engage in change planning, execution, and post-change validation
• Identify operational deficiencies and drive continuous improvement efforts
• Over 7 years of experience in NOC, data center, or infrastructure operations
• Proven experience in leading incidents or technical shifts
• Extensive hands-on experience with DCIM/BMS platforms (Distech, Schneider, RadixIOT)
• Advanced expertise in L2/L3 networking troubleshooting
• Strong knowledge of power, cooling, and monitoring systems
• Ability to read and interpret electrical one-lines and network diagrams
• Comfortable working shifts and participating in on-call rotations
• Preferred certifications include CCNA, JNCIA, CDCTP, CDCP, or equivalent
• Desired experience with edge/remote connectivity (VSAT, LTE/5G, Starlink)
• Familiarity with Linux/Windows server environments is preferred
• Basic scripting knowledge (Python, Bash, PowerShell) is preferred
• Competitive base salary and equity
• Medical, dental, and vision benefits with subsidized costs
• Health savings accounts (HSA), flexible spending accounts (FSA), and dependent care FSAs (DCFSA)
• Retirement plan options, including 401(k) and Roth 401(k)
• Unlimited paid time off (PTO)
• 14 paid company holidays each year
• Subsidized benefits
OnHires
Humana
AXON Networks
DYOPATH
Get handpicked remote jobs straight to your inbox weekly.