
Tools Development Engineer
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in North Carolina.
• Establish and document serviceability requirements for advanced non-x86 server platforms, encompassing AI, HPC, and GPU-accelerated systems.
• Create, sustain, and enhance serviceability requirements, technical specifications, design guidance, engineering documentation, and action plans.
• Provide support for hardware diagnostics, fault isolation, repair procedures, telemetry, event reporting, and service verification.
• Assess platform architectures, product designs, and engineering proposals to pinpoint serviceability gaps, risks, ambiguities, and conflicting requirements.
• Foster cross-functional alignment to resolve serviceability issues.
• Convert service and problem-determination needs into quantifiable and testable product requirements.
• Ensure traceability between service use cases, product requirements, diagnostic capabilities, design decisions, and published service information.
• Engage in architecture reviews, design reviews, readiness assessments, product readiness reviews, failure analysis activities, and technical discussions.
• Evaluate, create, update, consolidate, archive, or retire serviceability documentation.
• Assist Information Development and technical publications teams with technical content, engineering reviews, and subject matter expertise.
• Independently oversee assigned projects, dependencies, priorities, and deliverables.
• Bachelor's degree in Computer Engineering, Electrical Engineering, Computer Science, or a related technical field.
• At least three (3) years of experience in developing technical requirements, engineering specifications, architecture documentation, or similar technical content.
• Proven experience in defining product requirements related to hardware serviceability, diagnostics, technical support, repair processes, or related engineering functions.
• Familiarity with server hardware architecture and major subsystems, including processors, memory, storage, power, cooling, firmware, management controllers, and system interconnect technologies.
• Capability to comprehend complex technical designs and convey technical concepts to engineering, documentation, and cross-functional stakeholders.
• Experience with enterprise server technologies, data center infrastructure, rack-scale systems, or similar platforms.
• Background in serviceability, diagnostics, fault isolation, field-replaceable unit identification, or problem determination workflows.
• Acquainted with BMCs, Redfish, system event logs, telemetry, firmware diagnostics, hardware health monitoring, out-of-band management, DOORS, Jira, Confluence, Git, or similar platforms.
• Experience in participating in architecture reviews, design reviews, product readiness reviews, failure analysis activities, or comparable engineering governance processes.
• Understanding of Design for Serviceability (DFS), Design for Supportability (DFSu), Design for Reliability (DFR), or similar methodologies.
• Ability to work independently, manage ambiguity, take initiative, and drive cross-functional technical issues to resolution.
• Strong skills in organization, technical writing, document management, and content lifecycle management.
• Experience collaborating with geographically distributed engineering, development, support, and technical publications teams.
• Preferred experience in supporting ARM-based, AI, GPU-accelerated, high-performance computing (HPC), or other non-x86 server platforms.
• Equal Opportunity Employer.
• Reasonable accommodation available to complete the application.
Mercor
RTX
Expel
Qualus
Get handpicked remote jobs straight to your inbox weekly.