
Senior Solutions Architect – Customer Success, Partnership
Posted 11 hours ago

Posted 11 hours ago
This is a fully remote position, open to applicants in Poland, +3 more countries.
• Lead hands-on analysis, optimization, and performance tuning of intricate GPU-accelerated systems and AI workloads.
• Ensure high availability and operational efficiency across customer data centers.
• Act as a senior technical authority on NVIDIA technologies.
• Contribute to architecture reviews and guide large-scale infrastructure decisions.
• Establish and enhance monitoring and optimization methodologies through analytics, telemetry, and automation.
• Identify bottlenecks and enhance infrastructure resiliency.
• Participate in post-deployment assessments and incident retrospectives.
• Provide insights into NVIDIA’s infrastructure strategy and help shape the customer experience.
• Manage and lead complex technical projects from initial design through implementation and ongoing improvement.
• Ensure compliance with SLAs and address technical risks effectively.
• Recognize AI infrastructure opportunities within cloud and enterprise environments.
• Propel technical initiatives that highlight NVIDIA’s leadership.
• Collaborate with customers, partners, and internal teams on extensive Networking, System Design, and Automation projects.
• Over 10 years of experience in large-scale data center service operations with an emphasis on infrastructure.
• BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related disciplines.
• Strong analytical, problem-solving, and decision-making capabilities.
• Excellent communication, time management, and organizational skills.
• Preferred certifications in data center, server, or networking technologies.
• Proficient in system-level aspects, including Operating Systems, Linux kernel drivers, GPUs, NICs, and hardware architecture.
• Expertise in cloud orchestration software and job schedulers, such as Kubernetes, Docker Swarm, and Slurm.
• Familiarity with cloud-native technologies and their integration with traditional infrastructure.
• Extensive knowledge of AI infrastructure and workflows, including training/inference pipelines, MLOps/DevOps tools, containerization, and large-scale system deployments.
• Understanding of data center infrastructure operations, including safety, security, environmental controls, and standard operating procedures.
• Capability to lead discussions, influence outcomes, and foster positive relationships with both internal and external collaborators.
• Willingness to travel up to 25% for customer engagements and team collaboration.
Cisco
Gateway Ticketing Systems UK Ltd
MoneyGram
TCP Software
Get handpicked remote jobs straight to your inbox weekly.