About this role
Serves as a forward-deployed technical liaison to NVIDIA’s Partner Network, customers, and internal teams to deploy, manage, and validate large-scale AI Compute/HPC infrastructure in Linux-based environments. Responsibilities include system design, automation, validation, documentation, and providing feedback to improve reference architectures and partner enablement.
Skills for this role
Linux system administrationProcess managementPackage managementTask schedulingKernel managementBoot proceduresPerformance reportingPerformance optimizationPerformance loggingNetwork routingAdvanced networkingCluster managementBare-metal server provisioningBCM (Base Command Manager)BashPythonAnsibleSLURMLSFUGEHPLNCCLMLPerfKubernetesInfiniBandGPU hardwareMPILustreGPFSOEM GPU platforms familiarity""Channel sales