NCX Senior Engineer
AI summary of the role
This role is a senior customer-facing engineer on NVIDIA's AI Accelerator team, responsible for deploying and optimizing AI workloads on NVIDIA Cloud Platforms (NCP, Neo, DGX Cloud) for strategic partners.
What you’ll do
- Build and deploy custom AI solutions on NCP and Neo Cloud platforms, including distributed training, inference optimization, and MLOps pipelines.
- Act as the main technical contact for strategic NCPs, offering remote and on-site support and troubleshooting complex production problems.
- Deploy and manage AI workloads across DGX Cloud, NCP data centers, and major CSP environments using Kubernetes, containers, and GPU scheduling.
- Profile and tune large-scale training and inference workloads, implementing observability and SLO/SLA monitoring to reduce latency, cost, and risk.
What you’ll bring
- BS/MS/PhD in CS, CE, EE, or equivalent experience.
- 8+ years in customer-facing technical roles (Solutions Engineering, DevOps, SRE, ML Infra) supporting large-scale cloud or service provider environments.
- Strong expertise in Linux, distributed computing, Kubernetes, containers, and GPU scheduling on multi-tenant platforms.
- Demonstrated AI/ML experience supporting large-scale training and inference workloads (LLMs, generative models, recommendation systems) in production.
Technologies
Kubernetes · Docker · PyTorch · TensorFlow · CUDA · NeMo · Triton · NIM · InfiniBand · RoCE · Prometheus · Grafana
About NVIDIA
Designs and manufactures GPUs and system-on-chips powering data centers, AI workloads, gaming, autonomous vehicles, and HPC. The foundational hardware for modern deep learning.
Public
Employer postings · Data from · Sources