Technical Success Engineer
AI summary of the role
This role owns taking signed deployments from contract to live production for Lambda's largest, most strategic customers in the Superintelligence business unit.
What you’ll do
- Take signed deployments from contract to live production, validating configuration, connectivity, storage, and compute against promises
- Bring technical depth to validation and troubleshooting, working with engineering and infrastructure teams to resolve issues
- Coordinate with Infrastructure, Engineering, Product, and Data Center teams to close technical dependencies and resolve blockers
- Own a current, accurate technical picture of the deployment and keep stakeholders informed with RAG status and top risks
What you’ll bring
- 4+ years of hands-on technical experience with GPU/HPC infrastructure, cloud platforms, Kubernetes, or large-scale Linux systems
- Comfortable being the technical voice in the room, able to validate builds and hold engineering teams accountable
- Strong troubleshooting instincts in networking, storage, or compute issues
- Track record of coordinating across engineering and infrastructure teams to close out technical dependencies
Technologies
GPU · HPC · Kubernetes · Linux · cloud platforms · networking · storage · compute
About Lambda
GPU cloud serving AI researchers, frontier labs and hyperscalers with on-demand and reserved NVIDIA clusters for training and inference.
Series E · 500–1000 people
Employer postings · Data from · Sources