Solutions Architect - US
AI summary of the role
FuriosaAI seeks a US-based Solutions Architect to own end-to-end technical enablement for customers deploying AI models on its RNGD NPU using the Furiosa SDK.
What you’ll do
- Own end-to-end technical enablement for US customers deploying AI models on RNGD NPU using Furiosa SDK
- Develop POCs, benchmarking studies, and live debugging sessions in customer environments
- Act as technical authority to US BD/Sales during pre-sales and enterprise evaluations
- Onboard and train customers on integration patterns and optimization workflows post-purchase
What you’ll bring
- 2–5 years in a US customer-facing technical role (Solutions Architect, Sales Engineer, Forward Deployed Engineer) at AI infra, cloud, or semiconductor company
- Hands-on experience with modern inference stacks: vLLM, SGLang, TensorRT-LLM, Triton Inference Server, or similar
- Hands-on experience with agent/orchestration frameworks: LangChain, LlamaIndex, LangGraph, AutoGen, or MCP-based tooling
- Proficiency in Python; comfortable with PyTorch or TensorFlow
Technologies
RNGD · Furiosa SDK · vLLM · SGLang · TensorRT-LLM · Triton Inference Server · LangChain · LlamaIndex · LangGraph · AutoGen · MCP · Python
About FuriosaAI
Designs AI inference chips, servers, and compiler/runtime software for enterprises and cloud providers running LLM and multimodal workloads in standard air-cooled data centers.
Growth
Source and classification
Technical pre-sales · Evidence for this classification:
About FuriosaAI FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence for every enterprise. About the Role FuriosaAI is looking for a Solutions Architect to bring the full potential of our powerful RNGD chips/servers to our customers by acting as the primary technical authority in AI/LLM model deployments. From running POCs to
More from the job description
About FuriosaAI FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence for every enterprise. About the Role FuriosaAI is looking for a Solutions Architect to bring the full potential of our powerful RNGD chips/servers to our customers by acting as the primary technical authority in AI/LLM model deployments. From running POCs to benchmarking and debugging, you will translate RNGD’s powerful system to real-world deployments of customers’ models, empowering customers with FuriosaAI’s powerful solutions. If you are interested in providing the technical expertise in challenging the current status-quo of AI infrastructure in real-world environments, join us in our path to a sustainable future of AI. Key Responsibilities Own end-to-end technical enablement for US customers deploying AI models on FuriosaAI's RNGD NPU using the Furiosa [... source excerpt omitted ...] ue for engineering and C-suite audiences Develop deep, current expertise in FuriosaAI's hardware and software stack and demonstrate it at US technical forums, AI conferences, and customer workshops Onboard and train customers on integration patterns, optimization workflows, and best practices post-purchase Serve as a technical feedback loop from US customers back to Seoul HQ product and engineering teams Minimum Qualifications 2–5 years in a US customer-facing technical role: Solutions Architect, Sales Engineer, Forward Deployed Engineer, or equivalent at an AI infra, cloud, or semiconductor company Actively current on the AI/LLM landscape — tracking model releases, inferen [... source excerpt omitted ...] w) Strong written and verbal communication — able to engage credibly with ML engineers at frontier labs and VP/C-suite executives Authorized to work in the US; able to travel to customer sites and to Seoul HQ periodically Preferred Qualifications Prior experience at a US AI chip company, cloud silicon team, or AI infrastructure startup Familiarity with NPU/GPU accelerator ecosystems, PCIe integration, and data center hardware deployment Experience with inference optimization: quantization, kernel tuning, batching strategies, memory bandwidth optimization Proficiency in C, C++, or Rust Experience working with distributed or cross-timezone engineering teams Why Join Furios
Employer postings · Data from · Sources