
Senior Solutions Architect, Networking & GPU System
Beijing, Beijing, ChinaFull-timePosted 3w agoStill listed 2w ago
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Beijing, Beijing, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
NVIDIA seeks a Senior Solutions Architect to work with account and OEM teams, designing and validating SONiC‑based networking and GPU server solutions for large‑scale AI and HPC platforms, while providing production support and collaborating on complex technical issues.
Skills & qualifications
Skills
Qualifications
Full job description
NVIDIA is a world‑leading, fast‑growing AI computing company, delivering everything from the most powerful GPU‑accelerated supercomputers to gigawatt‑scale AI data centers. We build and optimize end‑to‑end GPU server platforms that combine accelerated computing, high‑performance networking, and software to power state‑of‑the‑art AI infrastructure. We believe in our people and products, and we are looking for outstanding talent to join us. We are looking for a Senior Solutions Architect who is both customer‑facing and deeply hands‑on with SONiC‑based networking and GPU server infrastructure. In this role, you will partner with account and OEM partner teams to qualify opportunities, design solutions, run technical evaluations, and prove value to customers building large‑scale AI and HPC platforms. Supporting production clusters, including monitoring and troubleshooting, is also part of this role. What you will be doing
- Work closely with account managers to understand customer requirements, position our SONiC + Spectrum whitebox switch solutions, and shape technical win strategies for AI and HPC opportunities.
- Design end‑to‑end architectures for customer proposals, including GPU server configurations, storage connectivity, and SONiC‑based leaf‑spine fabrics with RDMA/RoCE.
- Build and validate hands‑on demos and POCs: deploy SONiC switches and GPU servers, configure networking, install software stacks, and run benchmarks to prove performance, scalability, and reliability.
- Provide support for large-scale production SONiC switch clusters.
- Own and manage customized SONiC software projects on Spectrum white-box switches.
- Collaborate with internal engineering, product, and OEM partners to resolve complex issues in firmware, drivers, OS, routing, and GPU/network performance, then bring fixes back to active POCs.
What we need to see
- BS/BA in Computer Science, Electrical/Computer Engineering, or equivalent practical experience.
- 6+ years in data center or cloud infrastructure roles (solutions architect, systems engineer, network engineer) with direct exposure to presales or customer‑facing technical work.
- Solid background in data center networking for AI workloads: leaf‑spine designs, high‑bandwidth/low‑latency fabrics, RDMA/RoCE, and ideally InfiniBand; comfortable configuring and debugging these in lab and customer environments.
- Practical knowledge of SONiC: installing and upgrading, configuring interfaces and routing (BGP/EVPN/VXLAN), using monitoring/telemetry tools, and troubleshooting real incidents.
- Strong, practical understanding of GPU server architecture: CPU/GPU balance, memory bandwidth, PCIe/NVLink topology, storage and NIC placement, and power/cooling at rack level.
- Hands‑on experience designing, deploying, or operating AI/HPC clusters using GPU‑accelerated servers (on‑prem or cloud), including real involvement in sizing, configuration, and performance tuning.
- Excellent communication and presentation skills: able to explain complex technical topics to both highly technical engineers and non‑technical decision makers, and write clear design documents and POC reports. Fluent written and spoken English is required.
Ways to stand out from the crowd
- Contributions to open‑source SONiC, networking, or AI infrastructure projects, or published talks/whitepapers on AI data center design and performance tuning.
- Coding experience on SONiC.
- Coding experience with NCCL, NIXL or other collective communication library, RDMA applications, or performance‑critical distributed training frameworks, giving you credibility when discussing low‑level performance with customer engineers.
- Hands‑on deployments of SONiC in production or large lab environments, especially integrated with GPU clusters and RDMA/RoCE fabrics.
- Hands-on experience with NVIDIA networking products and solutions (Spectrum-X, InfiniBand, Cumulus Linux, etc.).
With competitive salaries and a generous benefits package, we are widely considered to be one of the world’s most desirable employers! We have some of the most forward-thinking and hardworking people in the world working for us and, due to outstanding growth, our best-in-class engineering teams are rapidly growing. If you're a creative and autonomous person with a real passion for technology, we want to hear from you.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
Senior Solutions Architect - CRISP SystemNVIDIA · Beijing, Beijing, ChinaPosted 1w agoPosted 1w ago
Senior Solution Architect, CBU TPM for NCPNVIDIA · Beijing, Beijing, ChinaPosted 1w agoPosted 1w ago
Senior Solutions Architect - ARM CPUNVIDIA · Beijing, Beijing, ChinaPosted 1w agoPosted 1w ago
Senior Solutions Architect, CSP SystemNVIDIA · Shenzhen, Guangdong Province, ChinaPosted 1w agoPosted 1w ago
Senior Software and System ArchitectNVIDIA · Shanghai, Shanghai, ChinaPosted 1w agoPosted 1w ago
You've read the whole posting — now see how you match it.