
Senior Solutions Architect, Networking and Compute Infrastructure
NVIDIA · Posted Sep 16
Graphics processing units, artificial intelligence platforms, accelerated computing, and data center solutions
Get a personal compatibility score
Add a resume for personal matches
About the role
NVIDIA is looking for a Senior Solution Architect, Networking and Compute Infrastructure to join its NVIDIA Infrastructure Specialist Team. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centres. This role will be interacting with customers, partners and internal teams to analyse, define and implement large scale Networking projects, combining Networking, System Design and Automation.
What you will do
- Primary responsibilities will include building AI/HPC infrastructure for new and existing customers.
- Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, real-time monitoring, logging, and alerting.
- Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation, and refinement.
- Develop tooling to automate and manage of large-scale infrastructure environments, to automate operational monitoring and alerting, and to enable self-service consumption of resources.
- Deploy monitoring solutions for the servers, network and storage.
- Perform troubleshooting bottom up from bare metal, operating system, software stack and application level.
- Being a technical resource, develop, re-define and document standard methodologies to share with customer and internal teams Support activities and engage in POCs/POVs for future improvements.
- Experience with system administration on Linux systems required (CentOS, RHEL, and Ubuntu preferred)
Skills used in this role
What the employer is looking for
- BS/MS/PhD or equivalent experience in Computer Science, Data Science, Electrical/Computer Engineering, Physics, Mathematics, other Engineering fields with at least 5+ years’ work or research experience in networking fundamentals, TCP/IP stack, and data centre compute architecture.
- Advance knowledge of HPC, AI & EVPN, BGP, OSPF, VXLAN protocols.
- Deep understanding of DC architecture fundamentals such as compute, storage (PFS) & InfiniBand, Ethernet, NVLink.
- Experience running HPC performance benchmarks, cluster health checks, and profiling tools to identify infrastructure bottlenecks.
- Python programming, bash scripting experience and Advance Linux knowledge.
- Extensive experience delivering automated network provisioning and comfortable with automation and configuration management tools including Jenkins, Ansible, Puppet. Chef, etc.
- Possess solid working knowledge of Ethernet/InfiniBand/RDMA core principles.
- Excellent customer-facing and communication skills (verbal and written in both languages), enabling effective engagement with customers, partners, and cross-functional teams across the India region. Listening skills in English are critical.
- Willingness to travel
Preferred qualifications
- Knowledge of CPU and/or GPU architecture including of Kubernetes, container related microservice technologies.
- Background with RDMA (InfiniBand or RoCE) fabrics.
- Linux or Networking Certifications (e.g., CCNP, CCIE) or NVIDIA-related certifications.
- Deep Knowledge on observability stack and build experience.
Benefits and support
- Compensation and benefits are detailed in the job posting
About NVIDIA
NVIDIA Corporation is a multinational technology company that designs graphics processing units (GPUs) for the gaming and professional markets, as well as system on a chip units (SoCs) for the mobile computing and automotive market. Pioneering accelerated computing, the company has become a driving engine of modern artificial intelligence, deep learning, and data center infrastructure.
- Industry
- Semiconductors
- Company size
- 42000+ employees
- Founded
- April 5, 1993
- Location
- Santa Clara, California, USA
- Funding stage
- Public Company
Funding
Public Company · $5M raised
Leadership
Founder, President and Chief Executive Officer
Executive Vice President and Chief Financial Officer
Founder and NVIDIA Fellow
Executive Vice President, Worldwide Field Operations
Recent coverage
NVIDIA Newsroom
NVIDIA Announces $1 Billion Commitment to Advance U.S. Super Intelligence Research and Quantum Leadership2026-10-08
NVIDIA Newsroom
NVIDIA Board of Directors Authorizes an Additional $150 Billion for Share Repurchases2026-10-03
NVIDIA Newsroom
NVIDIA Launches Open Agent Safety Platform to Secure Agents From Testing to Deployment2026-09-28
NVIDIA Newsroom
NVIDIA Reports Second Quarter Fiscal 2027 Financial Results with Record Revenue2026-08-26