Back to jobs

Senior Solutions Architect, Continuous Bring Up Networking

NVIDIA
Korea, Seoul
Full-time
AI tools:
Run:ai
You apply on NVIDIA's own careers site

This customer-facing networking architect helps stabilize and expand NVIDIA AI/HPC clusters after deployment. The role troubleshoots network issues and topology changes while coordinating with customers, partners, and internal infrastructure and DevOps teams.

Permanent
On-site
8+ years
Bachelor's, Master's, or PhD

Skills & Expertise

InfiniBand
Ethernet
EVPN
BGP
OSPF
VXLAN
Cumulus Linux
SONiC

Key Responsibilities

Stabilize AI Factory clusters after deployment and resolve customer networking questions.

Assess network topology changes and workload optimization requests to support cluster expansion.

Coordinate networking work with customers, partners, and internal infrastructure and DevOps teams.

Full Description

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect for the Continuous Bring Up (CBU) role. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer!

What you'll be doing:

• Primary responsibilities of the Continuous Bring Up (CBU) role will include stabilizing the NVIDIA AI Factory clusters after it is handed over to the customer once NVIDIA Infrastructure Specialist team deploys them.

• CBU Networking will focus on the customer’s questions or the change requests of the network topology or the workload optimization in the networking perspective in order ultimately to help customers expand their clusters.

• CBU Networking will also work internally with CBU DevOps responsible for the cluster orchestration layer and Infrastructure Solutions Architect responsible for the design of the cluster at the beginning.

• CBU Networking may work not only for post-sales support but also the pre-sales support as Infrastructure Solutions Architect from time to time depending on the situation.

What we need to see:

• BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields.

• At least 8 years of professional experience in networking fundamentals, especially in AI/HPC cluster with NVIDIA platform.

• Proficiency in configuring, testing, validating, and resolving issues in InfiniBand or Ethernet networks.

• Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols.

• Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS.

• Ability to develop CI/CD pipelines for network operations.

• Strong focus on customer needs and satisfaction.

• Self-motivated with leadership skills to work collaboratively with customers and internal teams.

• Strong written, verbal, and listening skills in English are essential.

Ways to stand out from the crowd:

• Familiarity with NVIDIA Reference Architecture or NVIDIA Reference Design consists of the compute fabric, storage fabric, or the management fabric.

• Linux or Networking Certifications.

• Experience with NVIDIA cluster orchestration software such as Mission Control, Base Command Manager, or Run:ai.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the world working for us. If you're creative and autonomous, we want to hear from you.

Applications are handled on NVIDIA's site