NVIDIA

Senior Solutions Architect, Continuous Bring Up Networking

NVIDIA  •  Seoul, KR (Remote)  •  3 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect for the Continuous Bring Up (CBU) role. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer!

What you'll be doing:

  • Primary responsibilities of the Continuous Bring Up (CBU) role will include stabilizing the NVIDIA AI Factory clusters after it is handed over to the customer once NVIDIA Infrastructure Specialist team deploys them.

  • CBU Networking will focus on the customer’s questions or the change requests of the network topology or the workload optimization in the networking perspective in order ultimately to help customers expand their clusters.

  • CBU Networking will also work internally with CBU DevOps responsible for the cluster orchestration layer and Infrastructure Solutions Architect responsible for the design of the cluster at the beginning.

  • CBU Networking may work not only for post-sales support but also the pre-sales support as Infrastructure Solutions Architect from time to time depending on the situation.

What we need to see:

  • BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields.

  • At least 8 years of professional experience in networking fundamentals, especially in AI/HPC cluster with NVIDIA platform.

  • Proficiency in configuring, testing, validating, and resolving issues in InfiniBand or Ethernet networks.

  • Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols.

  • Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS.

  • Ability to develop CI/CD pipelines for network operations.

  • Strong focus on customer needs and satisfaction.

  • Self-motivated with leadership skills to work collaboratively with customers and internal teams.

  • Strong written, verbal, and listening skills in English are essential.

Ways to stand out from the crowd:

  • Familiarity with NVIDIA Reference Architecture or NVIDIA Reference Design consists of the compute fabric, storage fabric, or the management fabric.

  • Linux or Networking Certifications.

  • Experience with NVIDIA cluster orchestration software such as Mission Control, Base Command Manager, or Run:ai.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the world working for us. If you're creative and autonomous, we want to hear from you.

NVIDIA

About NVIDIA

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Industry
Hardware & Semiconductors
Company Size
10,000+ employees
Headquarters
Santa Clara, CA
Year Founded
1993
Social Media