NVIDIA Korea · 채용 중 27건
Senior Solutions Architect, Continuous Bring Up Networking
Senior Solutions Architect, Continuous Bring Up Networking
솔루션 아키텍트정규직시니어 · 8년 이상
NVIDIA에서 AI/HPC 클러스터의 안정화 및 최적화를 담당할 Senior Solutions Architect를 채용합니다. 네트워킹 분야에서 8년 이상의 경력이 필수이며, InfiniBand 및 Ethernet 환경에서의 구성 및 문제 해결 능력이 요구됩니다. EVPN, BGP, CI/CD 파이프라인 구축 경험이 있는 분을 찾습니다. 고객 중심의 사고와 뛰어난 영어 커뮤니케이션 역량이 중요합니다.
NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect for the Continuous Bring Up (CBU) role. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer!
Primary responsibilities of the Continuous Bring Up (CBU) role will include stabilizing the NVIDIA AI Factory clusters after it is handed over to the customer once NVIDIA Infrastructure Specialist team deploys them.
CBU Networking will focus on the customer’s questions or the change requests of the network topology or the workload optimization in the networking perspective in order ultimately to help customers expand their clusters.
CBU Networking will also work internally with CBU DevOps responsible for the cluster orchestration layer and Infrastructure Solutions Architect responsible for the design of the cluster at the beginning.
CBU Networking may work not only for post-sales support but also the pre-sales support as Infrastructure Solutions Architect from time to time depending on the situation.
BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields.
At least 8 years of professional experience in networking fundamentals, especially in AI/HPC cluster with NVIDIA platform.
Proficiency in configuring, testing, validating, and resolving issues in InfiniBand or Ethernet networks.
Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols.
Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS.
Ability to develop CI/CD pipelines for network operations.
Strong focus on customer needs and satisfaction.
Self-motivated with leadership skills to work collaboratively with customers and internal teams.
Strong written, verbal, and listening skills in English are essential.
Familiarity with NVIDIA Reference Architecture or NVIDIA Reference Design consists of the compute fabric, storage fabric, or the management fabric.
Linux or Networking Certifications.
Experience with NVIDIA cluster orchestration software such as Mission Control, Base Command Manager, or Run:ai.
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the world working for us. If you're creative and autonomous, we want to hear from you.