AI Network System Architect

NVIDIA
Apply Now

Job Description

We seek a highly motivated Senior AI Network System Architect to join our team of experts and help shape the future of high-performance and ML / AI computing. Our next-generation Infiniband, NVLink, and Ethernet systems will be at the forefront of connecting and powering the world's most advanced AI clusters. As an AI system architect at NVIDIA, you will have the opportunity to work on some of the most cutting-edge technology and help drive the innovation of our next-generation networks that top researchers and engineers worldwide will use.

What You’ll Be Doing:

  • Investigating emerging technologies and methodologies in ML and AI to discern their interactions with network infrastructure.
  • Executing workloads on AI systems, conducting profiling, and analyzing bottlenecks and possible enhancements.
  • Conducting research and implementing optimizations for communication libraries like NCCL and UCX.
  • Spearheading the conceptualization of next-generation networking products tailored to support and accelerate state-of-the-art ML workloads.
  • Develop models for simulations, analyze simulation results, and develop optimization algorithms.
  • Collaborate with multi-functional teams, including other architecture teams, logic design, system software, firmware, and ML research teams, to ensure the successful execution of the project.

What We Need To See:

  • Sc, or Ph. D degree in Computer Science, Computer Engineering, or Electrical Engineering.
  • At least 2+ years of industry or research experience in computer networks.
  • Extensive expertise in ML/AI workloads, particularly in distributed training.
  • Excellent understanding of large-scale network behavior and the effect of distributed computing workloads on the network.
  • Experience in the development of simulation environments.
  • Great problem-solving and critical-thinking skills.
  • Ability to thrive in a fast-paced and dynamic environment is necessary.
  • Ability to work concurrently with multiple groups in the organization.

Ways To Stand Out Of The Crowd:

  • Knowledge of communication libraries such as NCCL, UCX, and UCC.
  • Good knowledge of network protocols - such as InfiniBand, IP, TCP, RoCE, and network topologies.
  • Experience with Python, C++, and dockers.
  • Expertise in system engineering, operations research, and intricate hardware-software integrated systems.
  • Demonstrated experience in DLRM, LLM or other generative AI.

NVIDIA has some of the most forward-thinking and hardworking people in the world working for us, and due to unprecedented growth, our world-class engineering teams are growing fast. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you.

We are committed to fostering a diverse work environment and are proud to be an equal-opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Company Info.

NVIDIA

NVIDIA’s invention of the GPU sparked the PC gaming market. The company’s pioneering work in accelerated computing—a supercharged form of computing at the intersection of computer graphics, high performance computing and AI—is reshaping trillion-dollar industries, such as transportation, healthcare and manufacturing, and fueling the growth of many others.

  • Industry
    Cloud computing,Video games,Computer software,Semiconductors,Computer hardware,Consumer electronics,Artificial intelligence
  • No. of Employees
    22,473
  • Location
    2701 San Tomas Expressway, Santa Clara, CA 95050, USA
  • Website
  • Jobs Posted

Get Similar Jobs In Your Inbox

NVIDIA is currently hiring AI Architect Jobs in Tel Aviv, Israel with average base salary of ₪260,000 - ₪400,000 / Year.

Similar Jobs View More