Job opportunity

Senior HPC and AI Network Software Architect / Senior HPC and AI Network Software Architectess

NVIDIA Switzerland AG Zürich September 6, 2026

Join NVIDIA as a Senior HPC and AI Network Software Architect

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for over 25 years. Our legacy of innovation is fueled by remarkable technology and extraordinary talent. Today, we are harnessing the limitless potential of AI to shape the next era of computing, where our GPU serves as the brains behind computers, robots, and self-driving cars capable of understanding their environment. Achieving what has never been done before demands vision, innovation, and the world’s top talent. As an NVIDIAN, you will thrive in a diverse and supportive environment that inspires everyone to excel. Join our team and discover how you can make a profound impact on the world.

Your Role and Responsibilities

We are seeking a Senior HPC and AI Network Software Architect to spearhead the development of the next generation of scalable AI infrastructure. This role focuses on distributed training, real-time inference, and enhancing communication efficiency across large systems. Your contributions will include:

  • Building and evolving the architecture of scalable software systems for distributed AI training and inference, with an emphasis on throughput, latency, resiliency, and memory efficiency in cluster-scale deployments.
  • Developing and evaluating next-generation communication and runtime capabilities in libraries such as NCCL, UCX, and UCC tailored for the rapidly evolving frontier of AI workloads.
  • Collaborating with AI framework teams (e.g., TensorFlow, PyTorch, JAX) and internal platform teams to create seamless integrations, explore innovative approaches, and enhance end-to-end performance and reliability.
  • Working on hardware and system-level features across GPUs, DPUs, and interconnects to accelerate data movement and enable new functionalities for training, inference, and model serving at scale.
  • Pioneering developments across runtime systems, communication libraries, and AI-specific protocol layers, helping transform new ideas into practical capabilities and reliable implementations.

What We Are Looking For

  • A Ph.D. or equivalent industry experience in computer science, computer engineering, or a closely related field.
  • 5+ years of experience in systems programming, parallel or distributed computing, high-performance networking, or large-scale data movement, including experience in designing and constructing complex systems.
  • A strong programming background in C++, Python, and ideally CUDA or similar GPU programming models, alongside a proven track record of developing production-quality, performance-critical software.
  • Extensive hands-on experience with AI frameworks such as PyTorch, TensorFlow, or JAX, and a solid understanding of how communication libraries and runtime systems facilitate large-scale training and inference.
  • Demonstrated success in developing and refining high-throughput, low-latency systems, with the ability to reason across software stacks, hardware capabilities, and system bottlenecks.
  • Strong collaboration skills in a multi-national, interdisciplinary setting, capable of contributing ideas, building momentum, and working effectively with senior engineers, researchers, and partner teams.

Ways to Stand Out

  • Deep expertise with NCCL, UCX, UCC, or similar communication libraries used in large-scale AI and HPC workloads.
  • A solid background in networking and communication protocols, RDMA, collective communications, congestion-aware transport, or accelerator-aware networking.
  • Comprehensive knowledge of large model training and inference serving at scale, including communication bottlenecks, scheduling challenges, and system-level tradeoffs among compute, memory, and fabric.
  • Experience in hardware-software co-design for distributed AI systems, including advancements in GPU, DPU, interconnect, or runtime capabilities.
  • Familiarity with deploying infrastructure for LLMs or transformer-based models, covering aspects like sharding, pipelining, expert parallelism, or hybrid parallelism.

Why NVIDIA?

At NVIDIA, you will collaborate with individuals dedicated to continuous learning and creative problem-solving in the industry, pushing the boundaries of what is possible in AI and high-performance computing. If you are passionate about architecting distributed systems, advancing AI infrastructure, and solving large-scale problems, we want to hear from you!

Regarded as one of the most desirable employers in the technology sector, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you envision your future, explore what we can offer you and your family.

Apply online using the form below. Please note that only applications matching the job profile will be considered.

Work locationZürich, Switzerland

Application Form

Please enter your information in the following form and attach your resume (CV)

Only pdf, Word, or OpenOffice file. Maximum file size: 3 MB.