
Senior High-Performance Storage Architect - NVIS
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.
NVIDIA is seeking a Senior High-Performance Storage Architect to join our Infrastructure Specialists team in Santa Clara, CA. Here, you'll be at the forefront of innovation, contributing to the development of the largest and fastest AI Factories in the world. This role offers an outstanding opportunity to work with modern technologies and interact with academic, commercial, and government groups globally. You'll be part of a dynamic, customer-focused team that thrives on flawless execution and remarkable collaboration!
What you'll be doing:
Deploying, managing, and validating High-Performance Storage infrastructure within AI Factory Linux-based environments.
Acting as the domain expert during customer planning calls through implementation phases.
Crafting and handing over documentation and performing knowledge transfers to support customers.
Providing feedback to internal teams by opening bugs, documenting workarounds, and suggesting improvements.
What we need to see:
8+ years of experience providing in-depth support, deployment, and validation services for hardware and software-based storage products.
Solid understanding and experience with storage concepts and technologies, including SDS, NFS, S3, NVMeoF, Lustre, GPFS, and RMDA (RoCEv2/InfiniBand).
Familiarity with storage benchmarking tools such as FIO, IO500, IOR, HPL, NCCL tests, and MLPerf storage.
Proficiency in Linux system administration, performance reporting/optimization/logging, and network-routing/advanced networking (tuning and monitoring).
A four-year degree in Computer Science, Electrical or Computer Engineering, or equivalent experience.
Scripting proficiency in Bash, Python, Ansible, etc.
Outstanding interpersonal skills and the ability to deliver resolutions for customer issues as they arise.
Strong organizational skills and ability to prioritize and multi-task with limited supervision.
Willingness to travel to customer sites within the region up to 20% of the time.
Ways to stand out from the crowd:
Experience with parallel and distributed filesystem products (Ceph, Weka.io, Vast, DDN, etc.).
Expertise in optimizing storage performance for AI training, checkpointing, inference, or large-scale data pipelines.
Proven foundation in networking: Spectrum-X Ethernet, InfiniBand, NVLink Switch fabrics, congestion control, and datacenter topologies.
You will be redirected to the company website to complete your application.






