JJobsSonar

Solutions Architect – AI Factory

NVIDIA · Hamburg, Hamburg, Germany

AI/MLRemote

About this role

NVIDIA is hiring a Solutions Architect – AI Factory with experience designing, building, and maintaining large-scale AI factories, helping NVIDIA AI Factory solutions bring large-scale AI to leading enterprise customers and operationalizing AI solutions at scale. The day-to-day involves guiding customers in adopting NVIDIA's compute, networking, and software stacks to deliver end-to-end GenAI and Agentic AI solutions, using cloud-native methodologies, low-latency networks, and accelerated compute to build modern AI factories, sharing knowledge through demos, proofs-of-concept, and developer blogs, and collaborating with executives and engineering to solve complex problems in the cloud and datacenter. Requirements include an MS or PhD in Engineering or Computer Science, an established track record with AI and HPC clusters (on-prem and cloud), 8+ years with cluster management tools (Docker, Slurm, Kubernetes, Ansible), and hands-on experience with datacenter MEP, network, storage, and cluster configuration and debugging; strong CUDA, Python, and C/C++ coding skills and GPU/InfiniBand experience stand out.

Skills & technologies

Must have

  • Kubernetes
  • Slurm
  • Ansible
  • CUDA
  • Python
  • GPU
  • InfiniBand
  • Docker

Nice to have

  • Base Command Manager
  • Run:ai
  • NVIDIA NIM

Read full description

About the job We are looking for a Solutions Architect – AI Factory with experience in designing, building, and maintaining large scale AI factories to join our team at NVIDIA. As Solution Architects on the AI Factory team, we are actively helping NVIDIA AI Factory solutions bring the benefits of large-scale AI to leading enterprise customers. We work closely with customers and partners to address unsolved problems in the industry and help to deploy and operationalize AI solutions at scale. What You'll Be Doing Our day-to-day work involves guiding customers in their adoption of NVIDIA's compute, networking, and software stacks to deliver end-to-end GenAI and Agentic AI solutions. Don't think this is a high-level slideshow job - we are the voice of experience, using cloud native methodologies, low latency networks, and accelerated compute to help build modern AI factories. We also excel at sharing knowledge with others, whether it's delivering demos, assisting with proof-of-concepts, or writing papers and developer blogs. By collaborating with executives and engineering, we solve complex problems and help bring NVIDIA's premiere technologies to life in the cloud and in the datacenter. Our mission is to solve the problems that nobody else has solved yet, and we need someone to be an instrumental part of that! What We Need To See MS, or PhD in Engineering, Computer Science, or a related field (or equivalent experience). Established track record working with AI and HPC clusters, both on-premises and cloud based. 8+ years of proven experience with cluster management and related tools, including Docker Containers, Slurm, Kubernetes, and Ansible. Hands-on experience with Datacenter MEP, network, storage, cluster configuration and debugging. Strong analytical and problem-solving skills, along with an ability to articulate what you know to others. Ability to multitask efficiently in a dynamic environment. Ways To Stand Out From The Crowd Strong coding and debugging skills, including experience with CUDA, Python, C/C++, Bash, AI frameworks and Linux utilities. Demonstrated expertise through projects or Open Source contributions involving GPU workloads, Kubernetes, InfiniBand, Ethernet, or other areas related to high-performance clusters and hybrid cloud solutions. Exhibit hands on experience with NVIDIA Enterprise software products, Base Command Manager, Run:ai and NVIDIA NIMs. Willingness and ability to learn quickly and solve advanced problems. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 292,500 PLN - 507,000 PLN for Level 4, and 375,000 PLN - 650,000 PLN for Level 5. , , JR2014615
PLN 292,500–PLN 650,000 / yearApply now

Similar AI/ML jobs

All AI/ML jobs