Senior Datacenter System Software Architect - DGX Cloud
NVIDIA Corporation - Santa Clara, California, us, 95053
Work at NVIDIA Corporation
Overview
- View job
Overview
Providing expertise in infrastructure workflows, including hardware, software release, workload orchestration and application tuning
Provide fast and creative solutions for complex problems and write effective, clear and reliable architecture specification
Translate requirements to vision, architecture and roadmap
Work with engineering teams across NVIDIA to ensure your software integrates seamlessly from the hardware all the way up to the AI training applications.
What we need to see: Masters or PhD in Computer Science, Computer Engineering, Physics or equivalent experience.
12+ years of experience in this field.
Data Sciences, Deep Learning, or Machine Learning coursework
Ability to seamlessly shift between Linux system environments to Python programming
Programming skills in 1 or more high-level languages (C, C++, Go, Rust, etc)
System-level experience with both hardware and software
Motivated self-starter with an equal balance of strong problem-solving skills and customer-facing communication skills
Strong design, coding, analytical, debugging and problem-solving skills
Passion for continuous learning and knowledge transfer. Ability to work concurrently with multiple groups locally and abroad in the organization
Ways to stand out from the crowd: Experience with GPU deep learning and data sciences. Experience using TensorFlow, PyTorch or other DL framework. Experience working with Docker containers, Slurm, Terraform and Kubernetes
CUDA programming and NCCL experience. HPC programming experience including MPI, OpenACC, or other parallel programming tools. Hands-on experience with DGX Cloud, NVIDIA AI Enterprise AI Software, Base Command Manager, NEMO and NVIDIA Inference Microservices.
Interest in crafting, analyzing and fixing large-scale distributed systems.
Systematic problem-solving approach, coupled with strong communication skills and a sense of ownership and drive.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD. You will also be eligible for equity and benefits . Applications for this job will be accepted at least until August 9, 2025.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. Similar Jobs (5)
Distinguished Engineer – Data Center System Software Architect locations US, CA, Santa Clara time type Full time posted on Posted 6 Days Ago Distinguished Engineer – Data Center System Software Architect locations 2 Locations time type Full time posted on Posted 3 Days Ago System Software Architect, HPC Networking locations US, CA, Santa Clara time type Full time posted on Posted 13 Days Ago NVIDIA is the world leader in accelerated computing. NVIDIA pioneered accelerated computing to tackle challenges no one else can solve. Our work in AI and digital twins is transforming the world's largest industries and profoundly impacting society.
#J-18808-Ljbffr