About the role
What will you do at NVIDIA?
Join NVIDIA as a Solution Architect on the Infrastructure Specialists team. Help
redefine deep learning, data analytics, and power data centers worldwide using
NVIDIA products. Collaborate on building the world's largest and fastest AI
Factories and supercomputers. We are seeking a candidate who can lead the
planning and deployment of large scale AI data centers, focusing on
infrastructure buildout including power and cooling systems, telemetry and
control systems, and large scale design, construction and delivery processes. In
this role, your main focus will be to support customers in the areas of
planning, design, construction, and deployment of large scale AI factories. You
will be a part of the team building capabilities to design, construct and
deliver large AI factories based on NVIDIA's reference designs. This includes
architectural systems, power distribution, cooling systems, integration of
telemetry and control systems, and all other physical infrastructure.
Collaboration with product and engineering teams, customers, and the
partner/provider ecosystem will be crucial to achieving successful deployments.
What you will be doing: NVIS Data Center deployment planning: Collaborate with
product and engineering teams to understand NVIDIA’s reference architectures for
data center infrastructure including power distribution, cooling systems,
controls and monitoring, and network/cabling architecture. Support customers and
partners in quickly implementing this architecture into advanced and reliable
data center designs. Building process capabilities: Collaborate across the org
to build processes, partner relationships and workflows to deliver and deploy
large AI factories at speed of light (SOL). Design and construction oversight:
Review and appraise customers' and partners' infrastructure design plans,
verifying their compliance with NVIDIA reference architecture, industry
standards, and regulatory requirements. Deliver guidance, expertise and
suggestions to optimize performance, scalability, and cost-effectiveness. Ensure
alignment with our customers and partners on reference architecture, guidelines
and processes to make their deployments successful. Assess the operational
efficiency, reliability, and readiness of data center infrastructure components
before deploying AI/HPC clusters. Develop and implement comprehensive audit
plans and conduct pre-deployment audits to identify potential issues, risks, and
areas for improvement. Partner and vendor ecosystem: Develop and sustain a
strong ecosystem of manufacturers, service providers and partners as needed, to
ensure customers can deploy NVIDIA solutions rapidly and reliably as well as be
the key liaison for customers and partners on matters of data center
infrastructure. Act as the NVIS mentor providing guidance, mentorship, and
support to ensure the team's success in their respective roles. Quality
Assurance: Implement and make quality assurance processes to ensure that
deployments meet established specifications and performance benchmarks. Conduct
detailed bring-up, testing, and commissioning to validate the functionality and
reliability of infrastructure components. Continuous Improvement: Drive
continuous improvement initiatives to improve data center infrastructure
reliability, resilience, and sustainability. Find opportunities to streamline
processes, automate repetitive tasks, and apply new technologies to optimize
infrastructure operations. Collaboration and Communication: Collaborate and
communicate across internal teams, external vendors, and customers to facilitate
the flawless integration of data center infrastructure solutions. Serve as a
domain authority and point of contact for infrastructure-related inquiries and
critical issues. What we need to see: Bachelor's degree or equivalent experience
in Engineering, or a related field. Advanced degree or equivalent experience or
relevant certifications are desirable. We need an expert professional with a
background in multiple aspects of infrastructure delivery, preferably of
hyperscale data centers. The ideal candidate will have at least 12 years of
experience, preferably in sophisticated, high-density AI/HPC data centers.
Proven experience in data center engineering, operations, deployment and/or
infrastructure management roles, focusing on large-scale data center
deployments. Strong technical knowledge and experience in data center systems
and processes- power distribution, liquid cooling, rack/server chassis, and
cabling. Proven technical and project leadership under fluid situations, and
ability to adapt to change. Strong communication at both ground execution as
well as executive levels, internally and with customers Excellent analytical,
problem-solving, communication and decision-making skills, keen attention to
detail, and a dedication to quality. Strong record of excellent partnership and
putting the mission's success first. Coordination & Time Management – proficient
at planning, scheduling, and coordinating tasks related to the job to accomplish
objectives within or ahead of designated time frames. Able to travel (25%). Way
to stand out from the crowd: Experience in hyperscale data center deployment,
operations process, safety, and security measures. Solid understanding of the
whole data center Infrastructure stack. Outstanding social skills. NVIDIA is
widely considered one of the world's most desirable employers in technology. We
have some of the world's most forward-thinking and passionate people working for
us. If you're creative and autonomous, we want to hear from you. Your base
salary will be determined based on your location, experience, and the pay of
employees in similar positions. The base salary range is 224,000 USD - 356,500
USD for Level 5, and 272,000 USD - 431,250 USD for Level 6. You will also be
eligible for equity and benefits. Applications for this job will be accepted at
least until July 3, 2026. This posting is for an existing vacancy. NVIDIA uses
AI tools in its recruiting processes. NVIDIA is committed to fostering an
inclusive work environment and proud to be an equal opportunity employer. As we
highly value diversity in our current and future employees, we do not
discriminate (including in our hiring and promotion practices) on the basis of
race, religion, color, national origin, gender, gender expression, sexual
orientation, age, marital status, veteran status, disability status or any other
characteristic protected by law. NVIDIA pioneered accelerated computing. Today,
our AI infrastructure powers global intelligence, transforming every industry.
Learn more about NVIDIA.
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
Technical Product Manager - AI Infra Resilience
NVIDIA · Santa Clara, California, United States
AI Principal Product Manager - Technical, Amazon Customer Service
Amazon · Seattle, Washington, United States
Staff Technical Product Manager (Platform)
Zeitview · Boston, Massachusetts, United States
Staff Technical Product Manager (Platform)
Zeitview · Mountain View, California, United States
Product Manager Insurance Services
Southland Steel Fabricators, Inc. · New Orleans, Louisiana, United States
Principal Product Manager- Technical - Performance, Demand Tech Products
Amazon · Seattle, Washington, United States
Role information can change. Confirm current details on the original application page.