NVIDIA Logo

NVIDIA

Senior Solutions Architect, AI Infrastructure Enterprise ISVs

Posted 5 Days Ago
Be an Early Applicant
In-Office or Remote
3 Locations
184K-288K Annually
Senior level
In-Office or Remote
3 Locations
184K-288K Annually
Senior level
Advise ISVs on designing, deploying, and optimizing large-scale accelerated AI infrastructure. Lead architecture reviews, POCs, benchmarks, and production deployment guidance across compute, networking, storage, containers, orchestration, observability, and CI/CD. Create reference architectures, sizing guidance, technical playbooks, demos, and whitepapers. Support cluster monitoring, reliability, and performance improvements. Up to 20% travel for customer engagements.
The summary above was generated by AI

NVIDIA is seeking outstanding AI Solutions Architects to assist and support customers that are building solutions with our newest AI and accelerated computing technologies. At NVIDIA, our solutions architects work across product, engineering, sales, developer relations, business development, and partner teams to help customers design, deploy and optimize AI infrastructure.

This role will focus on helping ISVs adopt NVIDIA accelerated infrastructure for training, fine-tuning, inference, retrieval, and agentic AI workloads. This role is an excellent opportunity to work in an interdisciplinary team at NVIDIA! You will serve as a technical advisor for accelerated systems architecture, GPU and networking systems, cluster design, architectures, orchestration, validation, and production deployment for AI data centers.

What You Will Be Doing:

  • Partner with ISVs on discovery, architecture reviews, technical deep dives, POCs, benchmarks, demos, and production deployment guidance

  • Advise on the design, build-out, and optimization of accelerated AI infrastructure, including large-scale clusters

  • Support infrastructure design across compute, networking, storage, containers, observability, security, power, and data center operations

  • Drive adoption of systems monitoring, telemetry, and management tools to improve cluster utilization, reliability, performance and workload insight

  • Build repeatable reference architectures, deployment guides, sizing guidance, benchmark reports, technical playbooks, demos and whitepapers

  • Travel up to 20% customer meetings may be required

What We Need To See:

  • BS, MS, or PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, other Engineering or related fields (or equivalent experience)

  • 8+ years of hands-on experience in AI infrastructure, accelerated computing, distributed systems, cloud infrastructure, high-performance computing, or machine learning platforms

  • Strong experience designing, deploying, and operating accelerated computing infrastructure at scale

  • In-depth knowledge of AI cluster orchestration, scheduling, automation and CI/CD deployment pipelines

  • Understanding of data center networking technologies such as InfiniBand, Ethernet, RDMA, network configuration or performance tuning

  • Familiarity with infrastructure requirements for AI workloads, including distributed training, inference serving, model deployment, storage performance, and cluster reliability

  • Excellent presentation, communication, problem-solving, documentation, and collaboration skills

Ways To Stand Out From The Crowd:

  • Experience architecting AI factories, large GPU clusters, multi-node training environments, production inference platforms

  • Experience deploying LLM training, fine-tuning, RAG, and inference workflows on large-scale AI infrastructure

  • Experience evaluating cluster performance using benchmarks such as MLPerf, HPL, or workload-specific performance tests

  • Applications and systems-level knowledge of OpenMPI, NCCL, distributed training frameworks, and GPU communication patterns

  • Experience delivering technical training, workshops, whitepapers, blogs, or mentoring engineers, researchers, and customers on AI/HPC infrastructure

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 20, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar Jobs

50 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
116K-137K Annually
Mid level
116K-137K Annually
Mid level
Insurance
Own state-level product strategy and market performance for assigned regions. Manage state filings and DOI relationships, collaborate with Technology and Operations on launches, analyze KPIs and profitability drivers, present strategic recommendations, support external data and competitor intelligence, and relay field feedback to product teams to inform product changes and new development.
57 Minutes Ago
Remote
USA
65K-72K Annually
Mid level
65K-72K Annually
Mid level
eCommerce • Retail
Manage marketing operations for eCommerce growth: enforce asset governance, prioritize creative production, traffic ads, ensure data integrity in reporting, analyze campaign performance, and present recommendations to improve campaign strategy and execution.
Top Skills: AsanaDriveFigmaGmailGoogle AdsGoogle DocsGoogle Dv360Google SheetsGoogle WorkspaceLookerMeta Ads Manager (Facebook Ads Manager)NotionSlackTiktok Ads ManagerTtcxYoutube
57 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
Senior level
Senior level
Enterprise Web • Mobile • Professional Services • Software
Design, build, and own scalable data pipelines and evaluation systems that power production AI features and internal analytics. Ensure data quality across ingestion, modeling, and reporting, collaborate with ML and analytics teams, deploy and monitor ML systems, and establish standards for data work and evaluation.
Top Skills: AirflowAWSDagsterGCPPostgresPythonSnowflake

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account