HOME Cloud & Infrastructure Engineering Senior Solutions Architect, Physical AI Cloud
  • NVIDIA
  • US,
  • Full-Time
  • 35 days ago
  • $152,000 – $287,500
NVIDIA VERIFIED EMPLOYER

Senior Solutions Architect, Physical AI Cloud.

Cloud & Infrastructure Engineering Full-Time

Senior Solutions Architect, Physical AI Cloud: our view in 3 lines...

  • The Role:This role is for a senior solutions architect focused on cloud infrastructure for Physical AI and robotics workloads.
  • The Person:The person will design and scale Kubernetes-native environments, support Physical AI data factories, improve GPU workload efficiency, accelerate distributed inference, and provide technical guidance to customers and partners.
  • Requirements:The role requires BS in Computer Science, Computer Engineering, or related field, 5+ years in Solution Architecture or Infrastructure Engineering, Kubernetes-based platforms, networking, storage technology, workflow orchestration softwares, and modern DevOps practices.

About the role

We're building a group of innovators to assist enterprises in deploying and accelerating NVIDIA’s three computer workloads for Physical AI. These include robotics simulation, synthetic data generation, multi-step model training, and inference, all on a large scale!

We are seeking a hands-on Solutions Architect with deep expertise in backend infrastructure, inference and cloud-native applications to design and scale Kubernetes-native environments for distributed Robotics workloads. This role offers an outstanding chance to build within the rapidly growing field of Robotics AI & Simulation. You’ll work closely with our product management, engineering, and business teams to drive the adoption of NVIDIA's groundbreaking Physical AI technologies with our key ecosystem partners!

What you’ll be doing:

  • Help partners build scalable, observable, GPU-accelerated Physical AI pipelines through agentic workflows, cloud-native technologies, and NVIDIA frameworks such as OSMO.

  • Support development of Physical AI data factories for data ingestion, preprocessing, annotation, filtering, synthetic data generation, training, simulation, and evaluation.

  • Develop a deep understanding of robotics workload scaling and translate customer requirements into optimized cloud-native architectures, improving scheduling, cost, storage access, networking, and GPU utilization across hybrid infrastructure.

  • Accelerate distributed inference using NVIDIA technologies such as NIM, TensorRT-LLM, vLLM, and SGLang.

  • Collaborate with business, engineering, and product teams while providing technical guidance and mentorship to customers implementing Physical AI at scale.

What we need to see:

  • BS in Computer Science, Computer Engineering, or a related field, or equivalent experience.

  • 5+ Years of experience in Solution Architecture or Infrastructure Engineering, advancing AI/ML systems from proof of concept to production on private/public cloud environments.

  • Experience with scaling Robotics workloads in one or more areas, such as multimodal model training, inference, robot learning and simulation, large scale data processing and generation.

  • Strong hands-on experience designing, deploying, and operating Kubernetes-based platforms for distributed GPU and AI workloads.

  • Expertise in networking (DNS, LB, TCP/IP, firewalls), storage technology, workflow orchestration softwares (Airflow, Argo, etc), modern DevOps practices (GitOps, IaC, Observability), and orchestrating efficient GPU workloads

  • Excellent communication skills to convey technical concepts to diverse audiences.

Ways to stand out from the crowd:

  • Hands-on experience with robotics frameworks (e.g., ROS2) and NVIDIA simulation and AI platforms such as Isaac Lab, Isaac Sim, GR00T or Cosmos.

  • Previous exposure to large scale Robotics data curation, annotation, filtering pipelines, including the use of AI models for data labeling.

  • Experience deploying NVIDIA inference technologies (Dynamo, NIM, Triton, vLLM) using acceleration techniques like quantization.

  • Proficiency using and developing agentic workflows to accelerate software development, infrastructure automation, troubleshooting, and deployment workflows.

  • Broad technical expertise across networking, compute, and storage systems (e.g., S3, NFS, Lustre), with hands-on experience building and debugging APIs (REST, gRPC).

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 14, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Published August 10, 2026
Location US, United States of America
Job Type Full-Time