Before you apply
Listed location: China, Beijing | China, Shanghai | China, Shenzhen
Work arrangement: onsite. A remote label does not confirm worldwide eligibility or visa sponsorship.
Read the employer’s description for qualifications, compensation and work eligibility. Confirm the position is still open on the application page.
Job description supplied by NVIDIA; category and skill labels may be inferred. How our listings work · Report a problem
Job Description
Join NVIDIA’s Cosmos Lab Infrastructure team to develop training and post-training systems for advanced Physical AI models, including world foundation models and robot policies. Our infrastructure connects training, inference, and evaluation with simulation and real-world robot interaction. You will work with a mentor on a focused project scoped to your experience and internship duration, implementing and evaluating systems improvements on real AI workloads using NVIDIA’s GPU infrastructure.
What you’ll be doing:
Develop and optimize training infrastructure for advanced Physical AI world models, supporting pre-training, supervised fine-tuning (SFT), and reinforcement learning (RL). Explore distributed parallelism, sharding, low-precision training, compute–communication overlap, and numerical consistency and efficient weight synchronization between training and inference.
Build Physical AI post-training and RL infrastructure supporting advanced training algorithms. Connect simulation or, where applicable, real-robot interaction with experience collection, rollout inference, reward computation, training, and evaluation. Optimize these workflows through partitioning, pipelining, data transfer, and synchronization across synchronous, asynchronous, or disaggregated execution.
Improve efficiency and scalability across training, inference, simulation, and evaluation through scheduling, placement, dynamic resource allocation, and load balancing, supporting heterogeneous resources, elasticity, and fault recovery.
Analyze and optimize system performance, working with researchers to investigate, support, and compare emerging Physical AI models, training workflows, and algorithms from a systems perspective. Use profiling, benchmarking, and performance modeling to identify bottlenecks and measure throughput, latency, GPU utilization, and policy freshness. Share findings through tested code, documentation, and technical presentations, and contribute to research publications where appropriate.
What we need to see:
Pursuing a Bachelor’s, Master’s, or PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related field.
Strong Python and debugging skills, with systems fundamentals in concurrency, distributed execution, memory management, or data movement.
Practical experience in at least one area: training infrastructure, RL infrastructure, simulation or robotics integration, or inference infrastructure. Coursework, research, open-source projects, and internships all count.
Strong analytical and communication skills, curiosity, and a willingness to learn.
Experience in every listed area, prior access to large GPU clusters, and model architecture or learning algorithm research are not required.
Ways to stand out from the crowd:
Experience optimizing training infrastructure, including distributed parallelism, low-precision training, GPU memory efficiency, or compute–communication overlap.
Experience optimizing scheduling, placement, resource allocation, or data transfer across training, rollout, simulation, and evaluation.
Experience extending RL pipelines, integrating simulation environments or robot interfaces, or optimizing inference; GPU profiling, C++/CUDA development, and open-source contributions or research in ML systems are also valued.
Skills mentioned
Categories
Frequently asked questions
Is the AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 position at NVIDIA remote?
The AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 role at NVIDIA does not have a confirmed remote arrangement in our data. Check the employer description for its work location.
What type of employment is the AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 role?
NVIDIA is hiring for a full-time AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 position.
Which skills are mentioned for the AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 job at NVIDIA?
Detected skill labels include Python, CUDA, Reinforcement Learning, GPU. Check the employer description to distinguish required skills from preferred experience.
How do I apply for the AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 position at NVIDIA?
You can apply for the AI Infrastructure and Frameworks Intern, Cosmos Lab - 2027 role directly through NVIDIA's official application link provided on this page.
Similar AI jobs
Director, Real Estate Development
NVIDIA · fulltime
Senior Research Engineer - Inference ML
Cerebras · fulltime
Marketing Recruiter
Baseten · fulltime
Senior Mission Systems Engineer, EO/IR Seekers
Anduril · fulltime
Mission Systems Engineer, Antenna Lead
Anduril · fulltime
Staff Embedded Engineer, EW
Anduril · fulltime