Before you apply
Listed location: US, CA, Santa Clara | US, CA, Remote
Work arrangement: remote. A remote label does not confirm worldwide eligibility or visa sponsorship.
Read the employer’s description for qualifications, compensation and work eligibility. Confirm the position is still open on the application page.
Job description supplied by NVIDIA; category and skill labels may be inferred. How our listings work · Report a problem
Job Description
NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale, low-latency AI services. You’ll lead architecture and performance work across Dynamo and open source frameworks. You’ll collaborate across research, software, systems, and hardware teams to make AI inference faster, more efficient, and easier to deploy. You’ll engage with the broader ecosystem, including vLLM, SGLang, and TensorRT-LLM as well as with external partners to build the best operating system for AI. If you’re excited by deep learning, performance engineering, and distributed systems, we’d love to hear from you.
What you'll be doing:
Design, build, and maintain Dynamo integrations for open source frameworks vLLM, SGLang, TRTLLM.
Partner with open source communities to land measurable gains in latency, throughput, reliability, and efficiency.
Showcase NVIDIA token/watt leadership by pushing the pareto frontier on public/private benchmarks
Find and remove bottlenecks across runtimes, kernels, networking, routing, and orchestration.
Develop inference optimizations for scheduling, disaggregation, KV caching, and autoscaling.
What we need to see:
BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field (or equivalent experience).
3+ years building, profiling, and debugging performance-critical distributed or ML systems.
Strong programming skills in Python and/or Rust, C++.
Understanding of modern ML architectures and inference techniques
Ways to stand out from the crowd:
High agency and a track record of leading ambiguous work end to end.
Experience with AI Accelerators
Open-source contributions / leadership
Research in ML inference or distributed systems.
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!
#LI-Hybrid
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.Skills mentioned
Categories
Frequently asked questions
Is the Senior Deep Learning Algorithm Engineer position at NVIDIA remote?
Yes. The Senior Deep Learning Algorithm Engineer role at NVIDIA is a remote position. Country eligibility is not specified here; check the employer listing.
What type of employment is the Senior Deep Learning Algorithm Engineer role?
NVIDIA is hiring for a full-time Senior Deep Learning Algorithm Engineer position.
Which skills are mentioned for the Senior Deep Learning Algorithm Engineer job at NVIDIA?
Detected skill labels include Python, Rust, LLM, Distributed Systems. Check the employer description to distinguish required skills from preferred experience.
How do I apply for the Senior Deep Learning Algorithm Engineer position at NVIDIA?
You can apply for the Senior Deep Learning Algorithm Engineer role directly through NVIDIA's official application link provided on this page.
Similar AI jobs
Senior Applied Scientist, Efficient LLM Inference & Model Optimization
Nebius · fulltime
Senior Manager, Physical AI Communications
CoreWeave · fulltime
Recruiting Operations Program Manager
Baseten · fulltime
Builder Community Manager
Nebius · fulltime
IT Software Engineer, Infrastructure
Databricks · fulltime
Product Lead, Inference - USA
Inworld AI · fulltime