Before you apply
Listed location: US, CA, Santa Clara | US, WA, Redmond
Work arrangement: onsite. A remote label does not confirm worldwide eligibility or visa sponsorship.
Read the employer’s description for qualifications, compensation and work eligibility. Confirm the position is still open on the application page.
Job description supplied by NVIDIA; category and skill labels may be inferred. How our listings work · Report a problem
Job Description
Artificial intelligence is moving from passive assistance to agents that can reason, use tools, and complete work on a person's device. NVIDIA is building the software foundation that helps developers and partners deliver these experiences privately, efficiently, and responsibly across GeForce RTX, NVIDIA RTX PRO, RTX Spark, DGX Spark, and DGX Station systems.
We are looking for an Engineering Manager to build and lead the team developing local and hybrid AI agent technology. You will set technical direction, grow engineers, and deliver production software spanning agent runtimes, local inference, Windows integration, developer tools, and security controls. You will turn lessons from Hermes, OpenClaw, Perplexity, and domain-specific agents into platform capabilities that work across NVIDIA systems.
What you'll be doing:
Lead the architecture and implementation of local AI agent runtimes on Windows and NVIDIA RTX systems, with clear interfaces across models, tools, GPU software, operating-system services, and applications.
Guide hands-on development through prototypes, design and code reviews, performance analysis, and debugging; help the team turn research ideas into reliable C++ and Python software.
Build core agent capabilities including MCP and tool routing, persistent memory, skills, scheduling, multi-agent coordination, computer use, and policy-aware local or hybrid model routing.
Own CI/CD, evaluation, compatibility, and regression systems across Windows releases, drivers, GPUs, models, and agent frameworks; add telemetry and diagnostics that make failures reproducible.
Partner with AI research, GPU and driver, Windows platform, product security, Microsoft, OEM, ISV, and open-source teams to land durable interfaces and production integrations.
Hire and develop engineers, set a high technical bar, allocate capacity, and create an environment where the team can make sound decisions and deliver maintainable software.
What we need to see:
Bachelor's degree in Computer Science, Computer Engineering, or a related field, or equivalent experience in practice.
10+ years of software engineering experience, including 3+ years managing engineers or leading a comparable technical organization.
Experience hiring and developing engineers, setting goals, providing useful feedback, and building an inclusive team environment.
A record of delivering complex software platforms or systems from architecture through production operation.
Technical fluency in AI agents, model inference, systems software, and local or hybrid deployment, with the judgment to guide architecture and execution.
Ability to align engineering, research, product, security, and external partners around clear decisions and measurable outcomes.
Proficiency with C++ or Python.
Ways to stand out from the crowd:
Hands-on experience with Windows internals, process isolation, virtualization, containers, or sandboxing.
Familiarity with OpenClaw, Hermes, LangChain or similar agent frameworks; MCP; memory, skills, multi-agent, or computer-use patterns.
Experience with TensorRT-LLM, Ollama, llama.cpp, vLLM, PyTorch, ONNX Runtime, Windows ML, model optimization, or quantization.
Experience with CUDA, GPU performance, local models on consumer hardware, or privacy-aware inference routing.
Experience contributing to open-source projects or working with Microsoft, OEMs, ISVs, and developer communities.
Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.Skills mentioned
Categories
Frequently asked questions
Is the Engineering Manager, Local AI Agents position at NVIDIA remote?
The Engineering Manager, Local AI Agents role at NVIDIA does not have a confirmed remote arrangement in our data. Check the employer description for its work location.
What type of employment is the Engineering Manager, Local AI Agents role?
NVIDIA is hiring for a full-time Engineering Manager, Local AI Agents position.
Which skills are mentioned for the Engineering Manager, Local AI Agents job at NVIDIA?
Detected skill labels include Python, PyTorch, CUDA, LLM, Spark, LangChain, GPU. Check the employer description to distinguish required skills from preferred experience.
How do I apply for the Engineering Manager, Local AI Agents position at NVIDIA?
You can apply for the Engineering Manager, Local AI Agents role directly through NVIDIA's official application link provided on this page.
Similar AI jobs
Member of Technical Staff - Engineering Lead, Data Platform
Reflection AI · fulltime
Product Strategist | Aeromechanical Engineering
Gecko Robotics · fulltime
Recruiting Operations Program Manager
Baseten · fulltime
IT/Support Operations Engineer (NY)
Baseten · fulltime
Software Engineer - Python, C++, Video Algorithms
NVIDIA · fulltime
Software Engineer - Partner Platform
Baseten · fulltime