NVIDIA Logo

NVIDIA

Senior Engineer - AI Agents and Systems

Posted 22 Days Ago
Be an Early Applicant
In-Office
Redmond, WA, USA
224K-431K Annually
Senior level
In-Office
Redmond, WA, USA
224K-431K Annually
Senior level
Build and optimize local LLM inference and agent runtimes for Windows and GeForce RTX GPUs. Profile and reduce latency/memory, integrate CUDA/TensorRT pipelines, implement sandboxed security and privacy routers, and harden multi-agent orchestration for constrained consumer hardware while collaborating across research, driver, and open-source teams.
The summary above was generated by AI

Artificial intelligence is moving from passive assistance to autonomous, always-on agentic workflows. Our mission is to make this transition flawless, high-performing, and secure for millions of users worldwide, running natively on the GPUs already sitting in their PCs.

We are looking for a Senior Software Engineer to build and optimize the local runtimes and agent frameworks that bring autonomous AI to Windows and NVIDIA GeForce RTX GPUs. You will be a hands-on individual contributor responsible for making open-source AI agents (like NemoClaw and OpenClaw) run locally, safely, and efficiently on consumer PCs. By combining high-performance local inference (Nemotron models) with robust privacy routers and sandboxed execution, you will help build the foundation of the desktop AI operating system. This is a deeply technical, code-first role. You will spend your days profiling inference pipelines, squeezing latency and memory out of local models, and hardening agent runtimes.

What you will be doing:

  • Local Inference Optimization: Optimize performance of local LLMs (Nemotron and others) on GeForce RTX hardware. Profile and optimize inference across Ollama, llama.cpp, and vLLM, minimizing latency and memory footprint using TensorRT and CUDA.

  • Agent Runtime Engineering: Build and optimize agentic harnesses (NemoClaw, OpenClaw) to run natively and reliably on Windows. Implement the orchestration logic that lets multi-agent systems plan, act, and use tools efficiently on constrained consumer hardware.

  • Sandboxing & Security: Implement policy-based privacy and security frameworks for autonomous agents, handling filesystem access, secure inference routing, and network egress within thorough sandboxed execution environments.

  • Hardware/Software Integration: Work close to the metal, integrating agent and inference stacks with NVIDIA's driver and middleware layers to extract maximum performance from RTX GPUs.

  • Cross-Team Collaboration: Partner with internal AI research teams, driver teams, and the open-source OpenClaw community to ensure our consumer hardware is the best possible platform for local agents.

  • Code Quality: Write reliable, production-ready code, contribute to engineering best practices, and raise the technical bar through code review and design input.

What we need to see:

  • Experience: 12+ years of relevant professional software engineering experience, with a track record of shipping performance-critical systems.

  • Education: BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field (or equivalent experience).

  • AI & GPU Infrastructure: Hands-on experience with LLM inference pipelines (Ollama, llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and running local models on consumer-grade hardware.

  • Agentic Frameworks: Practical experience with modern agentic frameworks (e.g., OpenClaw, LangChain, AutoGPT) and a working understanding of how multi-agent systems plan, act, and use tools.

  • Systems & OS Knowledge: Strong understanding of Windows OS internals, process isolation, sandboxing technologies, and system-level security.

  • Programming Languages: Proficiency in C++ (performance-critical systems and OS integration), Python (AI and orchestration logic), and TypeScript (agent plugins and tooling).

  • Communication: Ability to translate complex technical decisions into clear documentation and collaborate effectively across diverse engineering teams.

Ways to stand out from the crowd:

  • Demonstrated open-source contributions to AI agent platforms or inference/orchestration tools (especially OpenClaw or llama.cpp).

  • Deep knowledge of NVIDIA GeForce RTX architecture and its specific constraints and advantages for edge AI.

  • Experience building virtualization, containerization, or sandboxing tools natively for Windows.

  • Active technical community presence (blogs, talks, whitepapers) at the intersection of AI, security, and local compute.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most talented people on the planet working for us. As part of our team, you will have the opportunity to influence the future with your vision and expertise. Are you creative? Are you driven not just by data or the need to know why, but yearn to ask, 'why not'? We want to hear from you.

With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 5, and 272,000 USD - 431,250 USD for Level 6.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 10, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Seattle, Washington, USA Office

4545 Roosevelt Way NE 6th Floor, Seattle, Washington, United States, 98105

NVIDIA Bellevue, Washington, USA Office

Bellevue, United States

NVIDIA Redmond, Washington, USA Office

Redmond, United States

Similar Jobs

4 Hours Ago
In-Office
Redmond, WA, USA
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead development of local AI agent frameworks and runtimes on Windows for GeForce RTX PCs. Optimize GPU-accelerated LLM inference, enforce policy-based privacy and sandboxing, collaborate with research, driver, and open-source communities, mentor engineers, and produce production-ready code and deployment guidelines.
Top Skills: C++ContainerizationCudaGpu-Accelerated ComputingHermesLangchainLlama.CppNemoclawNemotronOllamaOpenclawPythonSandboxingTensorrtVirtualizationVllmWindows
4 Days Ago
In-Office
Redmond, WA, USA
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead engineering to build and optimize local AI agent frameworks and runtimes on Windows for GeForce RTX GPUs. Ensure secure, sandboxed execution, privacy-aware networking, and efficient local LLM inference. Collaborate with research, driver, and open-source communities, mentor engineers, and produce production-ready code and deployment guidelines.
Top Skills: C++ContainerizationCudaDevice DriversGeforce RtxHermesLangchainLlama.CppLlm InferenceNemoclawNemotronOllamaOpenclawProcess IsolationPythonSandboxingTensorrtVirtualizationVllmWindows
Yesterday
In-Office
Redmond, WA, USA
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead development of AI agent frameworks for Windows, optimizing performance, ensuring security, and collaborating with AI research teams. Mentor engineers and establish best practices.
Top Skills: C++CudaHermesLangchainLlamacppOllamaOpenclawPythonTensorrtVllm

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account