NVIDIA logo

Senior Engineer, Local AI - Agents and Systems

NVIDIA

Santa Clara, CAFull-time$184–357K/yrPosted 1mo agoVerified open 5 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

At a glance

Compensation
$184–357K/yr
Location
Santa Clara, CA
Schedule
Full-time
Work Authorization
Not specified

Requirements

Credentials this posting asks for.

Doctorate

Job overview

NVIDIA is hiring a Senior Engineer, Local AI - Agents and Systems. NVIDIA is advancing AI from passive assistance to autonomous, always‑on workflows. The company seeks a Senior Engineer to lead the development of agent frameworks and local runtimes on Windows and GeForce RTX GPUs, ensuring privacy‑focused, sandboxed execution on consumer PCs. The role involves building roadmaps, optimizing runtimes, collaborating with research and driver teams, and mentoring engineers to create a desktop AI operating system.

Key focus areas include Lead development of agent frameworks natively on Windows, creating roadmaps for always‑on AI assistants on RTX PCs., Optimize Windows agent runtimes to operate within policy‑based privacy and security frameworks, handling filesystem and network controls., and Partner with internal AI research, driver, and OpenClaw community teams to ensure robust ecosystem support for autonomous agents..

Important skills include Windows OS Internals, Process Isolation, Sandboxing Technologies, System-Level Security Architecture, LLM Inference Pipelines, and Ollama. Preferred (not required): Nemoclaw, OpenClaw, Nemotron Models, and Coaching.

Skills & qualifications

RequiredNice to have

Skills

Windows OS InternalsProcess IsolationSandboxing TechnologiesSystem-Level Security ArchitectureLLM Inference PipelinesOllamaLlama.cppvLLMGPU-Accelerated ComputingCUDATensorRTRunning Local Models on Consumer-Grade HardwareModern AI OrchestrationAgentic FrameworksHermesLangChainMulti-Agent SystemsC++PythonBuilding VirtualizationContainerizationRobust Sandboxing ToolsNemoclawOpenClawNemotron ModelsCoachingEstablishing Guidelines for AI Agent DeploymentWriting Reliable, Production-Ready Code

Qualifications

10+ Years Professional Software Engineering Experience3+ Years Staff, or Lead Architect RoleBS, MS, or PhD in Computer Science, Computer Engineering, or Related Technical Field or Equivalent Experience

Full job description

Artificial intelligence is shifting from passive help to autonomous, always-on workflows. Our mission is to make this change seamless, efficient, and secure for millions globally. We seek a Senior Engineer to lead technical efforts in deploying advanced AI agent frameworks and local runtimes on Windows and NVIDIA GeForce RTX GPUs. You will guide development so open-source AI agents (such as Nemoclaw and OpenClaw) operate locally, safely, and efficiently on consumer PCs. By combining powerful local inference (Nemotron models) with strong privacy routers and sandboxed execution, you will help develop the foundation of the desktop AI operating system.

What You Will Be Doing:

  • Act as the lead engineer for developing the agent frameworks natively on Windows environments. You will build the technical roadmap to bring always-on, self-evolving AI assistants to GeForce RTX PCs and laptops.

  • Lead the engineering efforts to optimize the agent runtimes for Windows. You will ensure that autonomous agents operate within detailed, policy-based privacy and security frameworks (e.g., handling filesystem access, secure inference routing, and network egress).

  • Partner closely with internal AI research teams, driver teams, and the open-source OpenClaw community. Ensure our consumer hardware provides an excellent ecosystem for autonomous agents.

  • Foster a collaborative engineering culture by mentoring other engineers, establishing guidelines for AI agent deployment, and writing reliable, production-ready code.

What We Need to See:

  • 10+ years of relevant professional software engineering experience, with at least 3+ years in Staff, or Lead Architect role.

  • BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field (or equivalent experience).

  • Deep understanding of Windows OS internals, process isolation, sandboxing technologies, and system-level security architecture.

  • Proven understanding of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience running local models on consumer-grade hardware.

  • Practical experience with modern AI orchestration and agentic frameworks (e.g., OpenClaw, Hermes, LangChain) and an understanding of how multi-agent systems plan, act, and use tools.

  • Proficiency in multiple languages, particularly C++ (for performance-critical systems/OS integration) and Python (for AI/blueprint logic).

  • Experience building virtualization, containerization, or robust sandboxing tools natively for the Windows ecosystem.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits .

Applications for this job will be accepted at least until July 31, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

You've read the whole posting — now see how you match it.