NVIDIA Logo

NVIDIA

Senior Engineer, Local AI - Agents and Systems

Reposted One Month Ago
Be an Early Applicant
In-Office
Santa Clara, CA, USA
184K-357K Annually
Senior level
In-Office
Santa Clara, CA, USA
184K-357K Annually
Senior level
Lead engineering to build and optimize local AI agent frameworks and runtimes on Windows for GeForce RTX GPUs. Ensure secure, sandboxed execution, privacy-aware networking, and efficient local LLM inference. Collaborate with research, driver, and open-source communities, mentor engineers, and produce production-ready code and deployment guidelines.
The summary above was generated by AI

Artificial intelligence is shifting from passive help to autonomous, always-on workflows. Our mission is to make this change seamless, efficient, and secure for millions globally. We seek a Senior Engineer to lead technical efforts in deploying advanced AI agent frameworks and local runtimes on Windows and NVIDIA GeForce RTX GPUs. You will guide development so open-source AI agents (such as Nemoclaw and OpenClaw) operate locally, safely, and efficiently on consumer PCs. By combining powerful local inference (Nemotron models) with strong privacy routers and sandboxed execution, you will help develop the foundation of the desktop AI operating system.

What You Will Be Doing:

  • Act as the lead engineer for developing the agent frameworks natively on Windows environments. You will build the technical roadmap to bring always-on, self-evolving AI assistants to GeForce RTX PCs and laptops.

  • Lead the engineering efforts to optimize the agent runtimes for Windows. You will ensure that autonomous agents operate within detailed, policy-based privacy and security frameworks (e.g., handling filesystem access, secure inference routing, and network egress).

  • Partner closely with internal AI research teams, driver teams, and the open-source OpenClaw community. Ensure our consumer hardware provides an excellent ecosystem for autonomous agents.

  • Foster a collaborative engineering culture by mentoring other engineers, establishing guidelines for AI agent deployment, and writing reliable, production-ready code.

What We Need to See:

  • 10+ years of relevant professional software engineering experience, with at least 3+ years in Staff, or Lead Architect role.

  • BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field (or equivalent experience).

  • Deep understanding of Windows OS internals, process isolation, sandboxing technologies, and system-level security architecture.

  • Proven understanding of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience running local models on consumer-grade hardware.

  • Practical experience with modern AI orchestration and agentic frameworks (e.g., OpenClaw, Hermes, LangChain) and an understanding of how multi-agent systems plan, act, and use tools.

  • Proficiency in multiple languages, particularly C++ (for performance-critical systems/OS integration) and Python (for AI/blueprint logic).

  • Experience building virtualization, containerization, or robust sandboxing tools natively for the Windows ecosystem.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 28, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

2 Minutes Ago
In-Office
120K-162K Annually
Mid level
120K-162K Annually
Mid level
Fintech • Real Estate • PropTech
Build, maintain, and optimize complex batch and streaming data pipelines across structured and unstructured sources. Partner with Engineering and Analytics teams to develop data exchange systems, translate business needs into technical solutions, ensure data and code quality, automate processes, support development environments, and participate in weekly on-call rotations.
Top Skills: AirflowSparkAPIsAWSAzureDockerGCPGitJavaJenkinsKubernetesMySQLPostgresPythonSnowflakeSQL
2 Minutes Ago
Hybrid
139K-170K Annually
Mid level
139K-170K Annually
Mid level
Fintech • Real Estate • PropTech
Build and maintain scalable, high-performance full-stack applications for Redfin’s personalized home-search experience using React, Java, AWS, and AI coding tools. Collaborate with engineering and product stakeholders, lead projects from design through delivery, balance infrastructure and business needs, support production systems through possible on-call rotations, and mentor junior engineers.
Top Skills: Anthropic Claude CodeAWSCursorGithub CopilotJavaReact
3 Minutes Ago
In-Office
96K-174K Annually
Senior level
96K-174K Annually
Senior level
Other • Utilities
Lead and own workstreams on high-impact strategy projects: define hypotheses, run quantitative and qualitative analyses, build financial and scenario models, synthesize insights, create executive-ready presentations, and partner cross-functionally to inform strategic decisions and recommendations for growth, partnerships, and product portfolio evolution.

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account