NVIDIA Logo

NVIDIA

Applied AI Engineer

Posted 3 Days Ago
In-Office or Remote
4 Locations
152K-288K Annually
Senior level
In-Office or Remote
4 Locations
152K-288K Annually
Senior level
Architect, build, and deploy LLM-powered AI systems and automation to accelerate post-silicon validation and toolchain workflows; integrate AI across engineering teams, evaluate emerging frameworks, measure impact, and operationalize solutions from prototype to production while supporting silicon bring-up and lab debug.
The summary above was generated by AI

NVIDIA's Silicon Co-Design Group is seeking an Applied AI Engineer to innovate, develop, and integrate innovative AI solutions into the design and automation infrastructure that powers our chips. Every CPU, GPU, and Tegra SoC NVIDIA has shipped in the past four years passed through our toolchain on its way to production — over 200 product SKUs were optimized during the Blackwell generation alone. Now we're rebuilding that toolchain around AI, and we're looking for the engineer to lead that charge. In this role, you will architect and implement solutions that enhance the efficiency, scalability, and intelligence of our workflows, driving initiatives from concept to deployment. If you combine deep technical expertise with a hands-on approach and an aim to push the boundaries of what's possible, this is your opportunity. At NVIDIA, we strive for perfection, encourage innovation, and provide opportunities to explore new ways to succeed!

What you'll be doing:

  • LLM-Powered Validation Pipelines: Design and deploy AI systems that make post-silicon validation faster, smarter, and more scalable across semiconductor environments. You're not maintaining what exists, you're building what comes next.

  • Cross-Team AI Integration: Work directly with multi-functional engineering teams across the organization to identify where AI can eliminate friction, and then build the solution. Your output will be felt across teams, products, and generations of silicon.

  • Technology Scouting & Evaluation: Evaluate emerging AI frameworks and architectures before the rest of the industry catches on. Be the person who spots what's worth adopting, and makes the case for it.

  • Impact Measurement & Continuous Improvement: Build the data systems that prove what's working. Establish clear, quantitative indicators of AI impact, close performance gaps, and drive iteration across the org to turn insight into lasting improvement!

What we need to see:

  • BS, MS, or PhD or equivalent experience in CS, EE, CE, or a related field, with 5+ years of hands-on experience building and deploying ML/AI systems or data-intensive backend services.

  • 2+ years of direct Applied AI experience independently owning an AI agent, LLM-powered workflow, or intelligent automation system end-to-end — from prototype through production deployment.

  • Strong Python skills and proficiency in at least one static language such as C, C++, C#, Java, or Scala.

  • Proven track record with deploying, monitoring, and debugging scalable AI/ML models.

  • Strong EE fundamentals, including computer architecture, high-speed interfaces, timing, power basics, and a solid understanding of firmware/driver structures and hardware interaction.

  • Experience working within a silicon development environment, with exposure to chip and system characterization methodologies.

  • Hands-on experience with silicon bring-up, characterization, or lab debug using standard tools (e.g., oscilloscopes, multimeters, logic analyzers).

  • Proven ability to balance multiple simultaneous projects with excellent problem-solving, communication, and collaboration skills.

Ways to stand out from the crowd:

  • Familiarity with modern AI technologies and methodologies for crafting and launching LLMs with ability to translate innovative AI research into practical, high-impact production tools.

  • Ability to translate innovative AI research into practical, high-impact production tools.

  • Demonstrated experience with deep learning frameworks like PyTorch or TensorFlow, and hands-on experience with agentic and orchestration tools including NeMo Agent Toolkit, LangChain, Semantic Kernel, AutoGen, CrewAI, or n8n.

  • Experience debugging complex system-level issues involving HW/SW interactions, including leadership or ownership in driving root cause analysis of silicon or feature-level issues.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 1, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

HQ

NVIDIA Santa Clara, California, USA Office

2701 San Tomas Expressway, Santa Clara, CA, United States, Santa Clara

NVIDIA San Francisco, California, USA Office

San Francisco, United States

NVIDIA San Jose, California, USA Office

San Jose, United States

Similar Jobs

22 Days Ago
Easy Apply
Remote
United States
Easy Apply
130K-140K Annually
Senior level
130K-140K Annually
Senior level
Artificial Intelligence • Consumer Web • Digital Media • Information Technology • Social Impact • Software
Build end-to-end full-stack AI features on a Ruby on Rails backend and React frontend, design and run experiments (A/B tests and AI evaluations), implement AI infrastructure for reliable LLM inference, and iterate quickly to scale production-ready LLM-powered products.
Top Skills: Ai AgentsLlmsMySQLPostgresReactRetrieval-Augmented Generation (Rag)Ruby On Rails
Yesterday
Remote
United States
Senior level
Senior level
Information Technology • Consulting
Design, prototype, and productionize applied AI capabilities (LLMs, RAG, agents) to improve reasoning, explainability, and developer productivity. Optimize retrieval, embeddings, multi-agent workflows, and evaluation frameworks; build benchmarks and diagnostics. Collaborate with architects and engineers to ingest enterprise knowledge, refine semantic search, and transition experiments into FedRAMP-aligned production for federal healthcare modernization.
Top Skills: AutogenCloud-Native ApplicationsCrewaiDistributed SystemsEmbeddingsLangchainLanggraphLarge Language Models (Llms)LlamaindexPrompt EngineeringPythonRest ApisRetrieval-Augmented Generation (Rag)Semantic KernelSemantic SearchVector Databases
2 Days Ago
Remote
United States
175K-200K Annually
Mid level
175K-200K Annually
Mid level
Blockchain
Owner of company-wide applied AI initiatives: discover high-leverage workflows, buy or build LLM/agent solutions, ship and operate production integrations, ensure security-by-design, drive adoption, and maintain/sunset tools with clear operational ownership.
Top Skills: Agent FrameworksCli ToolingDevOpsEgress ControlIdentity ManagementLangchainLinuxLlmsMcp ServersOpenhandsPythonRustSandboxingSecrets ManagementTypescriptVendor Llm IntegrationsVm Sandboxes

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account