NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / AI Engineer in United States of America
10 days ago
Apply with autofill
Apply with autofill
NVIDIA·Semiconductors·10 days ago
10 days ago

Senior High Performance AI Engineer, Agentic AI

Santa Clara, United States of AmericaFull-timeRemoteMid · 5-8 yearsAI Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at NVIDIA

How you compare FREE

?
Your scoreYour score: not yet known
→
70
Top 10%Top 10%: 70 out of 100

Top 10% of NextRaise users matched against AI Engineer roles in United States.

Must-have skills for this role

  • c++
  • python
  • cuda
  • gpu programming

PDF or DOCX · no account needed

Apply faster with autofill FREEThe NextRaise extension autofills your application in one click.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Design, build, and optimize agentic AI systems for the CUDA ecosystem, including agent architectures, multi-agent workflows, tool use, memory, and orchestration.
  • Build the data, environments, verification, reward, and evaluation systems needed to continuously improve model and agent capabilities.
  • Co-design and optimize agentic systems across NVIDIA's software and hardware stack, from models and inference through compilers, runtimes, libraries, kernels, and GPUs.

What they're looking for

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience; MS or PhD preferred.
  • 4+ years of relevant industry or academic experience in AI systems, machine learning, compilers, high-performance computing, or related areas.
  • Hands-on experience in one or more of the following: agent systems, coding agents, reinforcement learning, or modern AI inference systems.
  • Strong C/C++, Rust, and Python programming skills, with solid software engineering fundamentals.
  • Experience with GPU programming and performance optimization using CUDA or comparable accelerator platforms, and the ability to work effectively across system boundaries.

Nice to have

  • Track record of building high-impact coding agents, autonomous software-engineering systems, or developer tools.
  • Hands-on experience optimizing and deploying with TRT-LLM, SGLang, vLLM, or Transformer Engine.
  • Deep expertise in GPU systems and performance optimization, demonstrated through benchmark results, deployed systems, publications, or widely used software.
  • Publications or open-source leadership in deep learning, agentic AI, reinforcement learning, compilers, high-performance computing, or AI systems.

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

We are looking for outstanding Senior High Performance AI Engineers to build the next generation of agentic AI systems for the CUDA ecosystem. Our team works across the full agentic AI stack—from training and improving models, to designing agent architectures and multi-agent systems, to building the systems, runtimes, and evaluation frameworks that make them effective at scale.


You will help develop intelligent agentic systems that can reason about, generate, optimize, and operate across NVIDIA's accelerated computing stack. This includes advancing model and agent capabilities, building scalable agent systems and runtimes, developing high-fidelity training and evaluation environments, and co-designing with NVIDIA's libraries, runtimes, compilers, and hardware. You will collaborate closely with internal NVIDIA software, model, and hardware teams to bring new capabilities into NVIDIA products.


What you'll be doing:

  • Design, build, and optimize agentic AI systems for the CUDA ecosystem, including agent architectures, multi-agent workflows, tool use, memory, and orchestration.
  • Build the data, environments, verification, reward, and evaluation systems needed to continuously improve model and agent capabilities.
  • Co-design and optimize agentic systems across NVIDIA's software and hardware stack, from models and inference through compilers, runtimes, libraries, kernels, and GPUs.

What we need to see:

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience; MS or PhD preferred.
  • 4+ years of relevant industry or academic experience in AI systems, machine learning, compilers, high-performance computing, or related areas.
  • Hands-on experience in one or more of the following: agent systems, coding agents, reinforcement learning, or modern AI inference systems.
  • Strong C/C++, Rust, and Python programming skills, with solid software engineering fundamentals.
  • Experience with GPU programming and performance optimization using CUDA or comparable accelerator platforms, and the ability to work effectively across system boundaries.

Ways To Stand Out From The Crowd:

  • Track record of building high-impact coding agents, autonomous software-engineering systems, or developer tools.
  • Hands-on experience optimizing and deploying with TRT-LLM, SGLang, vLLM, or Transformer Engine.
  • Deep expertise in GPU systems and performance optimization, demonstrated through benchmark results, deployed systems, publications, or widely used software.
  • Publications or open-source leadership in deep learning, agentic AI, reinforcement learning, compilers, high-performance computing, or AI systems.

With highly competitive salaries and a comprehensive benefits package, NVIDIA is widely considered to be one of the technology industry's most desirable employers. We have some of the most brilliant and hardworking people in the world working with us and our product lines are growing fast in some of the hottest state of the art fields such as Virtual Reality, Artificial Intelligence, Deep Learning and Autonomous Vehicles.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 15, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Semiconductors

Company

NVIDIASemiconductors
Santa Clara, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from NVIDIA's careers site·first seen 11 Sept 2026·last verified 11 Sept 2026·How we source jobs

Similar jobs

  • Senior Founding AI Engineer at cleraSan Francisco, United States of America–match not yet calculated
  • Sr. AI Engineer, CX at gametimeunitedUnited States - Remote–match not yet calculated
  • AI Engineer (On-Site, Indiana Only) at alliedsolutionsCarmel, United States of America–match not yet calculated
  • GTM AI Engineer -Deal Desk at gomotiveUnited States - Remote–match not yet calculated
  • Sr. Engineer - GenAi at skechersManhattan Beach, United States of America–match not yet calculated

Browse more jobs

  • AI Engineer jobs in United States
  • Machine Learning Engineer jobs in United States
  • AI / ML Researcher jobs in United States
  • Computer Vision Engineer jobs in United States
  • AI Engineer jobs in India
  • AI Engineer jobs in Germany