You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
NVIDIA

Senior Deep Learning Scientist, Multimodal Agentic RL

  • Office
  • 6+ years

Salary

$184,000 - 287,500/ year

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

NVIDIA is widely regarded as one of the technology industry’s most desirable employers. We lead the way in High-Performance Computing, Artificial Intelligence, and Visualization. Our core invention, the GPU, serves as the visual cortex of modern computers and powers our entire product suite. GPU deep learning ignited the modern AI era—the next great computing age—with the GPU acting as the brain for everything from robots and autonomous cars to conversational AI. Today, we are known globally as "the AI computing company." We are looking to grow our teams by bringing in the smartest people in the world. Join us at the forefront of technological advancement.

NVIDIA is hiring Senior Deep Learning Scientists to advance our efforts in streaming and agentic multimodal AI. You will demonstrate foundational expertise in deep learning, reinforcement learning, and applied mathematics to help develop models capable of reasoning, planning, and acting across diverse modalities. This is a chance to define core algorithmic improvements for multimodal foundation models, scaling your ideas through our Nemotron Omni and VoiceChat platforms. You will work on high-impact, high-visibility large language models and multimodal AI products that improve the experience for millions of users. If you are creative and passionate about solving real-world agentic AI challenges, come join our Nemotron LLM team. For more details on Nemotron LLM, check https://www.nvidia.com/en-us/ai-data-science/foundation-models/nemotron/

What you’ll be doing:

Apply fundamental and applied research to develop, train, fine-tune, and deploy large language models for agentic systems encompassing audio-visual reasoning, tool usage, and document understanding.

Advance post-training and alignment methods including instruction tuning, preference optimization, and RLHF/RLVR/MOPD to improve multimodal agents for complex use cases.

Research and develop agentic reasoning and grounded perception capabilities, focusing on planning, tool execution, and long-horizon task completion across digital and physical environments.

Lead the collection, development, and benchmarking of multimodal datasets, ensuring high-quality evaluation of model accuracy, safety, and task completion success.

What we need to see:

Master’s degree (or equivalent experience) or PhD in Computer Science, AI, or Applied Math with 8+ years of relevant work experience.

Excellent programming skills in Python with strong fundamentals in scalable model development and deep learning frameworks like PyTorch.

Strong knowledge of ML/DL techniques and modern foundation model architectures, including Transformers and mixture-of-experts models.

Foundational understanding of reinforcement learning algorithms and implementation, including MDPs, policies, and reward design.

Hands-on experience in post-training multimodal models for omni-modality (audio-visual) reasoning, full-duplex voice chat, and human-AI interaction.

Proven ability to manage model development life cycles, including dataset versioning, experiment tracking, and evaluation pipelines.

Ways to stand out from the crowd:

Strong record of publications in top-tier AI and machine learning venues such as NeurIPS, ICML, ICLR, or CVPR.

Validated experience training and deploying multimodal foundation models using large-scale distributed infrastructure.

Experience applying deep reinforcement learning techniques to train multimodal agents in complex simulation or gaming environments.

Background in audio/speech AI, especially audio language models or audio generation.

Background in building embodied AI systems that integrate multimodal perception with backend action-fulfillment and long-horizon planning.

With highly competitive salaries and a comprehensive benefits package, NVIDIA is considered one of the industry’s most desirable employers. As you plan your future, see what we can offer you and your family at www.nvidiabenefits.com/ . If you are a creative and autonomous engineer with a genuine passion for state-of-the-art technology, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

Responsibilities

  • Applications for this job will be accepted at least until October 6, 2026.
  • This posting is for an existing vacancy.
  • NVIDIA uses AI tools in its recruiting processes.
  • NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Where you’d work

From the office

About the company

NVIDIA

Office in United States

Also hiring in Munich, Germany, France, Bristol, United Kingdom and 13 more places

273 of their 660 open roles are remote

Your chances

We've checked whether it's still hiring and how crowded it is.

  • 21 checks run
  • 0 red flags

Still hiring?

14 checks

No red flags

How crowded?

7 checks

No red flags

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details12 facts · Role, Location, Compensation, Employment
Tech stack
  • Python
  • PyTorch
  • LLMs
Seniority
Senior
Type
Full-time
Equity
Equity offered
Specialty
ML
Region
United States
Show 6 more factsShow less

Role

Category
Data & Analytics
Specialty
ML
Seniority
Senior
Experience
8+ years
Tech stack
  • Python
  • PyTorch
  • LLMs

Location

Work model
Office
Region
United States
Office
  • United States

Compensation

Salary
$184,000 - 287,500 / year
Pay period
Annual
Equity
Equity offered

Employment

Type
Full-time

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Data & AI Engineer Intern

    Salary by agreement

    • Hybrid · Paris
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    AI Engineer

    $80,000 - 210,000 / year

    • Office · CA, Austin
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior AI/ML Engineer

    $200,000 - 260,000 / year

    • Office · San Francisco
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Member of Technical Staff, ML Infra

    Salary by agreement

    • Office · San Francisco
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason