Logo NVIDIA

CUDA LLM Engineer

NVIDIA· Austin, TX· Ref. EXP-2026-0024 4.5 (5,400 reviews)
Easy applyFull-timeHybrid
View without an account · free sign-in only to apply
Salary
$165k – $230k / year
Type
Full-time
Location
Austin, TX · Austin (Texas) · Hybrid
Valid through
11/3/2026

Additional details

Posted
Posted 1 day ago
Applications
34
on Ganloss
Experience
Mid · 3+ years
Education
BS/MS CS, ML, or equivalent experience
Region
Texas
Team
NVIDIA — AI team
Company size
25000+
Founded
1993

Job description

NVIDIA is hiring a CUDA LLM Engineer (Full-time) hybrid. This llm engineer role is part of the AI hiring market in Austin (Texas). Kernel and inference optimizations for LLM serving — Austin AI research lab. Listed compensation: $165k – $230k / year. Stack: CUDA, LLM, Inference, Austin.

LLM Engineer profiles are in high demand in Austin: Austin’s tech scene keeps growing with AI product, robotics, and data roles at scale-ups and enterprise R&D labs.…

The team

NVIDIA — AI team

Responsibilities

  • Deliver production ML/LLM capabilities with measurable quality bars
  • Partner with product and infra on roadmap and reliability
  • Document evals, monitoring, and rollout practices

Requirements

  • 3+ years shipping ML, LLM, or data systems in production
  • Strong Python or TypeScript and clear written communication
  • Comfort with responsible AI and operational excellence

Tech stack

CUDALLMInferenceAustin

Benefits

Health, dental, vision
401(k) match
Equity
Learning budget
Flexible PTO

Hiring process

  1. 1Recruiter screen
  2. 2Technical interview
  3. 3Team conversation
  4. 4Offer

Experience: Mid · 3+ years · Education: BS/MS CS, ML, or equivalent experience

Salary guide — LLM Engineer in Austin

Go further

Prepare your application with our AI career guides — role sheets, resume, interview, and salary benchmarks.

FAQ — CUDA LLM Engineer

What is the salary for CUDA LLM Engineer at NVIDIA?
Listed range: $165k – $230k / year (Full-time). Packages often include PTO, remote options, and equity depending on the company.
Where is the CUDA LLM Engineer role based?
Austin, TX, Austin (Texas). Work mode: Hybrid.
How do I apply to NVIDIA?
Create a free Ganloss candidate account, then apply in one click. Job reference: EXP-2026-0024. Process: Recruiter screen → Technical interview → Team conversation → Offer.
Is NVIDIA hiring other AI roles?
NVIDIA (Hardware & AI) hires across LLM, ML, data, and AI product. See all openings on the Ganloss company page.

Apply in one click

Typical reply within 48h

Your profile is sent directly to NVIDIA's hiring team.

Apply to this job

Create a free candidate account (30 sec) to send your profile to NVIDIA. The job page stays public without signing up.