Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
NVIDIA

Senior Deep Learning Software Engineer, Inference and Model Optimization

NVIDIA

Senior Deep Learning Software Engineer at NVIDIA working on generative AI models and optimization techniques. Collaborating with teams to enhance AI software stack performance.

Posted 7/29/2026full-timeRemote • California • 🇺🇸 United StatesSenior💰 $184,000 - $356,500 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in developing and optimizing generative AI models, particularly using NVIDIA's AI software stack and the PyTorch ecosystem. Proficient in performance analysis and debugging, with strong communication skills for effective collaboration in a fast-paced environment.

Highest-signal resume keywords
Generative AI Model DevelopmentDeep Learning ExpertisePython ProficiencyPyTorch FrameworkPerformance Optimization Techniques

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Deep LearningSoftware DesignPerformance AnalysisDebuggingTest DesignAlgorithmsProgramming Fundamentals
Soft Skills
Written CommunicationVerbal CommunicationCollaborationIndependence
Tools & Technologies
NVIDIA AI Software StackTorch 2.0HuggingFace
Industry Keywords
Generative AILLMsDiffusion ModelsGPU Kernel PerformanceInference Performance

Tech Stack

Tools & technologies
PythonPyTorch

About the role

Key responsibilities & impact
  • Train, develop, and deploy state-of-the generative AI models like LLMs and diffusion models using NVIDIA's AI software stack
  • Leverage and build upon the torch 2.0 ecosystem to analyze and extract standardized model graph representation
  • Develop high-performance optimization techniques for inference
  • Collaborate with teams across NVIDIA to use performant kernel implementations within our automated deployment solution
  • Analyze and profile GPU kernel-level performance to identify hardware and software optimization opportunities
  • Continuously innovate on the inference performance

Requirements

What you’ll need
  • Masters, PhD, or equivalent experience in Computer Science, AI, Applied Math, or related field
  • 5+ years of relevant work or research experience in Deep Learning
  • Excellent software design skills, including debugging, performance analysis, and test design
  • Strong proficiency in Python, PyTorch, and related ML tools (e.g. HuggingFace)
  • Strong algorithms and programming fundamentals
  • Good written and verbal communication skills and the ability to work independently and collaboratively in a fast-paced environment

Benefits

Comp & perks
  • Highly competitive salaries
  • Comprehensive benefits package
  • Equity opportunities