FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Deep Learning Algorithm Engineer
NVIDIASenior engineer advancing NVIDIA’s Dynamo open-source distributed inference platform for large-scale, low-latency AI services. Optimizing runtimes, scheduling, networking, and ML inference performance.
Posted 8/5/2026full-timeSanta Clara • California • 🇺🇸 United StatesSenior💰 $152,000 - $287,500 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and optimizing AI inference systems, with strong programming skills in Python, Rust, and C++. Proven ability to collaborate across teams and engage with open-source communities to enhance performance and reliability.
Highest-signal resume keywords
Python ProgrammingRust ProgrammingC++ ProgrammingAI Inference OptimizationOpen-Source Contributions
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Dynamo IntegrationsPerformance ProfilingDebugging Distributed SystemsML ArchitecturesInference TechniquesLatency OptimizationThroughput ImprovementReliability EnhancementAutoscalingKV Caching
Soft Skills
High AgencyLeadership in Ambiguous Work
Tools & Technologies
AI AcceleratorsOpen Source Frameworks
Industry Keywords
Machine LearningDistributed SystemsPerformance-Critical SystemsResearch in ML Inference
Tech Stack
Tools & technologiesDistributed SystemsOpen SourcePythonRust
About the role
Key responsibilities & impact- Design, build, and maintain Dynamo integrations for open source frameworks vLLM, SGLang, and TRTLLM
- Partner with open source communities to improve latency, throughput, reliability, and efficiency
- Showcase NVIDIA token/watt leadership on public and private benchmarks
- Find and remove bottlenecks across runtimes, kernels, networking, routing, and orchestration
- Develop inference optimizations for scheduling, disaggregation, KV caching, and autoscaling
- Collaborate across research, software, systems, and hardware teams
- Engage with the broader ecosystem and external partners to advance AI inference deployment
Requirements
What you’ll need- BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field, or equivalent experience
- 3+ years building, profiling, and debugging performance-critical distributed or ML systems
- Strong programming skills in Python and/or Rust, and C++
- Understanding of modern ML architectures and inference techniques
- High agency and track record of leading ambiguous work end to end
- Experience with AI accelerators
- Open-source contributions or leadership
- Research in ML inference or distributed systems
Benefits
Comp & perks- Equity
- Benefits 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score