Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
NVIDIA

Senior Compute Platform Engineer, LSF

NVIDIA

Senior Compute Platform Engineer optimizing NVIDIA’s federated LSF infrastructure for semiconductor EDA workloads. Diagnosing scheduler and MultiCluster behavior across large-scale compute cells.

Posted 8/20/2026full-timeSanta Clara • California, Texas • 🇺🇸 United StatesSenior💰 $184,000 - $356,500 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Expertise in High-Performance Computing (HPC) and large-scale batch compute environments, with a strong focus on IBM Spectrum LSF and MultiCluster management. Proficient in system programming and scripting, particularly in Python, Perl, and shell, to optimize scheduler behavior and integration.

Highest-signal resume keywords
IBM Spectrum LSFHigh-Performance Computing (HPC)MultiCluster ManagementSystem Programming in Python, Perl, and ShellLinux Systems Fundamentals

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
IBM Spectrum LSF InternalsMultiCluster ExperienceScheduler Policy ConfigurationLSF Integration PointsContention AnalysisSystem ProgrammingScripting in PythonScripting in PerlShell ScriptingProduction Environment Management
Tools & Technologies
LSF APIsConfiguration SchemaScheduler Behavior TuningRemote Queue SizingCross-Cluster Forwarding
Industry Keywords
SemiconductorEDA ComputeBatch ComputeProduction Estate MigrationContention Patterns

Tech Stack

Tools & technologies
LinuxPerlPython

About the role

Key responsibilities & impact
  • Own scheduler behavior across 15–25 federated LSF cells, including mbatchd and mbschd tuning, scheduling cycle analysis, and contention patterns near host-count ceilings
  • Diagnose MultiCluster forwarding problems, including remote queue sizing, forwarding policy, and cross-cluster pending behavior
  • Set technical design for cell topology and federation as the farm grows
  • Decide what belongs in a cell versus a new cell
  • Work with the IaC engineer to encode scheduler policy into a configuration schema compatible with MultiCluster at scale
  • Partner with CAD and methodology teams on 500GB+ memory jobs, interactive-versus-batch contention, and tape-out crunch bursts

Requirements

What you’ll need
  • BS or MS in Computer Science, Computer Engineering, or equivalent experience
  • 8+ years in HPC or large-scale batch compute
  • 5+ years of experience with IBM Spectrum LSF
  • Demonstrated depth in LSF internals
  • Hands-on MultiCluster experience in a production, multi-site environment
  • Strong Linux systems fundamentals
  • System programming languages and scripting in Python, Perl, and shell
  • Experience with LSF integration points such as esub, eexec, elim, submit wrappers, RTM, or the LSF APIs
  • Background in semiconductor or EDA compute
  • Experience migrating a production estate off Slurm, PBS, or Grid Engine without a scheduled outage users noticed

Benefits

Comp & perks
  • Equity
  • Benefits 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score