FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Senior Compute Platform Engineer, LSF
NVIDIASenior Compute Platform Engineer optimizing NVIDIA’s federated LSF infrastructure for semiconductor EDA workloads. Diagnosing scheduler and MultiCluster behavior across large-scale compute cells.
Posted 8/20/2026full-timeSanta Clara • California, Texas • 🇺🇸 United StatesSenior💰 $184,000 - $356,500 per yearWebsite
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Expertise in High-Performance Computing (HPC) and large-scale batch compute environments, with a strong focus on IBM Spectrum LSF and MultiCluster management. Proficient in system programming and scripting, particularly in Python, Perl, and shell, to optimize scheduler behavior and integration.
Highest-signal resume keywords
IBM Spectrum LSFHigh-Performance Computing (HPC)MultiCluster ManagementSystem Programming in Python, Perl, and ShellLinux Systems Fundamentals
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
IBM Spectrum LSF InternalsMultiCluster ExperienceScheduler Policy ConfigurationLSF Integration PointsContention AnalysisSystem ProgrammingScripting in PythonScripting in PerlShell ScriptingProduction Environment Management
Tools & Technologies
LSF APIsConfiguration SchemaScheduler Behavior TuningRemote Queue SizingCross-Cluster Forwarding
Industry Keywords
SemiconductorEDA ComputeBatch ComputeProduction Estate MigrationContention Patterns
Tech Stack
Tools & technologiesLinuxPerlPython
About the role
Key responsibilities & impact- Own scheduler behavior across 15–25 federated LSF cells, including mbatchd and mbschd tuning, scheduling cycle analysis, and contention patterns near host-count ceilings
- Diagnose MultiCluster forwarding problems, including remote queue sizing, forwarding policy, and cross-cluster pending behavior
- Set technical design for cell topology and federation as the farm grows
- Decide what belongs in a cell versus a new cell
- Work with the IaC engineer to encode scheduler policy into a configuration schema compatible with MultiCluster at scale
- Partner with CAD and methodology teams on 500GB+ memory jobs, interactive-versus-batch contention, and tape-out crunch bursts
Requirements
What you’ll need- BS or MS in Computer Science, Computer Engineering, or equivalent experience
- 8+ years in HPC or large-scale batch compute
- 5+ years of experience with IBM Spectrum LSF
- Demonstrated depth in LSF internals
- Hands-on MultiCluster experience in a production, multi-site environment
- Strong Linux systems fundamentals
- System programming languages and scripting in Python, Perl, and shell
- Experience with LSF integration points such as esub, eexec, elim, submit wrappers, RTM, or the LSF APIs
- Background in semiconductor or EDA compute
- Experience migrating a production estate off Slurm, PBS, or Grid Engine without a scheduled outage users noticed
Benefits
Comp & perks- Equity
- Benefits 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score