Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Nvidia logo

Senior Compute Platform Engineer, LSF

Nvidia
Posted 5 hours ago
🇺🇸United States🏢Hybrid📁Engineering & Development
Is this job info correct?

NVIDIA's silicon does not tape out without the farm behind it. Our EDA compute environment runs millions of cores across federated LSF cells, and every simulation, synthesis run, and timing signoff on our roadmap passes through it. We are consolidating a dual-scheduler estate onto a single LSF platform, and we are looking for an engineer who knows LSF at the level of its internals — not just its configuration files. This is a deep-specialist role. You will be the person the team escalates to when scheduling latency creeps up and nothing in the logs explains why. What you'll be doing: Owning scheduler behavior across 15–25 federated LSF cells, including mbatchd and mbschd tuning, scheduling cycle analysis, and the contention patterns that appear as cells approach host-count ceilings. Diagnosing MultiCluster forwarding problems — remote queue sizing, forwarding policy, cross-cluster pend behavior — where the symptom reported by users is "the farm is slow" rather than a clear failure. Setting the technical design for cell topology and federation as the farm grows, and deciding what belongs in a cell versus what belongs in a new one. Working with our IaC engineer to encode scheduler policy into a config schema that survives contact with MultiCluster, rather than one that looks clean and breaks at scale. Partnering directly with CAD and methodology teams on workloads that break normal assumptions: 500GB+ memory jobs, interactive-versus-batch contention, and tape-out crunch bursts. What we need to see: BS or MS in Computer Science, Computer Engineering, or equivalent experience. 8+ years in HPC or large-scale batch compute, with 5+ years of that on IBM Spectrum LSF. Demonstrated depth in LSF internals — you have debugged scheduler behavior beyond what the documentation covers, and you can explain a scheduling cycle from submission to dispatch. Hands-on MultiCluster experience in a production, multi-site environment. Strong Linux systems fundamentals, system programming languages and scripting in Python, Perl, and shell. Ways to stand out from the crowd: You have worked on LSF as a developer or in escalation engineering, rather than only as a consumer of it. Experience with LSF integration points: esub, eexec, elim, submit wrappers, RTM, or the LSF APIs. Background in semiconductor or EDA compute, where license constraints and job constraints compete. You have migrated a production estate off Slurm, PBS, or Grid Engine without a scheduled outage users noticed. #LI Hybrid Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits . Applications for this job will be accepted at least until August 24, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar jobs

Similar jobs

LD

Staff Software Engineer

Latent Defense

🇺🇸United States8 minutes ago
Strava logo

Senior Engineering Manager, Community Engagement

Strava

🇺🇸United States9 minutes ago
Sargent & Lundy logo

Structural Engineering Senior Consultant - Nuclear

Sargent & Lundy

🇺🇸United States49 minutes ago
Ccs Medical logo

Application Developer

Ccs Medical

🇺🇸United States1 hour ago
Healthequity logo

Sr Tech Integration Manager

Healthequity

🇺🇸United States1 hour ago
Usap logo

Process Automation Engineer - Quickbase

Usap

🇺🇸United States1 hour ago