Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Liquid Ai logo

Member of Technical Staff - GPU Infrastructure Engineer

Liquid Ai
Posted 3 days ago
🇺🇸United States🏢Hybrid📁Engineering & Development
Is this job info correct?

About Liquid AI Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there. The Opportunity Our Cluster Infrastructure team owns the compute environments that power foundation model training and research at Liquid AI. We are looking for a hands-on software engineer to keep our GPU clusters reliable, improve resource efficiency, and build the tooling that allows researchers to focus on model development rather than infrastructure. This role matters because infrastructure issues can delay training by days, while improvements in utilization, storage management, and automation can significantly increase research velocity and reduce compute costs. You will work closely with researchers and infrastructure engineers, owning problems from immediate operational response through long-term platform improvements. What We’re Looking For We need someone who: Brings order to complex systems: You identify root causes and build durable fixes rather than repeatedly firefighting. Is an engineer first: You can go deep across Linux, networking, storage, schedulers, and distributed systems. Balances operations and engineering: You handle urgent issues while steadily replacing manual work with automation. Owns outcomes: You communicate clearly, prioritize effectively, and drive problems to resolution across internal teams and external providers. The Work Own the reliability and operation of the GPU clusters used for training and research. Debug issues across compute, storage, networking, schedulers, and distributed workloads. Improve CPU, GPU, and storage utilization through better tooling and automation. Onboard and migrate workloads across GPU providers and hardware platforms. Build monitoring, validation, and platform abstractions that reduce operational work for researchers. Contribute to the longer-term architecture of Liquid AI’s training infrastructure and GPU platform. Desired Experience Must-have Strong software engineering experience, with the ability to build production-quality infrastructure tooling and automation. Deep knowledge of distributed systems, Linux, networking, and storage. Experience operating a shared compute cluster or distributed training platform. A track record of supporting production users and turning recurring failures into durable solutions. The technical depth to partner effectively with senior research and infrastructure engineers. Nice-to-have Experience with SLURM, Kubernetes, Ray, Hadoop, or another distributed compute platform. Experience supporting GPU, HPC, or large-scale AI training infrastructure. Experience with distributed storage, cluster schedulers, cloud providers, or infrastructure control planes. What Success Looks Like (Year One) Researchers spend less time resolving infrastructure and resource-allocation issues. GPU, CPU, and storage resources are used more efficiently across the fleet. Recurring operational problems are replaced with automation, monitoring, and dependable platform tooling. Liquid AI has the beginnings of a durable internal platform that hides infrastructure complexity from researchers. What We Offer High-impact ownership: Own infrastructure that directly affects how quickly and efficiently we train foundation models. Compensation: Competitive base salary with equity in a unicorn-stage company. Health: We pay 100% of medical, dental, and vision premiums for employees and dependents. Financial: 401(k) matching up to 4% of base pay. Time Off: Unlimited PTO plus company-wide Refill Days throughout the year.

Similar jobs

Similar jobs

Nvidia logo

Senior Software Engineer, Infrastructure and Tooling Lead - Automation

Nvidia

🇺🇸United States8 hours ago
Gray Swan AI logo

Software Engineer, Infrastructure

Gray Swan AI

🇺🇸United States16 hours ago
AA

Infrastructure Engineer

Applied Atomics

🇺🇸United States16 hours ago
Brook & Whittle logo

Senior Infrastructure Engineer

Brook & Whittle

🇺🇸United States23 hours ago
Bold New Solutions - BNS Power logo

Sales Engineer Data Center & AI Infrastructure

Bold New Solutions - BNS Power

🇺🇸United States23 hours ago
Nvidia logo

Senior Firmware Engineer - Development, Verification and Infrastructure

Nvidia

🇺🇸United States23 hours ago