Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Liquid AI logo

Member of Technical Staff - Inference Systems

Liquid AI
Posted 5 hours ago
🇺🇸United States🏢Hybrid📁Engineering & Development
Is this job info correct?

Liquid AI Job Description Role: Member Of Technical Staff, Infrastructure Department: Research & Engineering Location: Boston Location Type: Hybrid Employment Type: Full-time About Liquid AI Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there. The Opportunity Our inference stack is central to everything we ship. You'll be a core part of the team responsible for the engine layer that runs our models in production and in partner environments, and for the benchmarking infrastructure we use to evaluate our own work and verify what partners bring to us. Day to day, that means working closely with research and product, but also directly with external engineering teams. What We're Looking For We need someone who: Can pick up unfamiliar tools quickly and knows how to assess whether they're worth using. Designs AI benchmarks and holds methodology to a high standard. Cares about inference details, understands the tradeoffs, and checks what changed across the board before calling something done. Doesn’t consider a model port finished until you can prove the outputs are correct. The Work Design and build benchmark suites that cover inference performance, model quality, and knowledge evaluation across different hardware targets. Run external partner verifications: evaluate their solutions against our benchmarks, identify gaps, and clearly deliver findings. Port models like LFM2 onto different runtimes and frameworks, and verify correctness end-to-end. Maintain and extend the inference engine layer built on llama.cpp, ONNX, and MLX as new model architectures emerge from research. Make benchmark results explainable and verifiable, so internal teams and partners can trust and reproduce them independently. Desired Experience Must-have: Hands-on experience with at least one inference framework like llama.cpp, ONNX Runtime, or MLX, going beyond basic usage into internals and modification. Experience designing and building benchmarking pipelines, including methodology, validation, and reproducibility. Strong C++ and Python in performance-sensitive contexts. Solid understanding of inference fundamentals: quantization, decoding strategies, memory layout, and how they interact. Nice-to-have: Experience porting models across runtimes and verifying numerical correctness. Prior work with external partners or clients in a technical validation or evaluation capacity. Familiarity with edge inference targets and the constraints that come with them. What Success Looks Like (Year One) You've ported LFM2 onto multiple runtimes and platforms, you know the model inside out, and new ports take you a fraction of the time they did at the start. You've run multiple partner verifications end-to-end and built enough context to spot weak evaluations quickly and push back with evidence. The benchmark suite covers inference performance and model quality across the platforms we care about, and both internal teams and partners are using it as a reference. What We Offer Compensation: Competitive base salary with equity in a unicorn-stage company Health: We pay 100% of medical, dental, and vision premiums for employees and dependents Financial: 401(k) matching up to 4% of base pay Time Off: Unlimited PTO plus company-wide Refill Days throughout the year

Similar jobs

Similar jobs

d-Matrix logo

Principal System Software Engineer, AI Inference Execution

d-Matrix

🇺🇸United StatesMay 27, 2026, 10:13 PM UTC
Elloe Ai logo

Principal Engineer – Distributed Systems (GPU Edge + Inference)

Elloe Ai

🇺🇸United StatesMay 27, 2026, 9:26 PM UTC
Bbva logo

Cybersecurity Regulatory Lead

Bbva

🇺🇸United States3 hours ago
Takeda logo

AIRx Director, Computational & AI Biologics Design Lead

Takeda

🇺🇸United States3 hours ago
I8Is logo

TIBCO Integration Developer

I8Is

🇺🇸United States3 hours ago
Hpe logo

Senior Routing Specialist

Hpe

🇺🇸United States3 hours ago