Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education & Training jobs
  • Remote Healthcare & Nursing jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact mahmoud@relomote.com · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Inferact logo

Member of Technical Staff, Inference

Inferact
Posted 3 hours ago
🌍Singapore, United States🏠Remote📁Engineering & Development
Is this job info correct?

Inferact's mission is to grow vLLM as the world's AI inference engine and accelerate AI progress by making inference cheaper and faster. Founded by the creators and core maintainers of vLLM, we sit at the intersection of models and hardware—a position that took years to build. About the Role We're looking for an inference runtime engineer to push the boundaries of what's possible in LLM and diffusion model serving. Models grow larger. Architectures shift: mixture-of-experts, multimodal, agentic. Every breakthrough demands innovations on the inference engine itself. You'll work at the core of vLLM, optimizing how models execute across diverse hardware and architectures. Your work will directly impact how the world runs AI inference. Skills and Qualifications Minimum qualifications: Bachelor's degree or equivalent experience in computer science, engineering, or similar. Deep understanding of transformer architectures and their variants. Strong programming skills in Python with experience in PyTorch internals. Experience with LLM inference systems (vLLM, TensorRT-LLM, SGLang, TGI). Ability to read and implement model architectures and inference techniques from research papers. Demonstrate the ability to contribute performant and maintainable code and debug in complex ML codebases. Preferred qualifications: Deep understanding of KV-cache memory management, prefix caching, and hybrid model serving. Familiarity with RL frameworks and algorithms for LLMs. Experience with multimodal inference (audio/image/video/text). Contributions to open-source ML or system infrastructure projects. Bonus points if you have: Implemented core features in vLLM or other inference engine projects. Contributed to vLLM integrations (verl, OpenRLHF, Unsloth, LlamaFactory, etc). Written widely-shared technical blogs or side projects on vLLM or LLM inference. Logistics Location: Fully remote, worldwide. We're timezone-flexible but expect regular overlap with Pacific Time for critical syncs. Compensation: We offer competitive compensations (salary + equity) compared to the local market conditions. Visa sponsorship: We sponsor visas on a case-by-case basis. Benefits: Inferact offers competitive benefits appropriate to your location, including health coverage where applicable.

Similar jobs

Similar jobs

Liquid AI logo

Member of Technical Staff - Inference Systems

Liquid AI

🇺🇸United States1 weeks ago
Cohere logo

Lead Member of Technical Staff, Inference Infrastructure

Cohere

🌍Canada, United StatesMay 27, 2026, 10:06 PM UTC
Prime Intellect logo

Member of Technical Staff - Inference

Prime Intellect

🇺🇸United StatesMay 27, 2026, 8:50 PM UTC
Cerebras Systems logo

AI Inference Core - SDET Technical Lead, Release Integration Testing

Cerebras Systems

🌍Canada, United StatesJul 27, 2026, 9:56 PM UTC
Together AI logo

Technical Support Engineer (Inference) - US Weekends

Together AI

🇺🇸United StatesAug 5, 2026, 1:29 AM UTC
Ameriprise logo

Senior Manager-Technology Delivery

Ameriprise

🇺🇸United States31 minutes ago