Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Photalabs Com logo

ML Engineer - Inference

Photalabs Com
Posted May 27, 2026, 9:28 PM UTC
📦Relocation support🛂Visa sponsorship
🇺🇸United States
📁Data & Analytics
Is this job info correct?

About us: At Phota Labs, we’re building visual GenAI that helps people capture, express, and relive their memories — in ways that feel effortless, personal, and emotionally resonant. Our core technology enables personalized image generation that faithfully reflects who you are and the moments you experienced. Our first goal is to bring visual GenAI into everyday photography. We're a small team of researchers, engineers, and designers who have always been at the forefront of how people capture, edit, and share images and videos. We build with our hands and hearts. We believe GenAI is the next shift for photography, and are seeking builders who share this vision — people like us, like you. We're just getting started! The role: As our first ML Engineer specializing in inference and optimization, you'll bridge the gap between cutting-edge research models and production systems. Your expertise will transform PyTorch research code into highly optimized, low-latency inference solutions that power our user-facing applications. You'll work closely with our GenAI researchers, vision ML engineers, and backend team to deliver exceptional performance. What you’ll do: Deploy and integrate researcher-trained model checkpoints into our cloud infrastructure and production pipelines. Conduct thorough performance profiling and benchmarking to identify and eliminate computational bottlenecks. Implement neural network optimization techniques including quantization, pruning, and architectural refinements while preserving model accuracy. Develop efficient training and fine-tuning strategies with optimal precision trade-offs and parallelism. Build and maintain scalable multi-GPU inference solutions with sophisticated model parallelism and serving architectures. Collaborate with the research team to ensure optimization integrate smoothly with model development workflows. You may be a strong fit if you: Have experience deploying and optimizing deep learning models for production environments, particularly with multi-GPU inference and large-scale model serving. Are well-versed in cutting-edge techniques for optimizing both inference and training workloads. Possess strong knowledge of efficient attention mechanisms and algorithms. Have hands-on experience implementing model quantization and working with inference frameworks. Can write production-quality code and successfully integrate ML models into robust inference pipelines. Are familiar with various cloud platforms, storage solutions, and modern training frameworks. Logistics: This role is based in San Jose, where we work in person. We believe the best ideas come from being in the same room. We sponsor visas. We are committed to working through the process together for the right candidates. If you're currently outside the US, we're also committed to helping you relocate to the US throughout this process. We offer generous health, dental, and vision coverage, unlimited PTO, paid parental leave, and relocation support as needed. Don't meet every single qualification? That’s okay — we care more about your trajectory than checking every box. If the role excites you and the mission resonates, we'd love to hear from you. Note: In the event your application is successful and an offer of employment is made to you, any offer of employment will be conditional on the results of a background check, performed by a third party acting on our behalf.

Similar jobs

Similar jobs

Elorian Ai Inc logo

Inference Infrastructure Engineer, Serving

Elorian Ai Inc

🇺🇸United States1 weeks ago
PU

Software Engineer, Inference

Pulse

🇺🇸United States1 weeks ago
Pulse logo

Software Engineer, Inference

Pulse

🇺🇸United StatesMay 28, 2026, 12:57 AM UTC
Thinking Machines Lab logo

Research Engineer, Infrastructure, Inference

Thinking Machines Lab

🇺🇸United StatesMay 27, 2026, 8:53 PM UTC
The Pokémon Company International logo

Director, Enterprise Data Platform and AI

The Pokémon Company International

🇺🇸United States6 hours ago
Merck logo

Director - Pharmacometrics AI Lead (Remote or Hybrid)

Merck

🇺🇸United States10 hours ago