Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Lumalabs Ai logo

Research Scientist / Engineer — Multimodal Agent

Lumalabs Ai
Posted May 27, 2026, 9:26 PM UTC
🇺🇸United States🏢Hybrid📁Engineering & Development
Is this job info correct?

About Luma AI: Luma’s mission is to build multimodal AGI. Through our research on video, 3D, and now multimodal models at Luma, we believe that AI needs to be jointly trained over all signal modalities – text, video, audio, images – analogous to the human brain. To advance our mission, we build and operate the full stack end-to-end, spanning foundation models, inference systems, and products. This integrated approach powers technologies like Ray3, which is seeing rapidly growing adoption among Fortune 500 companies across media, entertainment, and advertising. Backed by a recent $900M Series C and our partnership with Humain to build a 2 GW compute supercluster (Project Halo) , our models and the Dream Machine platform are now enabling creatives worldwide to tell some of the most impactful stories of our time. Where You Come In: This is a rare and foundational opportunity to define the future of multimodal AI. You will be at the forefront of building and training large-scale multimodal models, directly impacting how users interact with pixels. This role offers the chance to bridge cutting-edge research with magical, shipped products, working end-to-end on novel problems with no existing playbook. What You'll Do: This opportunity involves both the “science” and “engineering” parts of research, two aspects that are of equal importance. This is a multi-stack opportunity where you will work on the intersection of modeling, data, systems, and evaluation. Modeling : Architect large-scale multimodal agentic models that use reasoning, planning, coding, and tool calling to achieve complex, multi-step multimodal work. Data : Hillclimbing existing tasks and formulating new tasks through data. Design, implement, and run robust data pipelines for constructing, enriching, and filtering massive pixel datasets. Systems : Train large-scale multimodal models on massive datasets and GPU clusters. Evaluation : Define and build novel evaluation frameworks to measure multimodal agents. Who You Are: Strong foundation in machine learning, foundation models and agentic systems. Deep understanding of agentic systems and approaches in LLM/VLM reasoning, coding models, LLM/VLM tool calling. Hands-on experience with PyTorch and large-scale training (distributed, mixed precision, large datasets). What Sets You Apart (Bonus Points): Experience in the following around data, modeling, or evaluation: State-of-the-art foundation models in reasoning State-of-the-art foundation models in coding State-of-the-art foundation models in tool calling State-of-the-art multimodal agents Your application are reviewed by real people. Compensation The base pay range for this role is $250,000 – $450,000 per year. About Luma Luma’s mission is to build unified general intelligence that can generate, understand, and operate in the physical world. We believe that multimodality is critical for intelligence. To go beyond language models and build more aware, capable and useful systems, the next step function change will come from vision. So, we are working on training and scaling up multimodal foundation models for systems that can see and understand, show and explain, and eventually interact with our world to effect change.

Similar jobs

Similar jobs

Nvidia logo

Senior Deep Learning Scientist, Multimodal Agentic RL

Nvidia

🇺🇸United States1 weeks ago
OpenAI logo

Research Scientist - Multimodal Agent, Consumer Devices

OpenAI

🇺🇸United States1 weeks ago
NS

Associate/Journey Salesforce Developer, Information Technology (REMOTE AVAILABLE) [R0151805]

Nshe

🇺🇸United States1 hour ago
VO

Platform Engineer

Volarisgroup

🇺🇸United States1 hour ago
VO

QA Automation Engineer

Volarisgroup

🇺🇸United States1 hour ago
ON

Bswift System Analyst - Remote (EST)

Onedigital

🇺🇸United States1 hour ago