Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Tavus logo

Multimodal AI Model Optimization Research Engineer

Tavus
Posted 3 weeks ago
📦Relocation support
🇺🇸United States
📁Engineering & Development
Is this job info correct?

About Us At Tavus, we're building the human layer of AI. Our mission is to make human-AI interaction as natural as face-to-face interaction, enabling the human touch where it has been previously unscalable. We achieve this through pioneering research in multimodal AI for modeling human-to-human communication (language, audio, and video), as well as generating audio-visual avatar behavior. Our models power everything from text-to-video AI avatars to real-time conversational video experiences across industries like healthcare, recruiting, sales, and education. By enabling AI to see, hear, and communicate with human-like authenticity, we're creating the foundation for the next generation of AI employees, assistants, and companions. We are a Series B company backed by top investors, including Sequoia, Y Combinator, and Scale VC. Join us in driving the future of human-AI interaction. The Role We’re looking for an experienced Research Scientist/Engineer with a focus on model optimization to join our core AI team. Our ideal partner-in-crime thrives in startup environments, is comfortable prioritizing independently, and is willing to take calculated risks. We’re moving fast and looking for people who can help pave the path. Your Mission Take cutting-edge research models and make them fast, efficient, and production-ready using sparsification, distillation, and quantization Own the optimization lifecycle for key models: define metrics, run experiments, and benchmark trade-offs across latency, cost, and quality Partner closely with researchers and engineers to turn new ideas into deployable systems Requirements Strong experience in deep learning using PyTorch Hands-on experience with model optimization and compression, including knowledge distillation, pruning/sparsification, quantization, and mixed precision Understanding of efficient architectures such as low-rank adapters Strong understanding of inference performance and GPU/accelerator fundamentals Strong Python coding skills and reliable research engineering practices Experience working with large models and datasets in cloud environments Ability to read ML papers, reproduce results, and adapt ideas Clear communication and collaboration skills Preferred Experience Optimization of diffusion models, video/audio generative models, or large language models Experience with real-time or streaming systems (low-latency APIs, WebRTC, streaming TTS/video) Familiarity with TensorRT, ONNX Runtime, TVM, Triton, or XLA Experience writing custom Triton/CUDA kernels or low-level performance tuning Experience with experiment tracking, benchmarking, and profiling at scale Prior experience in research engineering or applied science roles Location This position is preferably hybrid in San Francisco, with relocation support offered. Remote candidates are also considered. Benefits When you join Tavus, you’re joining a family. We offer flexible work schedules, unlimited PTO, competitive healthcare and gear stipends, and a collaborative environment focused on learning and impact. Culture & Diversity We are not looking for cultural fits — we are looking for culture creators. Diversity drives our success, and we combine varied backgrounds, skills, and perspectives to build the best experiences for our clients..

Similar jobs

Similar jobs

Mistral logo

Applied Scientist/Research Engineer - Palo Alto/NYC

Mistral

🇺🇸United StatesYesterday
Colossal Biosciences logo

Research Associate, Stem Cell Engineering

Colossal Biosciences

🇺🇸United StatesYesterday
NR

Research Engineer - Machine learning applications to power system operations

Nrel

🇺🇸United States2 days ago
Gevernova logo

Research Engineer, Power System

Gevernova

🇺🇸United States2 days ago
Gevernova logo

Research Engineer, Power System

Gevernova

🇺🇸United States2 days ago
Eaton logo

Engineering Specialist - DC Power Systems Research

Eaton

🇺🇸United States3 days ago