Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Cantina logo

Machine Learning Engineer, Ops

Cantina
Posted 4 days ago
🌍Europe, United States🏠Remote💰$125.0K–$165.0K📁Data & Analytics
Is this job info correct?

About Cantina: Cantina Labs is a social AI company, developing a suite of advanced real-time models that push the boundaries of expression, personality, and realism. We bring characters to life, transforming how people tell stories, connect, and create. We build and power ecosystems. Cantina, our flagship social AI platform, is just the beginning. If you're excited about the potential AI has to shape human creativity and social interactions, join us in building the future! About the Role: We are looking for an MLOps Engineer to build and scale the inference infrastructure for our generative audio models, including Text-to-Speech (TTS), voice conversion, and Automatic Speech Recognition (ASR). You will be responsible for designing and deploying high-performance systems that ensure low-latency, reliable, and scalable model serving for both streaming and batch inference. This role is central to bridging the gap between research and production, ensuring our audio models are optimized for performance and cost-efficiency as we scale. What You’ll Do: Design and maintain inference infrastructure for generative audio model architectures. Implement and manage high-performance inference engines. Orchestrate service deployments using Kubernetes (K8S), implementing advanced autoscaling paradigms to handle varying traffic loads efficiently. Develop and automate robust CI/CD pipelines to streamline the testing and deployment of model artifacts and inference configurations. Monitor production systems, establishing observability practices to track latency, resource utilization, and overall model performance. Collaborate closely with research teams to optimize model serving paths and evaluate various inference strategies. Optimize inference performance for both streaming and batch applications. What You’ll Bring: Deep understanding of modern audio model architectures (e.g., TTS, ASR) and their specific inference requirements. Strong hands-on experience with Kubernetes (K8S), container orchestration, and implementing autoscaling strategies for production workloads. Solid background in MLOps, including CI/CD automation and managing scalable cloud infrastructure. Proficiency in software engineering principles and experience with Python or Go for infrastructure tooling and backend services. Experience with GPU-accelerated inference and performance profiling techniques. Familiarity with high-performance inference engines (e.g., Triton Inference Server, vLLM-Omni) is a plus. Compensation: The anticipated annual base salary range for this role is between $125,000-$165,000 (€110,000-€145,000). When determining compensation, a number of factors will be considered, including skills, experience, job scope, location, and competitive compensation market data. Benefits for U.S.-based roles: Competitive salary and generous company equity Medical, dental, and vision insurance – 99.99% of premiums covered by Cantina 42 days of paid time off, including: 15 PTO days 10 sick days 15 company holidays 2 floating holidays Generous parental leave & fertility support 401(k) retirement savings plan Lifestyle spending account – $500/month to use however you’d like Complimentary lunch and snacks for in-office employees One Medical membership, and more!

Similar jobs

Similar jobs

Cars logo

Senior DevOps Engineer

Cars

🇺🇸United States6 hours ago
KO

Senior MLOps Engineer (Remote)

Kohls

🇺🇸United States6 hours ago
Costar logo

Matterport - Senior ML Ops Engineer

Costar

🇺🇸United States6 hours ago
DD

Integration and DevSecOps Engineer

Ddcdine

🇺🇸United States6 hours ago
Cloudzero logo

Senior CloudOps Engineer

Cloudzero

🇺🇸United States6 hours ago
AR

DevOps Engineer

Archesys

🇺🇸United States6 hours ago