Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
ITRex Group logo

AI Inference Engineer (f/m/d)

ITRex Group
Posted 3 weeks ago
🌍Europe, Turkey🏠Remote📁Engineering & Development
Is this job info correct?

About ITRex THE PLACE ITRex - AI pioneers who build systems that actually work in the real world, not just in demos. We're 250+ people spread across the US and Europe, creating solutions for companies like Procter & Gamble and Shutterstock. We keep it simple, build it right, and focus on what works. THE PEOPLE We're the kind of people who don't ignore messages in Slack, who jump in to help when you're stuck on a problem, and who offer solutions instead of blame when things go sideways. We believe in openness, accountability, and having each other's backs. No office politics, no hidden agendas - just people who care about doing good work together and supporting each other to get there. THE ROLE We are looking for a strong C++ Engineer with hands-on experience deploying and optimizing modern AI models for production. The ideal candidate combines deep systems programming expertise with practical experience working with LLMs and modern deep learning architectures. Rather than building or training models from scratch, this engineer focuses on integrating, evaluating, profiling, and optimizing AI inference pipelines for high-performance on-device execution. Responsibilities Work on deploying machine learning models to edge devices using the frameworks: llama.cpp, ggml Collaborate closely with researchers to assist in coding, training and transitioning models from research to production environments Integrate AI features into existing products, enriching them with the latest advancements in machine learning Core Software Engineering 4+ years of professional experience in Modern C++ (C++17/20) Strong knowledge of memory management, multithreading, profiling and performance optimization Experience debugging low-level issues (memory leaks, fragmentation, OOM, concurrency) Experience working with Linux development environments AI Inference / ML Systems Experience integrating machine learning models into production applications Experience deploying and optimizing AI inference pipelines Hands-on experience with AI inference frameworks such as: llama.cpp (strong plus), ggml (strong plus), ONNX Runtime, TensorRT / TensorRT-LLM, OpenVINO, MLC LLM, ExecuTorch, TVM Experience profiling inference performance and optimizing memory usage and latency Deep Learning Knowledge Strong understanding of modern AI model architectures, including: Transformer architecture Large Language Models (LLMs) Diffusion Models Tokenization Attention mechanisms KV Cache Quantization techniques Model conversion and deployment Practical AI Experience Experience working with one or more of the following: LLM deployment, Computer Vision models, OCR models, Multimodal models, Speech models, Image generation models Experience evaluating new models and integrating them into existing products is highly desirable Nice to Have CUDA Vulkan Compute Metal OpenCL Typescript Python Experience contributing to open-source AI infrastructure projects Why people stay First, the foundation: Remote flexibility : Work where and how you work best - we trust you to deliver Fair compensation : Competitive salary + benefits that matter (medical, learning) Then, the growth: Ownership opportunities: See a problem worth solving? Own it. We back smart risks over bureaucratic safety AI enhancement : We leverage AI to make you faster and stronger - complementing your abilities, not replacing them Learning investment : English classes, professional development Career progression : Real paths up, not just sideways shuffling Finally, the people: Responsive teammates : No ignored Slacks, no "not my problem" attitudes Supportive culture : When you're stuck, people help. When things break, we fix them together Human connections : Regular meetups, tech talks, and actual relationships beyond work Curious? We are too. Let's talk

Similar jobs

Similar jobs

Cast AI logo

Senior ML Engineer - Kimchi (LLM Inference Optimization)

Cast AI

🌍EuropeJun 4, 2026, 6:37 AM UTC
Nebius logo

Senior Site Reliability Engineer — Token Factory (Inference Platform)

Nebius

🌍EuropeMay 28, 2026, 12:34 AM UTC
Tether Operations Limited logo

AI Inference Engineer QVAC (100% remote Worldwide)

Tether Operations Limited

🌍Argentina, Colombia, Europe, India, Israel, Pakistan, United Arab Emirates, Uruguay, VietnamMay 27, 2026, 7:36 PM UTC
Tether Operations Limited logo

AI Research Engineer (Kernel & Inference Optimization) - 100% Remote Worldwide

Tether Operations Limited

🌍Asia, Europe, Israel, Latin America, Turkey, United Arab EmiratesMay 27, 2026, 7:36 PM UTC
VO

Precon / FEP Lead

Volta

🌍Europe5 hours ago
VO

Global Electrical Design Manager

Volta

🌍Europe5 hours ago