Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
BN

Applied AI/ML Engineer

Boundless Networks Inc
Posted 7 hours ago
🇺🇸United States🏠Remote💰$250.0K📁Data & Analytics
Is this job info correct?

Boundless is coordinating GPU compute at scale and building toward becoming a leader in AI. As an Applied AI/ML Engineer, you'll ship AI-powered products end-to-end on top of our growing GPU inference fleet — owning everything from serving low-latency inference to standing up reinforcement-learning post-training pipelines. This is a builder's role: you take an idea from prototype to production, tune it for throughput and cost on real GPUs, and iterate fast on customer and internal feedback. You should be comfortable operating with a high degree of autonomy, navigating ambiguity, and defaulting to a strong bias for action. What You'll Do End-to-End AI Product Delivery : Own AI features and products from prototype through production — model selection, serving, evaluation, and iteration — shipping working software rather than research artifacts. Inference Serving: Deploy and optimize LLM inference across the fleet using vLLM and SGLang. Tune continuous batching, KV-cache management, quantization, speculative decoding, and multi-model routing to maximize throughput and minimize latency and cost per token. RL & Post-Training Harnesses : Build and operate reinforcement-learning and post-training pipelines using slime (Megatron-LM + SGLang) and Prime Intellect (prime-rl + the Environments Hub / verifiers). This includes reward and verifier design, rollout orchestration, weight synchronization, and keeping long-running training stable. Evaluation & Iteration: Build eval harnesses and benchmarks that measure quality, throughput, and cost together, and use them to drive fast, data-informed iteration. Work Across the Stack: Partner with Infrastructure on GPU scheduling and fleet utilization, and with Product on what to build next and why. 3+ years shipping ML/AI systems to production Hands-on experience serving LLM inference with vLLM, SGLang, or TensorRT-LLM Experience with RL / post-training methods (GRPO, PPO, DPO, or SFT), or strong adjacent experience and a clear desire to go deep here Strong Python and PyTorch Working understanding of GPU execution: batching, memory, and basic CUDA concepts Comfort operating in ambiguity with a strong bias for action Nice to Have Direct experience with slime, prime-rl, the verifiers library, or Megatron-LM Distributed training experience (FSDP, TP/PP/DP parallelism) Quantization (FP8/INT8), P/D disaggregation, or speculative decoding Experience with verifiable inference or large-scale distributed systems Kubernetes and container-based deployment Familiarity with GPU fleet orchestration (Ray, SkyPilot, Slurm) Additional Requirements Candidates must include a public GitHub profile in their application. The GitHub profile should demonstrate a minimum of 1 year of activity/history. Applications that do not include a GitHub profile, or show insufficient activity, will not be considered. At Boundless, we take care of our people, because building the future of AI compute starts with an empowered team. Here's what you can expect when you join us: Competitive salary (proposed band b/t US$175k and $250k annually) + equity allocation Health, dental, vision (for U.S. employees; region-adjusted globally) Flexible PTO Professional development and conference travel budget Remote-first with regular off-sites and a high-trust, high-velocity team environment We are a global team, and applicants from around the world are welcome to apply.

Similar jobs

Similar jobs

Gdit logo

AI/ML Delivery Engineer

Gdit

🇺🇸United States6 hours ago
GE

Senior ML Systems Engineer

Generalmotors

🇺🇸United States7 hours ago
GE

ML Systems Engineer

Generalmotors

🇺🇸United States7 hours ago
Chooch logo

Backend Software Engineer - Web/ML Dev Ops

Chooch

🇺🇸United States7 hours ago
A2Z Sync logo

Senior Engineer, AI/ML

A2Z Sync

🇺🇸United States7 hours ago
NMDP logo

AI/ML Engineer

NMDP

🇺🇸United States7 hours ago