Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Friendliai logo

Solutions Architect - AI Inference Specialist

Friendliai
Posted May 28, 2026, 3:38 AM UTC
🇺🇸United States🏢Hybrid📁Engineering & Development
Is this job info correct?

About the job FriendliAI is seeking a Solution Architect to assist enterprises in deploying, scaling, and operating generative and agentic AI workloads on FriendliAI infrastructure. You will work directly with customers to solve and implement production-grade applications using our products, such as Serverless Endpoints, Dedicated Endpoints, or Container. Friendli Container is our service that allows customers to download our inference engine as Docker images and deploy it in their chosen environment, such as private clouds or on-premises. Our Friendli Container can be adopted directly to AWS EKS clusters using our EKS add-on product. You will work directly on our customers’ projects, collaborating with their engineering teams to solve AI inference challenges like scaling, orchestration, and monitoring. This is a hands-on, customer-embedded role. If you have worked in DevOps, platform engineering, or SRE for AI applications, this is your ideal position. Key Responsibilities Design and implement large-scale deployment architectures for LLM and multimodal inference Deploy and manage containerized workloads across Kubernetes clusters Diagnose production issues, such as performance bottlenecks, and implement temporary fixes as needed Collaborate with customers’ DevOps teams to integrate FriendliAI’s infrastructure into their CI/CD workflows Develop scripts, Helm charts, and Terraform modules that simplify repeated deployments Contribute field insights to shape our platform reliability, observability, and scaling strategies Lead workshops, technical sessions, or webinars to help customers master infrastructure best practices. Qualifications 3+ years of experience in cloud infrastructure, DevOps, or reliability engineering Bachelor’s or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent Proficiency with Kubernetes, Docker, Terraform, and Helm Strong foundation in distributed systems, networking, and performance tuning Experience with GPU-based computing and generative AI model serving workloads Strong technical background in backend systems or AI tooling Experience operating workloads on AWS, GCP, or OCI Excellent problem-solving and debugging skills in real-world environments Preferred Experience Experience deploying large models (LLMs, diffusion models) on GPUs or clusters Familiarity with inference frameworks (Triton, vLLM, TensorRT, DeepSpeed-Inference) Familiarity with observability stacks (Prometheus, Grafana, Loki, ELK, OTEL) Understanding of networking security and compliance frameworks (e.g., SOC 2) Experience supporting on-prem or hybrid-cloud deployments Benefits A front-row seat to the generative AI infrastructure revolution Competitive compensation and benefits package Daily lunch and dinner provided; unlimited snacks and beverages Health check-up and top-tier hardware support Flexible working hours and a highly collaborative environment About us FriendliAI is building the next-generation AI inference platform that accelerates the deployment of large language and multimodal models with unmatched performance and efficiency. Our infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 500,000 open-source models. We are on a mission to deliver the world’s best platform for AI inference.

Similar jobs

Similar jobs

JG

Machine Learning Specialist - Inference

Jackson Green Recruitment Limited

🇺🇸United States19 hours ago
Agileengine logo

Application Security Engineer (Senior) ID71668

Agileengine

🌍Latin America, United States1 hour ago
Cvshealth logo

Sr Software Development Engineer

Cvshealth

🇺🇸United States3 hours ago
Cvshealth logo

Senior API Platform Engineer (Apigee X/Edge)

Cvshealth

🇺🇸United States3 hours ago
Aptiv logo

Principal Technologist (Austin, TX)

Aptiv

🇺🇸United States3 hours ago
Crowdstrike logo

Platform Professional Services Senior Consultant, NG-SIEM (Remote)

Crowdstrike

🇺🇸United States3 hours ago