Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
ZA

Applied Research Engineer

Zep AI
Posted 1 weeks ago
🇺🇸United States🏠Remote📁Engineering & Development
Is this job info correct?

Zep is the memory and context layer for AI agents. As a Senior Applied Research Engineer, you'll explore novel approaches to memory, context, and context generation, then own those ideas all the way to production. This is a research role with a hard applied bent. We're not hiring ML researchers chasing publications. We're hiring engineers who can run rigorous experiments, train and evaluate models, and ship the result as production code our customers depend on. How we work We're a small, distributed team that works closely together. We pair on hard problems, review each other's designs, and treat learning as part of the job rather than something that happens after hours. We ask a lot of questions: of customers, of teammates, of our own assumptions. When we find pain, we go fix it. We expect the same back: ask questions early, push back when you disagree, and care about the people on the other end of the API. What you'll do Explore novel approaches to memory, context, and context generation. Define the problem, run the experiments, ship the result. Own research to production end-to-end: dataset creation and curation, experiment design, evaluation, training and finetuning, and production deployment. Train, finetune, and evaluate models on Zep's domain. Build the eval harnesses that catch regressions before they ship. Work with our model serving stack to operate inference at low latency and reasonable cost on AWS. What we're looking for 6+ years of production engineering with a strong backend systems background. You've shipped services with real throughput and latency requirements. Master's in Computer Science or equivalent. Strong research skills: methodology, dataset creation and curation, experiment design, and evaluation. You can frame an open problem and design experiments that actually answer the question. Hands-on experience with model finetuning. Working familiarity with transformer architectures, training and finetuning workflows, and evaluation. PyTorch and OpenAI Triton for experimentation. Working experience with model serving technologies: vLLM, SGLang, or Triton Inference Server. You've operated inference in production. Python, plus high proficiency in one of Rust, C++, or Go. You can work in critical-path code and on performance. Python-only is not enough. Hands-on AWS experience in production: deployments, monitoring, scaling, cost and reliability tradeoffs. Nice to have Published or open-source work in retrieval, memory systems, or LLM evaluation. Tech stack: Python, Rust/C++/Go, PyTorch, vLLM/SGLang, AWS. This role is probably NOT a fit if: You're an ML researcher or model trainer who hasn't shipped research to production. Your background is primarily Python application work without lower-level systems experience. You haven't operated production backend systems with real latency or throughput requirements.

Similar jobs

Similar jobs

Cotiviti logo

Intern - Generative AI Research Engineer

Cotiviti

🇺🇸United States6 hours ago
Nvidia logo

Security Research Engineer, AI Safety and Security Engineering

Nvidia

🇺🇸United States7 hours ago
DA

Senior Machine Learning Research Engineer

David AI

🇺🇸United States9 hours ago
Pinterest logo

Staff Machine Learning Engineer, Applied Research

Pinterest

🇺🇸United States9 hours ago
Mistral logo

Applied Scientist/Research Engineer - Palo Alto/NYC

Mistral

🇺🇸United States20 hours ago
Thomsonreuters logo

Lead Research Engineer

Thomsonreuters

🇺🇸United States20 hours ago