Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Lifted, an Upwork Company™ logo

LLM Model Response Evaluation

Lifted, an Upwork Company™
Posted Yesterday
🇺🇸United States🏠Remote📁Data & Analytics
Is this job info correct?

Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria. Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities. The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities. The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more. Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions. Each task will include detailed project guidelines within the evaluation platform. 3+ years of hands-on experience in LLM / GenAI data evaluation. Master's or PhD required (PhD candidates strongly preferred). Ability to research unfamiliar topics using trusted sources and make well-supported judgments. Comfortable evaluating content across multiple modalities Flexible and remote work Variable workload: Accept or decline tasks based on your availability No guaranteed hours: Workload may vary weekly

Similar jobs

Similar jobs

Elsevier logo

Faculty Nurse Educator - AI Response Evaluation (Part-Time, Fixed Term Contract)

Elsevier

🇺🇸United StatesMay 28, 2026, 12:03 AM UTC
Elsevier logo

Clinical Practice Nurse Educator - AI Response Evaluation (Part-Time, Fixed Term Contract)

Elsevier

🇺🇸United StatesMay 28, 2026, 12:03 AM UTC
Mercor logo

AI Red Team Specialist - Remote | Upto $22/hr

Mercor

🌍Australia, Canada, India, United Arab Emirates, United Kingdom, United States1 hour ago
Mercor logo

AI Safety Specialist - Fully Remote | Upto $22/hr

Mercor

🌍Asia, Australia, Canada, United Kingdom, United States1 hour ago
WI

Data Warehouse Engineer

Wisconsin

🇺🇸United States5 hours ago
Bms logo

Senior Manager, AI and Data Analytics

Bms

🇺🇸United States5 hours ago