RE

Germany-Based English & German AI Generalist Trainer (Remote, Full-Time)

RemoExpertsApplies on LinkedInData & Analytics
Salary
$35–$40/hr
Hiring from
Germany
Work type
Remote
Posted
Is this job info correct?
Show job description
Role Summary

Rex.zone is hiring Germany-based, bilingual (English/German) AI Generalist Trainers to support RLHF and large language model evaluation. You will evaluate and rank model outputs, perform QA checks, and write clear rationales that improve training data quality and drive measurable model performance improvement.

About The Role

As a Germany-based English & German AI Generalist Trainer, you will work remotely to evaluate, rank, and QA model-generated responses across multiple task types. Your feedback directly supports LLM evaluation workflows and training data quality initiatives.

Key Responsibilities

  • Evaluate and rank model outputs in both English and German
  • Perform prompt evaluation, response ranking, and rubric-based validation
  • Complete QA evaluation of labeled data and peer work
  • Write concise, structured reasoning/rationales to support RLHF decisions
  • Ensure annotation guidelines compliance and handle edge cases consistently
  • Label content for safety/policy adherence and escalate sensitive items when needed
  • Run training data quality checks to reduce noise and improve consistency
  • Participate in calibration to maintain high inter-annotator agreement

Basic Qualifications

  • Must be based in Germany; remote, full-time role
  • Fluent in German and English (reading and writing)
  • Strong analytical skills and attention to detail
  • Ability to follow annotation guidelines with high consistency
  • Comfort with web-based annotation tools and spreadsheets
  • Reliable internet connection and ability to meet production and QA targets

Preferred Qualifications

  • Experience with data labeling, QA evaluation, or editorial review
  • Familiarity with LLM evaluation concepts (helpfulness, harmlessness, factuality, style)
  • Experience writing structured rationales for ranking decisions
  • Understanding of RLHF/model training feedback loops
  • Self-driven and organized; sound judgment in ambiguous cases

Compensation And Schedule

  • Pay rate: $35–$40 USD per hour (range may vary by project needs)
  • Full-time, remote (Germany-based)

How To Apply

Apply through Rex.zone with an up-to-date resume and a brief summary of your bilingual (English/German) experience. You may be asked to complete a short skills assessment focused on large language model evaluation and training data quality decision-making.

Similar jobs

Apply on LinkedIn