Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
2M

Remote | AI Safety Evaluation Specialist — $55–$65/hour

24 Mag
Posted 2 hours ago
🇺🇸United States🏠Remote💰$55.0–$65.0/hr📁Data & Analytics
Is this job info correct?

We are sharing a specialised part-time consulting opportunity for experienced AI safety, trust and safety, public policy, journalism, scientific research, security, and content-evaluation professionals with strong judgment across complex and policy-sensitive subject matter. This role supports a frontier AI initiative focused on evaluating the safety, quality, factual accuracy, and alignment of advanced models. Selected professionals will review AI-generated responses across sensitive and ambiguous scenarios, apply structured safety policies and rubrics, identify behavioural failures, and provide detailed feedback that supports safer and more reliable model performance. Key Responsibilities AI Safety & Quality Evaluation Evaluate AI-generated responses for safety, factual accuracy, policy compliance, relevance, and overall quality Assess whether outputs demonstrate appropriate judgment across nuanced and ambiguous scenarios Identify unsafe, misleading, incomplete, or poorly reasoned responses Compare alternative outputs and determine which response better satisfies safety and quality standards Sensitive-Domain Content Review Review content involving misinformation, political persuasion, self-harm, violence, cybersecurity, biosecurity, fraud, and other sensitive areas Apply appropriate evaluation standards across high-risk and grey-area scenarios Distinguish between legitimate informational requests, potentially harmful content, and clear policy violations Evaluate whether model responses remain useful while handling sensitive subject matter responsibly Rubric Application & Development Apply structured rubrics used in AI safety benchmarking, RLHF, and supervised fine-tuning workflows Assess outputs against defined criteria covering safety, accuracy, reasoning, and instruction adherence Identify ambiguity, gaps, or inconsistencies within evaluation guidelines Contribute to the refinement of scoring standards, policy interpretations, and reviewer instructions Failure Analysis & Structured Feedback Identify hallucinations, unsafe outputs, reasoning failures, and policy-compliance issues Classify recurring model weaknesses and behavioural patterns Provide clear written explanations supporting each evaluation decision Collaborate with researchers and safety teams on calibration and ongoing evaluation initiatives Ideal Profile Strong candidates may have: At least 5 years of professional experience in AI safety, trust and safety, journalism, public policy, scientific research, security, content integrity, or a related field Strong analytical reasoning and the ability to assess nuanced, policy-sensitive scenarios consistently Excellent written English and the ability to explain complex evaluation decisions clearly Experience reviewing sensitive, high-risk, or ambiguous content Strong attention to factual accuracy, context, and policy interpretation Ability to work independently while applying detailed evaluation standards Professional residence in one of the eligible countries listed below Educational Background A bachelor's degree or higher in journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science, or a related discipline is highly relevant Graduate-level education in policy, behavioural science, security, law, life sciences, or artificial intelligence may be valuable Equivalent specialist experience in safety evaluation, content integrity, scientific review, or risk analysis may also be considered Relevant research, policy, moderation, or AI evaluation work may strengthen an application Nice to Have Experience with AI safety, reinforcement learning from human feedback, supervised fine-tuning, trust and safety, or model evaluation Familiarity with content policies, safety standards, moderation frameworks, or rubric development Experience evaluating frontier AI models or language-model outputs Background in misinformation, political content, cybersecurity, biosecurity, scientific safety, or behavioural risk Experience participating in reviewer calibration or quality-assurance programmes Familiarity with structured annotation, safety benchmarking, or human-feedback workflows Previous collaboration with researchers, engineers, policy specialists, or safety teams Why This Opportunity Help shape the safety and behaviour of advanced AI systems Work across challenging real-world scenarios involving complex and sensitive topics Apply professional judgment to improve model alignment, factual quality, and policy compliance Collaborate with experienced AI researchers and safety specialists Participate in flexible remote work with competitive hourly compensation Contract Details Independent contractor role Fully remote with flexible scheduling Competitive rates between $55–$65 per hour depending on expertise and project scope Weekly payments via Stripe or Wise Eligible locations include Albania, Austria, Belgium, Bosnia and Herzegovina, Bulgaria, Croatia, the Czech Republic, Denmark, Estonia, Finland, France, Germany, Greece, Hungary, Iceland, Ireland, Italy, Kosovo, Latvia, Liechtenstein, Lithuania, Luxembourg, Malta, Moldova, Monaco, the Netherlands, North Macedonia, Norway, Poland, Portugal, Romania, San Marino, Serbia, Slovakia, Slovenia, Spain, Sweden, Switzerland, the United Kingdom, and the United States Projects may be extended, shortened, or adjusted depending on scope and performance Work will not involve access to confidential or proprietary information from any employer, client, or institution About the Platform This opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams. By submitting this application, you acknowledge that your information may be processed by 24-MAG LLC for recruitment and opportunity matching in accordance with our Privacy Policy: https://www.24-mag.com/privacy-policy .

Similar jobs

Similar jobs

2M

Remote | AI Safety Specialist (English & Swedish) — $45–$60/hour

24 Mag

🇺🇸United States2 hours ago
2M

Remote | Corporate Accounting & Controllership Specialist — $75–$115/hour

24 Mag

🇺🇸United States2 hours ago
Mercor logo

AI Safety Specialist - Fully Remote | Upto $62/hr

Mercor

🌍Belgium, United Kingdom, United States2 hours ago
Mercor logo

AI Safety Specialist - Fully Remote | Upto $22/hr

Mercor

🌍Australia, Canada, India, Pakistan, Singapore, United Kingdom, United States2 hours ago
SV

Business Intelligence Senior Analyst, Information Technology

Svclnk

🇺🇸United States2 hours ago
Affiliates Commonspirit logo

RN Supervisor UM Prior Auth

Affiliates Commonspirit

🇺🇸United States1 hour ago