AI Safety Specialist - Evaluation Expert

mercor

Quelle: HimalayasStandort: EU/EMEAAktiv bestätigt: 10. Sept. 2026
Freelance / Vertrag

Unsere Einschätzung

  • Unsere Prüfung des gesamten Anzeigentexts bestätigt: vollständig remote.
  • 25 weitere offene Stellen dieses Arbeitgebers in unserem Index. Davon 25 vollständig remote.

Nur dieser Abschnitt: automatisch von nomado24 berechnet, aus unserem Stellenindex und unserer eigenen Prüfung des Anzeigentexts. Keine Angabe des Arbeitgebers.

Stellenbeschreibung

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Practitioner
Type: Contract
Compensation: $60–$70/hour
Location: Remote

Role Responsibilities

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF , SFT , and AI safety benchmarking .
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Qualifications

Must-Have

  • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
  • 5+ years of professional experience in AI Safety , Trust & Safety, journalism, public policy, scientific research, security, or a related field.
  • Excellent written English , critical thinking, and analytical reasoning skills.
  • Ability to consistently evaluate nuanced and policy-sensitive scenarios.

Preferred

  • Experience with AI Safety , RLHF , SFT , Trust & Safety, or AI evaluation.
  • Familiarity with safety policies, content moderation, or evaluation rubric development.
  • Experience reviewing complex, high-risk, or ambiguous content.

Application Process (Takes 20–30 mins to complete)

  • Upload resume
  • AI interview based on your resume
  • Submit form

Resources & Support

  • For details about the interview process and platform information, please check:
  • For any help or support, reach out to:

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Originally posted on Himalayas

AI SafetyTrust and SafetyContent EvaluationAI Safety PractitionerAI Research SupportAI Safety EvaluatorAI Safety SpecialistAI Safety Expert

Diese Stelle wird von einer externen Quelle bereitgestellt. Die Bewerbung erfolgt auf der Website der Quelle.

Neue Remote-Jobs wie „Safety Specialist“ per E-Mail

Nur wenn etwas Passendes erscheint. Kostenlos, jederzeit abbestellbar.

Du bestätigst die Anmeldung per E-Mail (Double-Opt-in) und kannst dich jederzeit über den Link in jeder E-Mail abmelden. Es gilt unsere Datenschutzerklärung.

Alle Remote Data Jobs

Bewerbung mit KI vorbereiten

Lass dir vom KI-Berater ein Anschreiben entwerfen, deine Profil-Lücken zur Stelle analysieren und dich aufs Interview vorbereiten.

Anmelden & mit KI vorbereiten

Kostenlos: nach der Anmeldung geht es direkt weiter.

Oder lass uns für dich suchen

Nicht der richtige Job?