Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) - #2727266

Rex.zone


Date: vor 50 Minuten
Stadt: München
Vertragstyp: Ganztags
Arbeitsplan: Volle Tag
Rex.zone
Overview

Rex.zone is hiring Germany-based English/German AI Generalist Trainers to support RLHF and large language model evaluation by assessing, ranking, and quality-checking model outputs to strengthen training data quality and drive model performance improvement.

About The Role

In this role, you will evaluate and improve AI/LLM behaviors through RLHF-style workflows. You will review prompts and model responses, validate factuality and instruction-following, apply annotation guidelines compliance standards, and write clear rationales that improve training data quality.

Key Responsibilities

  • Perform large language model evaluation across English and German prompts and outputs
  • Rank multiple model-generated responses using defined rubrics and consistent reasoning
  • Conduct QA evaluation checks by auditing items for annotation guidelines compliance
  • Write concise rationales that justify rankings and identify reasoning errors
  • Validate outputs for safety, policy adherence, and content safety labeling needs
  • Identify edge cases and failure modes; document findings for model performance improvement
  • Apply data labeling standards to create reliable feedback signals for RLHF pipelines
  • Track discrepancies, escalate guideline questions, and support consistent quality targets

Basic Qualifications

  • Must be based in Germany and able to work remotely from Germany
  • Fluency in English and German (strong reading and writing skills in both)
  • Strong analytical skills; ability to evaluate arguments, logic, and evidence
  • Exceptional attention to detail and consistency in applying rubrics
  • Comfortable working with sensitive or policy-relevant content under safety labeling rules

Preferred Qualifications

  • Experience with AI data labeling, RLHF, prompt evaluation, or LLM evaluation
  • Familiarity with LLM failure patterns (hallucinations, instruction drift, bias, safety issues)
  • Background in linguistics, translation, writing/editing, QA, research, or content moderation

Compensation

Hourly pay range: $30–$50 USD per hour (base), based on skills, language proficiency, and evaluation quality.

How To Apply

Please apply with an updated resume and a brief note describing your English/German proficiency and any experience with evaluation, QA, annotation, or AI/LLM workflows.

Wie bewerbe ich mich?

Um sich für diesen Job zu bewerben, müssen Sie auf unserer Website autorisieren. Wenn Sie noch kein Konto haben, registrieren Sie sich bitte.

Veröffentlichen Sie einen Lebenslauf

Ähnliche Jobs

Account Director (Public Sector DACH)

deepset, makers of Haystack,
vor 20 Minuten
TL;DR You'll open and build deepset's public sector business across Germany, Austria and Switzerland, selling sovereign AI to federal, state and municipal administrations, both directly and through partners. It's a new-business role. There's already strong activity on the patch, and...
deepset, makers of Haystack

Mitarbeiter Backoffice (m/w/d) am Standort München

psyprax GmbH,
vor 20 Minuten
Wir sind einer der führenden Dienstleister im Bereich Praxisverwaltungssoftware in Deutschland und liefern IT-Lösungen aus einer Hand. Wir unterstützen Praxen dabei, die wachsende Komplexität von E-Health zu meistern, damit Ärzte:innen und Psychotherapeut:innen Zeit für das Wesentliche haben: ihre Patient:innen. Unsere...
psyprax GmbH

SAP HCM Spezialist HR-IT / Systembetreuung (m/w/d)

SBK Siemens-Betriebskrankenkasse,
vor 19 Stunden
Sie kennen SAP HCM nicht nur aus Anwendersicht, sondern wissen, wie HR-Prozesse im System konfiguriert, integriert und weiterentwickelt werden? Dann übernehmen Sie bei uns eine zentrale Rolle in der Weiterentwicklung unserer HR-IT. Bei der SBK (Siemens-Betriebskrankenkasse) gestalten Sie die digitale...
SBK Siemens-Betriebskrankenkasse