[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-jobs-berlin-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation, RLHF-style ranking, prompt evaluation, QA evaluation, validation, and data labeling tasks, including writing rationales and following annotation guidelines compliance to maintain training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. Strong analytical skills, attention to detail, and the ability to apply rubrics and write clear reasoning are essential; we provide task guidelines and feedback loops.","Do I need AI experience?",{"A":17,"Q":18},"Fluent German and English are required because you will evaluate and rank outputs in both languages.","What languages are required?",{"A":20,"Q":21},"Tasks can span general knowledge, writing quality, instruction following, reasoning, and content safety labeling, all focused on training data quality and model performance improvement within AI\u002FLLM workflows.","What domains are covered?","ai-generalist-trainer-jobs-berlin-germany",{"desc":24,"title":25,"content":26},"Remote, full-time AI Generalist Trainer role at Rexzone for Germany-based bilingual (English\u002FGerman) professionals to perform large language model evaluation, RLHF-style ranking, and QA evaluation to improve training data quality and support model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","Rexzone is hiring Germany-based English & German AI Generalist Trainers to support AI\u002FLLM workflows by evaluating, ranking, and validating model-generated outputs. You will follow annotation guidelines compliance requirements, write clear rationales, and perform QA evaluation to ensure training data quality for large language model evaluation and model performance improvement. This is a remote, full-time role requiring fluent German and English and strong analytical judgment.",{"h2":31,"desc":32},"Key Responsibilities","Evaluate and compare LLM responses in English and German using defined rubrics; Rank outputs for RLHF-style preference data and prompt evaluation; Perform QA evaluation and validation checks to ensure training data quality and consistency; Write concise reasoning and rationales that justify rankings and highlight errors, hallucinations, or policy issues; Apply annotation guidelines compliance across tasks, documenting edge cases and updating notes for reviewers; Label and categorize content including content safety labeling when required; Escalate unclear prompts, ambiguous instructions, or systematic model failures with actionable examples.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and able to work remotely from Germany; Fluency in German and English (reading and writing) for bilingual evaluation; Strong analytical skills, critical thinking, and structured reasoning; High attention to detail with consistent annotation guidelines compliance; Comfortable working with web-based tooling and handling repeated evaluation and ranking tasks; Ability to meet quality targets and follow feedback to improve training data quality.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with data labeling, prompt evaluation, QA evaluation, or content moderation\u002Fcontent safety labeling; Familiarity with LLM evaluation, RLHF concepts, and common failure modes (hallucinations, safety gaps, instruction-following issues); Experience writing clear rationales and validation notes for reviewers; Self-driven, reliable, and able to manage time in a remote environment.",{"h2":40,"desc":41},"Compensation and Employment Details","Pay: $35–$40\u002Fhour (USD). Remote, full-time. You will contribute directly to training data quality, large language model evaluation, and model performance improvement through high-quality evaluation and ranking work. Apply if you are Germany-based, bilingual in English and German, and ready to support RLHF-style workflows.","AI Data Operations"]