[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-jobs-onsite-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation work such as evaluation and ranking of responses, QA evaluation, prompt evaluation, writing rationales, validation checks, data labeling, and content safety labeling to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. If you can follow annotation guidelines compliance, apply strong analytical reasoning, and maintain quality, you can succeed; prior annotation or LLM evaluation experience is a plus.","Do I need AI experience?",{"A":17,"Q":18},"Fluent English and German are required, including the ability to read and write with high accuracy.","What languages are required?",{"A":20,"Q":21},"You will work across general domains commonly used in AI\u002FLLM workflows, including helpfulness, safety, factuality, reasoning quality, and policy-aligned content, supporting model performance improvement.","What domains are covered?","ai-generalist-trainer-jobs-onsite-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based AI Generalist Trainers to support large language model evaluation through RLHF-style ranking, prompt evaluation, and QA evaluation. You will label and review model outputs in English and German, follow annotation guidelines compliance, and produce clear rationales that strengthen training data quality and drive model performance improvement across AI\u002FLLM workflows.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve AI systems by assessing model-generated responses, ranking alternatives, validating factuality and reasoning, and applying annotation guidelines to ensure training data quality. This is a remote, full-time role focused on large language model evaluation, RLHF-aligned feedback, and consistent QA evaluation.",{"h2":31,"desc":32},"Responsibilities","Evaluate and rank model-generated outputs in English and German using defined rubrics and prompt evaluation criteria; perform QA evaluation and validation checks for consistency, policy adherence, and content safety labeling; write concise, evidence-based rationales that explain ranking and reasoning; apply data labeling standards with strict annotation guidelines compliance to improve training data quality; identify edge cases, ambiguity, and failure patterns and escalate issues with clear examples; track and report quality metrics and contribute feedback to improve workflows for model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and authorized to work as a contractor\u002Femployee per local requirements; fluent in English and German (reading and writing) with strong grammar and nuance; strong analytical skills with the ability to compare alternatives and justify decisions; high attention to detail and ability to follow annotation guidelines compliance; reliable internet connection and ability to work independently in a remote setting.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in AI data labeling, content moderation, QA evaluation, or large language model evaluation; familiarity with RLHF concepts, prompt evaluation, and training data quality best practices; comfort working with ambiguous tasks and rapidly evolving guidelines; self-driven, organized, and able to maintain consistency at scale.",{"h2":40,"desc":41},"How to Apply","Apply through Rexzone with your resume\u002FCV and a brief note confirming Germany-based location and English\u002FGerman proficiency. Qualified candidates may be invited to complete an evaluation task focused on ranking, reasoning, and validation.","AI Data Operations"]