[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-jobs-germany-hiring-now":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a full-time remote role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation by reviewing model outputs, ranking and comparing responses (RLHF-style), completing QA evaluation and validation, and writing rationales while following annotation guidelines compliance to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. We value strong analytical skills, attention to detail, and the ability to apply guidelines consistently; training is provided for evaluation rubrics and workflows.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required, including reading and writing for prompt evaluation and response assessment.","What languages are required?",{"A":20,"Q":21},"You may evaluate general knowledge, reasoning, writing quality, instruction following, and content safety labeling across a variety of consumer and professional scenarios, with an emphasis on training data quality and model performance improvement.","What domains are covered?","ai-generalist-trainer-jobs-germany-hiring-now",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (English\u002FGerman) AI Generalist Trainers to support RLHF and large language model evaluation by assessing, ranking, and validating model outputs to strengthen training data quality and drive model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","In this full-time remote role at Rexzone, you will evaluate and improve AI\u002FLLM workflows by reviewing model-generated responses, ranking outputs, performing QA evaluation, and writing clear rationales that support model performance improvement. You will follow annotation guidelines compliance and contribute to training data quality through consistent large language model evaluation across English and German content.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation for English and German outputs; rank and compare responses using RLHF-style criteria; execute QA evaluation and validation checks to ensure training data quality; identify errors in reasoning, factuality, safety, and instruction-following; write concise rationales to justify rankings and evaluations; apply annotation guidelines compliance consistently and flag guideline gaps; conduct prompt evaluation and support data labeling and content safety labeling where needed; track edge cases and report recurring failure modes to improve model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and authorized to work remotely from Germany. Fluent in English and German (reading and writing). Strong analytical skills, critical thinking, and attention to detail. Ability to evaluate reasoning and provide clear, structured written rationales. Comfortable working with web tools, spreadsheets, and task queues while meeting quality and throughput targets.",{"h2":37,"desc":38},"Preferred Qualifications","Experience with AI training data, data labeling, or annotation work. Familiarity with LLM evaluation, RLHF concepts, prompt evaluation, or QA evaluation methods. Background in linguistics, technical writing, translation, or content moderation. Self-driven, dependable, and able to follow evolving annotation guidelines compliance with minimal supervision.",{"h2":40,"desc":41},"How to Apply","Apply to Rexzone with a short summary of your English\u002FGerman proficiency, your availability for full-time remote work from Germany, and any experience related to data labeling, QA, or large language model evaluation. Qualified candidates may be asked to complete an evaluation task focused on training data quality and ranking consistency.","AI Data Operations"]