[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-rater-jobs-hamburg-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":36},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a full-time remote role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation tasks including RLHF-style ranking, prompt evaluation, QA evaluation, label validation, and writing rationales in both English and German while following annotation guidelines compliance.","What tasks will I do?",{"A":14,"Q":15},"AI or annotation experience is helpful but not always required. Strong analytical skills, attention to detail, and the ability to apply guidelines consistently are essential.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in English and German is required because you will evaluate, compare, and write rationales across bilingual content.","What languages are required?",{"A":20,"Q":21},"You may evaluate content across general knowledge, reasoning, instruction-following, writing quality, and content safety labeling scenarios to support training data quality and model performance improvement.","What domains are covered?","ai-rater-jobs-hamburg-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based AI Generalist Trainers to support large language model evaluation through RLHF, prompt evaluation, and training data quality work. You will assess, rank, and QA model outputs in English and German, write clear rationales, and follow annotation guidelines compliance to drive model performance improvement. This full-time remote role focuses on high-precision evaluation, reasoning, and validation across diverse domains while ensuring consistent, safe, and reliable labeling outcomes.","Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) 2026 May",[27,30,33],{"h2":28,"desc":29},"About the Role","As an AI Generalist Trainer at Rexzone, you will evaluate and improve AI systems by performing large language model evaluation tasks in bilingual English\u002FGerman workflows. Your work will directly influence training data quality, model performance improvement, and annotation guidelines compliance through RLHF-style preference ranking, QA evaluation, and detailed rationale writing.",{"h2":31,"desc":32},"What You Will Do","You will review model-generated responses, compare alternatives, rank outputs based on instructions and quality rubrics, and provide concise reasoning. You will validate labels, perform QA checks, flag edge cases, and apply content safety labeling requirements to ensure high-quality, policy-aligned datasets for downstream training and evaluation.",{"h2":34,"desc":35},"How Success Is Measured","Success is measured by accuracy of evaluations, consistency of ranking decisions, strength and clarity of written rationales, and adherence to annotation guidelines compliance. High-performing trainers demonstrate strong analytical reasoning, reliable QA habits, and careful validation that improves training data quality and supports model performance improvement.","AI Data Operations"]