[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-rater-jobs-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will evaluate and rank model outputs, perform QA evaluation and validation, write reasoning-based rationales, follow annotation guidelines, and complete content safety labeling as needed to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. Strong analytical skills, attention to detail, and the ability to follow annotation guidelines compliance are essential; we provide task instructions and rubrics.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required for bilingual large language model evaluation and prompt evaluation tasks.","What languages are required?",{"A":20,"Q":21},"Tasks can span general knowledge, customer support-style interactions, safety and policy scenarios, reasoning and instruction-following, and other domains relevant to model performance improvement.","What domains are covered?","ai-rater-jobs-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, remote, full-time AI Generalist Trainers to support RLHF and large language model evaluation by assessing, ranking, and validating model outputs. You will apply annotation guidelines compliance to improve training data quality, document reasoning, and drive model performance improvement across English and German workflows. This role focuses on LLM evaluation, prompt evaluation, QA evaluation, and content safety labeling to help teams ship safer, more accurate AI systems.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate model-generated responses, rank alternatives, write clear rationales, and perform validation checks that improve training data quality. Your work directly supports RLHF pipelines, large language model evaluation, and model performance improvement through consistent application of annotation guidelines compliance.",{"h2":31,"desc":32},"Responsibilities","Evaluate and rank model-generated outputs in English and German using defined rubrics and annotation guidelines.\nPerform QA evaluation on labeled datasets to validate consistency, completeness, and policy adherence.\nWrite concise reasoning and rationales that justify rankings and identify failure modes.\nValidate prompt-response pairs, detect hallucinations, and flag safety and compliance issues.\nApply content safety labeling and ensure annotation guidelines compliance across domains.\nCollaborate asynchronously with remote reviewers to resolve edge cases and improve rubrics.\nTrack errors, suggest rubric updates, and contribute to model performance improvement initiatives.\nMaintain high throughput while meeting training data quality targets and review SLAs.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and eligible to work as a remote contractor\u002Femployee per local requirements.\nFluency in English and German (reading and writing) with the ability to evaluate nuanced tone and intent.\nStrong analytical skills and structured thinking for comparative evaluation and ranking tasks.\nHigh attention to detail and consistent adherence to annotation guidelines compliance.\nComfort working with web-based labeling tools, spreadsheets, and written QA checklists.\nAbility to explain reasoning clearly and consistently in written rationales.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in data labeling, QA evaluation, content moderation, or annotation operations.\nFamiliarity with RLHF concepts and large language model evaluation practices.\nExperience with prompt evaluation, rubric-based ranking, and error taxonomy creation.\nSelf-driven, reliable, and able to manage workload independently in a remote setting.\nInterest in improving training data quality and contributing to model performance improvement.",{"h2":40,"desc":41},"Compensation and Schedule","Pay: $35–$40 USD per hour (hourly). Full-time, remote. Work is performed from Germany with bilingual English\u002FGerman task requirements.",{"h2":43,"desc":44},"How to Apply","Apply through Rexzone with your resume\u002FCV and a brief note highlighting bilingual English\u002FGerman experience, analytical evaluation work, and any LLM evaluation or data labeling exposure.","AI Data Operations"]