[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-work-from-home-jobs-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a full-time remote role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation tasks such as RLHF-style ranking, prompt evaluation, QA evaluation, validation of outputs, content safety labeling, and writing reasoning rationales while following annotation guidelines compliance.","What tasks will I do?",{"A":14,"Q":15},"AI experience is preferred but not required. Strong analytical skills, attention to detail, and the ability to follow guidelines consistently are essential; training is provided for the evaluation workflow.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required, including professional reading and writing in each language.","What languages are required?",{"A":20,"Q":21},"You will evaluate model outputs across general domains such as reasoning, summarization, instruction following, customer-support style interactions, and safety-related scenarios, with a focus on training data quality and model performance improvement.","What domains are covered?","ai-generalist-trainer-work-from-home-jobs-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based AI Generalist Trainers to support large language model evaluation through RLHF, ranking, QA evaluation, and bilingual (English\u002FGerman) review of model outputs to drive training data quality and model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate large language model outputs across varied tasks, apply annotation guidelines compliance, and provide clear rationales that support RLHF workflows. Your work directly impacts training data quality, large language model evaluation, and ongoing model performance improvement by validating responses, ranking alternatives, and flagging safety or policy issues.",{"h2":31,"desc":32},"Responsibilities","Evaluate and rank model-generated outputs in English and German using defined rubrics; perform QA evaluation and validation checks to ensure training data quality; write concise reasoning rationales explaining rankings and corrections; follow annotation guidelines compliance and document edge cases for guideline updates; conduct prompt evaluation and identify failure patterns impacting model performance improvement; label content for content safety labeling and policy adherence; collaborate asynchronously to resolve disagreements, calibrate scoring, and improve consistency.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany; fluent in English and German (professional reading and writing); strong analytical skills with the ability to evaluate nuanced reasoning and compare alternatives; exceptional attention to detail and consistency in applying rubrics; ability to follow strict annotation guidelines compliance and handle sensitive content within policy; reliable internet connection and ability to work full-time remotely.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with data labeling, prompt evaluation, QA evaluation, or RLHF-style ranking; familiarity with LLM evaluation methods and common LLM failure modes; experience writing clear, structured rationales; self-driven and able to manage throughput, accuracy, and feedback cycles independently; comfort working across multiple domains (general knowledge, customer support style prompts, summarization, reasoning, and safety).",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly, full-time).",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with a short summary of your bilingual (English\u002FGerman) experience, availability in Germany, and any background in evaluation, annotation, or QA. Selected candidates may complete a paid skills assessment focused on large language model evaluation, ranking, and reasoning.","AI Data Operations"]