[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-llm-evaluator-jobs-munich-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, but you must be based in Germany.","Is This Role Remote?",{"A":11,"Q":12},"You will perform large language model evaluation, rank model-generated outputs, complete QA evaluation checks, validate responses against rubrics and policies, and write clear rationales that support RLHF and training data quality.","What Tasks Will I Do?",{"A":14,"Q":15},"AI experience is helpful but not required. If you can follow annotation guidelines compliance requirements, apply consistent reasoning, and deliver high-quality evaluations, Rexzone will provide role-specific guidance.","Do I Need AI Experience?",{"A":17,"Q":18},"Fluency in English and German is required, including strong reading and writing skills in both languages.","What Languages Are Required?",{"A":20,"Q":21},"You may evaluate prompts and outputs across general knowledge, writing quality, summarization, translation, safety and content safety labeling, instruction-following, and reasoning tasks aimed at model performance improvement.","What Domains Are Covered?","llm-evaluator-jobs-munich-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based English\u002FGerman AI Generalist Trainers to support RLHF and large language model evaluation by assessing, ranking, and quality-checking model outputs to strengthen training data quality and drive model performance improvement.","Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve AI\u002FLLM behaviors through RLHF-style workflows. Your work focuses on large language model evaluation, prompt evaluation, and QA evaluation, including ranking model-generated responses, validating factuality and instruction-following, and writing clear rationales. You will follow annotation guidelines compliance standards to ensure consistent training data quality and measurable model performance improvement.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation across English and German prompts and responses; rank multiple model outputs using defined rubrics and reasoning; conduct QA evaluation by auditing items for annotation guidelines compliance; write concise rationales that justify rankings and highlight reasoning errors; validate outputs for safety, policy adherence, and content safety labeling needs; identify edge cases, ambiguity, and failure modes, then document them for model performance improvement; apply data labeling standards to create reliable feedback signals for RLHF pipelines; track discrepancies, escalate guideline questions, and support consistent training data quality.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and able to work remotely from Germany; fluent in both English and German (written and reading proficiency required); strong analytical skills with the ability to evaluate arguments, logic, and evidence; exceptional attention to detail and consistency in applying rubrics; ability to follow structured processes and meet productivity and quality targets; comfortable working with sensitive or policy-relevant content under content safety labeling rules.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in AI data labeling, RLHF, prompt evaluation, or large language model evaluation; familiarity with LLM capabilities and common failure patterns (hallucinations, instruction drift, bias, safety issues); background in linguistics, translation, writing\u002Fediting, QA, research, or content moderation; self-driven and reliable in a remote environment, with strong time management and documentation habits.",{"h2":40,"desc":41},"Compensation","This is a full-time remote role. Pay is $35–$40 USD per hour, based on skills, language proficiency, and evaluation quality.",{"h2":43,"desc":44},"How to Apply","Apply through Rexzone with an updated resume and a brief note describing your English\u002FGerman proficiency and any experience with evaluation, QA, annotation, or AI\u002FLLM workflows. Candidates may be asked to complete a short skills assessment focused on ranking, reasoning, and training data quality.","AI Data Operations"]