[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-remote-munich-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a fully remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will evaluate and rank model outputs, perform QA evaluation, write reasoning\u002Frationales, validate work against annotation guidelines compliance, and contribute to training data quality in RLHF and large language model evaluation workflows.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. You must be able to follow annotation guidelines, apply consistent judgment, and produce high-quality evaluations that support model performance improvement.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in English and German is required, as you will evaluate and write rationales across both languages.","What languages are required?",{"A":20,"Q":21},"Tasks span general assistant behaviors, prompt evaluation, QA evaluation, content safety labeling, and other areas relevant to training data quality and RLHF-based large language model evaluation.","What domains are covered?","ai-generalist-trainer-remote-munich-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based English\u002FGerman AI Generalist Trainers to support AI\u002FLLM workflows through RLHF, large language model evaluation, and training data quality improvements by evaluating, ranking, and QA-checking model outputs with clear rationales.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve large language model behavior by assessing model-generated responses, ranking alternatives, and writing concise rationales that support model performance improvement. This role is fully remote and full-time, focused on RLHF, LLM evaluation, data labeling, prompt evaluation, and QA evaluation to raise training data quality. You will follow strict annotation guidelines compliance, validate edge cases, and flag content safety issues to ensure reliable, high-quality training signals.",{"h2":31,"desc":32},"Key Responsibilities","Evaluate and rank model-generated outputs in English and German; perform QA evaluation to verify correctness, helpfulness, and policy adherence; write clear reasoning and rationales that justify rankings; validate tasks against annotation guidelines compliance and escalate ambiguities; perform prompt evaluation and response comparison for RLHF workflows; label and categorize content for training data quality, including content safety labeling; run consistency checks, error analysis, and spot regressions affecting model performance improvement; document decisions and contribute to guideline updates to strengthen large language model evaluation.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany; fluent in English and German (written and reading comprehension required for evaluation work); strong analytical skills with the ability to compare nuanced answers and detect contradictions; exceptional attention to detail for training data quality and QA evaluation; comfortable writing short, structured rationales and applying annotation guidelines compliance; reliable internet connection and ability to work independently in a remote environment.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in AI data labeling, LLM evaluation, RLHF, prompt evaluation, or QA evaluation; familiarity with large language model evaluation concepts (hallucinations, grounding, instruction-following, safety); experience applying annotation guidelines at scale and maintaining high training data quality; self-driven, organized, and able to handle repeated evaluation tasks with consistent judgment.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly). Full-time, remote. Exact rate within the range depends on task complexity and demonstrated evaluation quality.",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with a short summary of your bilingual English\u002FGerman experience, your location in Germany, and any relevant background in evaluation, QA, data labeling, or working with AI\u002FLLM systems. Selected candidates may complete a short skills assessment focused on large language model evaluation and annotation guidelines compliance.","AI Data Operations"]