[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-remote-vacancy-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation tasks including evaluating and ranking model outputs, writing rationales, completing prompt evaluation, running QA evaluation checks, and doing data labeling with annotation guidelines compliance to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. We value strong analytical skills, attention to detail, and the ability to follow rubrics; training is provided for RLHF-style workflows and evaluation standards.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required, including strong reading and writing skills for nuanced evaluation and rationale writing.","What languages are required?",{"A":20,"Q":21},"Tasks can span general knowledge, customer support-style writing, summarization, reasoning, and content safety labeling. You will apply consistent evaluation criteria to support training data quality and model performance improvement.","What domains are covered?","ai-generalist-trainer-remote-vacancy-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (English\u002FGerman) AI Generalist Trainers to support RLHF and large language model evaluation by assessing, ranking, and QA-checking model outputs to strengthen training data quality and drive model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will work remotely and full-time to evaluate and improve AI systems used in modern AI\u002FLLM workflows. Your work will focus on RLHF-style evaluation, prompt evaluation, and QA evaluation: you will compare model-generated responses, rank outputs, validate factuality and reasoning, and write clear rationales. You will apply annotation guidelines compliance to produce consistent labels that improve training data quality and support model performance improvement across multiple domains.",{"h2":31,"desc":32},"Key Responsibilities","Evaluate and rank model-generated outputs in English and German using defined rubrics; perform large language model evaluation for helpfulness, harmlessness, and honesty; write concise, evidence-based rationales explaining ranking decisions and reasoning; conduct QA evaluation by reviewing peer work for accuracy, consistency, and annotation guidelines compliance; label and validate training data (data labeling) including content safety labeling and policy-based tagging; identify edge cases, ambiguity, and failure patterns to support model performance improvement; validate prompts and responses for instruction-following, tone, and clarity (prompt evaluation); maintain high training data quality through careful self-checks and systematic validation; document issues, propose rubric clarifications, and help refine annotation guidelines.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and authorized to work as a contractor\u002Femployee as applicable; fluent in both German and English (reading and writing) with the ability to evaluate nuanced meaning; strong analytical skills to compare responses, detect logical gaps, and justify decisions; exceptional attention to detail to ensure consistent labels and training data quality; ability to follow annotation guidelines compliance and apply rubrics consistently; reliable internet connection and ability to work independently in a remote environment.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in data labeling, content moderation, QA evaluation, search evaluation, or other annotation workflows; familiarity with LLM evaluation, RLHF concepts, or prompt evaluation frameworks; comfort explaining reasoning clearly and consistently across varied tasks and domains; self-driven, organized, and responsive when handling feedback and iterative guideline updates; interest in AI safety, content safety labeling, and model performance improvement.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour, full-time remote. Exact rate within the range depends on assessment performance, language proficiency, and task alignment.",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with an up-to-date CV that highlights bilingual English\u002FGerman writing skills, evaluation or QA experience, and any exposure to AI\u002FLLM workflows. Selected candidates may complete a short qualification task focused on large language model evaluation, ranking, and rationale writing.","AI Data Operations"]