[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-llm-evaluator-remote-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation, ranking and comparison of model outputs, prompt evaluation, QA evaluation, validation checks, data labeling per annotation guidelines, and write concise rationales that explain your reasoning.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. Rexzone values strong analytical skills, attention to detail, and the ability to follow annotation guidelines compliance for training data quality and model performance improvement.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required, including reading and writing in both languages.","What languages are required?",{"A":20,"Q":21},"Projects commonly include general knowledge, customer-style conversations, instruction following, content safety labeling, and reasoning-focused tasks aligned with RLHF and training data quality goals.","What domains are covered?","llm-evaluator-remote-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (English\u002FGerman) AI Generalist Trainers to support RLHF and large language model evaluation. You will evaluate, rank, and QA model outputs to strengthen training data quality, enforce annotation guidelines compliance, and drive model performance improvement across real-world AI\u002FLLM workflows.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will improve AI systems by performing large language model evaluation and RLHF-style assessments. You will review model-generated responses in English and German, compare and rank outputs, write clear rationales, and validate data to ensure training data quality. This remote, full-time role focuses on consistent evaluation practices, annotation guidelines compliance, and measurable model performance improvement.",{"h2":31,"desc":32},"What You Will Do","You will evaluate and rank model responses, perform QA evaluation and validation checks, and document reasoning that explains preferences and errors. Your work will include prompt evaluation, data labeling decisions, content safety labeling when required, and maintaining high training data quality standards for LLM evaluation workflows.",{"h2":34,"desc":35},"How Success Is Measured","Success is measured by accuracy and consistency in large language model evaluation, strong annotation guidelines compliance, high-quality rationales, effective QA evaluation outcomes, and reliable validation that supports model performance improvement.",{"h2":37,"desc":38},"Compensation","This role pays $35–$40 USD per hour (hourly), depending on project needs and performance in evaluation and QA tasks.",{"h2":40,"desc":41},"How to Apply","Apply to Rexzone with a brief summary of your bilingual English\u002FGerman experience, your Germany-based availability, and examples of analytical work that demonstrate attention to detail and structured reasoning.","AI Data Operations"]