[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-annotation-remote-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":48},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time position, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation work such as evaluating outputs, ranking responses, completing QA evaluation and validation checks, writing reasoning-based rationales, and following annotation guidelines compliance to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is preferred but not required. We value strong analytical skills, attention to detail, and the ability to consistently apply guidelines; training is provided for RLHF and LLM evaluation workflows.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both German and English is required, including the ability to write clear rationales and evaluate nuanced meaning in both languages.","What languages are required?",{"A":20,"Q":21},"You will cover generalist domains such as reasoning, instruction following, summarization, professional writing, and content safety labeling, contributing to training data quality and model performance improvement across multiple task types.","What domains are covered?","ai-annotation-remote-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based English & German AI Generalist Trainers to support AI\u002FLLM workflows through RLHF, large language model evaluation, and training data quality improvements by evaluating, ranking, and QA-checking model outputs with clear rationales and strict annotation guidelines compliance.","Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) 2026 May",[27,30,33,36,39,42,45],{"h2":28,"desc":29},"About the Role","As a Germany-Based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve AI systems by assessing, ranking, and validating model-generated outputs in both German and English. Your work directly supports RLHF pipelines, large language model evaluation, and training data quality, helping drive model performance improvement through consistent judgments, annotation guidelines compliance, and high-quality written rationales. This is a remote, full-time role for candidates based in Germany.",{"h2":31,"desc":32},"Responsibilities","Evaluate and rank model-generated responses for correctness, relevance, completeness, tone, and safety; perform QA evaluation and validation of labeled datasets to ensure training data quality; write clear reasoning and rationales that justify rankings and decisions in English and German; apply annotation guidelines compliance consistently and flag ambiguities or edge cases for guideline refinement; conduct prompt evaluation and error analysis to identify recurring failure modes and propose improvements for model performance improvement; perform content safety labeling and policy-based reviews to reduce harmful or non-compliant outputs; verify multilingual accuracy (German\u002FEnglish) and ensure faithful translation, intent preservation, and terminology consistency; track issues, document decisions, and contribute to calibration sessions to maintain inter-annotator agreement.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and able to work remotely from Germany; fluent in both German and English (reading, writing, and nuanced comprehension); strong analytical skills with the ability to evaluate arguments, logic, and evidence; exceptional attention to detail and consistency when following annotation guidelines; comfortable writing concise, well-structured rationales that explain evaluation and ranking decisions; reliable internet access and ability to meet productivity and quality targets in a full-time schedule.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience in data labeling, QA evaluation, or content review for AI\u002FML systems; familiarity with RLHF, LLM evaluation, prompt evaluation, or human-in-the-loop workflows; experience applying content safety labeling or policy-based moderation guidelines; ability to work self-driven in a remote environment, manage time effectively, and maintain consistent output quality; comfort with iterative feedback, calibration, and continuous guideline updates focused on training data quality and model performance improvement.",{"h2":40,"desc":41},"Skills and Domains Covered","This role covers generalist evaluation domains such as everyday knowledge, professional communication, reasoning, summarization, instruction following, and multilingual (German\u002FEnglish) content. You will work with training data quality practices, annotation guidelines compliance, and large language model evaluation processes, including ranking, validation, and QA across multiple task types.",{"h2":43,"desc":44},"Compensation and Schedule","Compensation is $35–$40 per hour (USD), full-time, remote (Germany-based). Work involves structured evaluation queues, quality targets, and regular calibration to ensure consistent labeling and model performance improvement outcomes.",{"h2":46,"desc":47},"How to Apply","Apply to Rexzone with your resume\u002FCV and a short note confirming you are based in Germany and fluent in English and German. If selected, you will complete a brief skills assessment focused on large language model evaluation, ranking, reasoning, and annotation guidelines compliance.","AI Data Operations"]