[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-trainer-jobs-hamburg-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a full-time remote role for candidates based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation including evaluation, ranking, QA evaluation, validation, prompt evaluation, and writing reasoning rationales to improve training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI experience is preferred but not required. You must be able to follow annotation guidelines compliance, apply consistent judgment, and deliver high-quality evaluations.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required for bilingual evaluation and ranking tasks.","What languages are required?",{"A":20,"Q":21},"You will evaluate outputs across general knowledge, customer-support style content, writing quality, instruction-following, and content safety labeling scenarios to support model performance improvement.","What domains are covered?","ai-trainer-jobs-hamburg-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (English\u002FGerman) AI Generalist Trainers to support RLHF and large language model evaluation. You will assess, rank, and QA model outputs using annotation guidelines compliance to strengthen training data quality and drive model performance improvement across real-world AI\u002FLLM workflows.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will improve AI systems by performing RLHF-style evaluation, ranking, and QA evaluation of model-generated responses. Your work directly impacts training data quality and model performance improvement by producing consistent judgments, validating outputs against annotation guidelines, and writing clear rationales for decisions. This is a full-time remote role focused on large language model evaluation and prompt evaluation across multiple domains.",{"h2":31,"desc":32},"Key Responsibilities","Evaluate and rank model-generated outputs in English and German using defined rubrics and task instructions; Perform RLHF-style preference ranking and comparative evaluation to identify the best responses; Conduct QA evaluation and validation checks for accuracy, completeness, tone, and policy adherence; Write concise reasoning rationales that justify rankings and support reviewer alignment; Apply annotation guidelines compliance to ensure consistency and reduce label noise; Flag ambiguous cases, escalate edge scenarios, and propose clarifications to annotation guidelines; Validate training data quality by auditing samples, tracking error patterns, and correcting mislabeled items; Support content safety labeling and policy-based decisions for sensitive or restricted content; Collaborate asynchronously with ops and QA leads to calibrate scoring and improve inter-annotator agreement.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and authorized to work as a remote contractor\u002Femployee per local requirements; Fluent in English and German (written and reading comprehension required for evaluation tasks); Strong analytical skills with the ability to compare options, detect subtle errors, and apply consistent judgment; Excellent attention to detail and ability to follow annotation guidelines with high precision; Comfortable writing short, structured rationales that explain reasoning and validation decisions; Reliable internet connection and ability to meet full-time throughput and QA targets.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with data labeling, prompt evaluation, QA evaluation, or content safety labeling; Familiarity with LLM evaluation, RLHF workflows, and common LLM failure modes (hallucinations, unsafe content, instruction-following issues); Experience applying rubrics, taxonomies, or annotation guidelines compliance in production environments; Self-driven, organized, and able to work independently in a remote setting while maintaining quality and consistency; Interest in AI safety, training data quality, and continuous model performance improvement.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly), depending on skills and task alignment. Full-time remote.",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with a short summary of your bilingual (English\u002FGerman) experience, availability for full-time remote work in Germany, and any relevant work in evaluation, ranking, QA, or data labeling. Selected candidates may complete an online skills calibration focused on large language model evaluation and annotation guidelines compliance.","AI Data Operations"]