[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-jobs-%2440-per-hour-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will evaluate and rank model outputs, perform QA evaluation and validation, write rationales explaining reasoning, and follow annotation guidelines compliance to improve training data quality and support model performance improvement.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. We value strong analytical skills, attention to detail, and the ability to learn RLHF and large language model evaluation workflows quickly.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both German and English is required, with strong reading and writing skills in each.","What languages are required?",{"A":20,"Q":21},"Domains may include general knowledge, customer-style assistance, safety and content policy scenarios, and bilingual writing tasks, including content safety labeling, prompt evaluation, and training data quality reviews.","What domains are covered?","ai-generalist-trainer-jobs-$40-per-hour-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based bilingual (German\u002FEnglish) AI Generalist Trainers to support RLHF and large language model evaluation by rating, ranking, and validating model outputs. You will apply annotation guidelines compliance to improve training data quality and drive model performance improvement through consistent, well-reasoned evaluations and QA evaluation across diverse content types.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","As a Germany-Based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve AI systems by assessing model-generated responses in English and German. Your work supports RLHF, large language model evaluation, and training data quality initiatives by producing high-signal ratings, pairwise rankings, validations, and written rationales that enable model performance improvement.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation by scoring and ranking responses against rubrics; execute QA evaluation through audits, consistency checks, and error triage; write clear rationales explaining reasoning and trade-offs; validate outputs for factuality, instruction-following, tone, and safety; follow annotation guidelines compliance and escalate unclear cases; contribute to prompt evaluation and dataset reviews to improve training data quality; support content safety labeling and policy-driven decisions; track edge cases and document patterns to enable model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and able to work remotely from Germany; fluent in both German and English (reading and writing); strong analytical skills with the ability to compare nuanced outputs and justify decisions; exceptional attention to detail and consistency; comfortable following annotation guidelines compliance and structured rubrics; reliable internet connection and ability to meet full-time productivity and quality targets.",{"h2":37,"desc":38},"Preferred Qualifications","Experience in data labeling, RLHF, prompt evaluation, or LLM evaluation; familiarity with LLM behavior (hallucinations, instruction-following, safety constraints) and large language model evaluation workflows; prior work with annotation tools, QA checklists, and training data quality processes; self-driven, organized, and comfortable working independently in a remote setting.",{"h2":40,"desc":41},"How to Apply","Apply to Rexzone with a concise resume highlighting bilingual German\u002FEnglish writing ability, evaluation or QA experience, and any AI\u002Fannotation background. Candidates may be asked to complete a short skills assessment focused on ranking, reasoning, validation, and annotation guidelines compliance.","AI Data Operations"]