[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-jobs-bremen-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will evaluate and rank model-generated outputs, perform QA evaluation and validation, write reasoning\u002Frationales, follow annotation guidelines compliance requirements, and support training data quality and model performance improvement in RLHF workflows.","What tasks will I do?",{"A":14,"Q":15},"AI experience is preferred but not required. Strong analytical skills, attention to detail, and the ability to follow annotation guidelines are essential; Rexzone provides task instructions and rubrics.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both English and German is required, as you will complete large language model evaluation tasks in both languages.","What languages are required?",{"A":20,"Q":21},"Tasks commonly cover general knowledge and everyday professional writing, including prompt evaluation, safety and content safety labeling, and quality-focused assessments designed to improve training data quality.","What domains are covered?","ai-generalist-trainer-jobs-bremen-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual English\u002FGerman AI Generalist Trainers to support AI\u002FLLM workflows through RLHF, large language model evaluation, and training data quality improvements by evaluating, ranking, and validating model outputs with clear rationales.","Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As an AI Generalist Trainer at Rexzone, you will remotely evaluate and improve large language model evaluation workflows by assessing model-generated responses in English and German. Your work will directly support RLHF pipelines, training data quality, annotation guidelines compliance, and model performance improvement through careful evaluation, ranking, QA evaluation, and rationale writing.",{"h2":31,"desc":32},"Key Responsibilities","Perform prompt evaluation and response evaluation for LLM outputs; rank multiple model responses using project rubrics; write concise, evidence-based reasoning and rationales for rankings; execute QA evaluation to validate labels and ensure annotation guidelines compliance; detect and flag content safety labeling issues and policy violations; validate edge cases, ambiguity, and hallucinations through structured checks; apply annotation guidelines consistently and escalate unclear cases with proposed guideline refinements; contribute to training data quality by identifying systematic model errors and documenting trends that enable model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and able to work remotely from Germany; fluent in English and German (written and verbal); strong analytical skills with the ability to compare outputs against criteria; high attention to detail and consistency when following annotation guidelines; able to explain evaluation decisions clearly with defensible reasoning; comfortable working with web-based labeling tools and handling repetitive evaluation tasks while maintaining quality.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with data labeling, QA evaluation, or content review; familiarity with RLHF concepts, LLM evaluation, and prompt evaluation; experience improving training data quality through validation and error analysis; self-driven, reliable, and able to manage time independently in a remote environment; interest in content safety labeling and policy-based decision making.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly).",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with a short summary of your bilingual English\u002FGerman experience, your location in Germany, and any relevant evaluation, annotation, or QA work. Include availability and confirm you can perform large language model evaluation tasks remotely.","AI Data Operations"]