[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-bilingual-ai-trainer-jobs-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation, including prompt evaluation, RLHF-style ranking of model outputs, QA evaluation, validation of labels, and writing reasoning-based rationales while maintaining training data quality.","What tasks will I do?",{"A":14,"Q":15},"AI or annotation experience is preferred but not strictly required if you can follow guidelines precisely, reason clearly, and produce consistent evaluation judgments.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in English and German is required, as tasks involve bilingual evaluation and rationale writing.","What languages are required?",{"A":20,"Q":21},"Domains vary by project and may include general knowledge, reasoning, instruction-following, content safety labeling, and other areas used to improve training data quality and model performance improvement.","What domains are covered?","bilingual-ai-trainer-jobs-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (English\u002FGerman) AI Generalist Trainers to support large language model evaluation through RLHF-style ranking, prompt evaluation, and QA evaluation to improve training data quality and drive model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate model-generated responses across tasks and domains, applying annotation guidelines compliance to deliver high-quality judgments and rationales. You will work in AI\u002FLLM workflows that include RLHF, large language model evaluation, data labeling, and content safety labeling, focusing on training data quality and measurable model performance improvement.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation by reviewing prompts and model outputs; rank and compare multiple responses using RLHF-style preference signals; execute QA evaluation to validate labels, catch inconsistencies, and ensure annotation guidelines compliance; write clear reasoning and rationales that justify rankings and decisions in both English and German; validate edge cases, ambiguity, and safety-sensitive content using content safety labeling policies; follow detailed project instructions, maintain training data quality targets, and report issues or guideline gaps; contribute to prompt evaluation and test set creation to support model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and authorized to work from Germany; fluent in English and German (written and reading comprehension required for evaluation and rationale writing); strong analytical skills with the ability to break down instructions and apply consistent decision logic; exceptional attention to detail and comfort following strict annotation guidelines compliance; reliable internet connection and ability to work independently in a remote environment.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with AI data labeling, LLM evaluation, or QA evaluation (including RLHF or preference ranking); familiarity with prompt evaluation, safety policies, and training data quality best practices; self-driven, organized, and comfortable handling iterative feedback to support model performance improvement.",{"h2":40,"desc":41},"Pay Range","USD $35–$40 per hour (remote, full-time). Final rate within the range depends on assessment performance, task complexity, and quality metrics.",{"h2":43,"desc":44},"How to Apply","Apply through Rexzone with your resume\u002FCV and a brief note confirming Germany location and English\u002FGerman fluency. Qualified candidates will complete an evaluation to assess reasoning quality, ranking consistency, and annotation guidelines compliance.","AI Data Operations"]