[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-evaluator-hybrid-jobs-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will evaluate and rank model outputs, perform QA evaluation and validation, write reasoning-based rationales, and follow annotation guidelines compliance to improve training data quality and model performance improvement.","What tasks will I do?",{"A":14,"Q":15},"AI experience is helpful but not required. We value analytical judgment, attention to detail, and the ability to learn RLHF and large language model evaluation workflows.","Do I need AI experience?",{"A":17,"Q":18},"Professional fluency in both English and German is required for bilingual evaluation and writing tasks.","What languages are required?",{"A":20,"Q":21},"You may cover general knowledge, writing quality, reasoning, summarization, instruction following, and content safety labeling across English and German prompts within large language model evaluation.","What domains are covered?","ai-evaluator-hybrid-jobs-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based English and German AI Generalist Trainers to support AI\u002FLLM workflows through RLHF, large language model evaluation, and training data quality improvements. You will evaluate, rank, and QA model-generated responses, write clear rationales, and validate outputs against annotation guidelines compliance to drive model performance improvement. This remote, full-time role requires bilingual fluency (English + German) and strong analytical judgment to produce consistent, high-quality evaluation signals for training data quality and safety.","Germany-Based English & German AI Generalist Trainer (Remote, Full-Time) 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As an AI Generalist Trainer at Rexzone, you will assess and rank model outputs, perform QA evaluation, and provide reasoning-heavy feedback used in RLHF and large language model evaluation pipelines. Your work directly impacts training data quality, annotation guidelines compliance, and model performance improvement across multilingual (English\u002FGerman) tasks.",{"h2":31,"desc":32},"Responsibilities","Evaluate and rank AI-generated responses for accuracy, completeness, helpfulness, and policy alignment; perform QA evaluation and validation to ensure consistency with annotation guidelines compliance; write concise rationales that explain reasoning and support reviewer auditability; identify edge cases, ambiguity, and failure patterns to improve training data quality; review bilingual (English\u002FGerman) content and apply content safety labeling when required; follow prompt evaluation protocols and maintain high throughput without sacrificing quality; escalate unclear instructions, propose guideline clarifications, and support calibration to improve inter-annotator agreement.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and authorized to work there; fluent in English and German (professional reading and writing required); strong analytical skills and structured reasoning; exceptional attention to detail and consistency under guidelines; comfortable making judgment calls and documenting rationale; reliable internet connection and ability to work remotely in a full-time schedule.",{"h2":37,"desc":38},"Preferred Qualifications","Prior experience with data labeling, prompt evaluation, or QA evaluation; familiarity with RLHF concepts and large language model evaluation; experience applying content safety labeling and taxonomy-based policies; strong self-driven workflow management, responsiveness to feedback, and comfort with iterative guideline updates.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly pay), full-time, remote.",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with an English (or bilingual) resume\u002FCV and a brief note describing your experience with evaluation, ranking, QA, and writing rationales. Qualified candidates may be asked to complete a short bilingual assessment aligned to annotation guidelines compliance.","AI Data Operations"]