[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-trainer-onsite-jobs-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":42},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, and you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation tasks such as evaluation and ranking of model outputs, QA evaluation and validation of labels, prompt evaluation, writing rationales, and ensuring training data quality through annotation guidelines compliance.","What tasks will I do?",{"A":14,"Q":15},"AI\u002Fannotation experience is preferred but not required. Strong analytical reasoning, attention to detail, and the ability to follow guidelines consistently are essential.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in both German and English is required, including the ability to read, write, and evaluate nuanced content in both languages.","What languages are required?",{"A":20,"Q":21},"Tasks can span general knowledge, writing quality, instruction-following, content safety labeling, and other areas that impact training data quality and model performance improvement.","What domains are covered?","ai-trainer-onsite-jobs-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based, bilingual (German\u002FEnglish) AI Generalist Trainers to support RLHF and large language model evaluation by ranking model outputs, writing clear rationales, and driving training data quality for model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will evaluate and improve AI\u002FLLM systems through RLHF-style feedback workflows. You will assess, rank, and QA model-generated responses in both German and English, document reasoning, and ensure annotation guidelines compliance to strengthen training data quality and enable model performance improvement. This is a remote, full-time role focused on large language model evaluation, prompt evaluation, and training data quality validation.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation by reviewing, scoring, and ranking model outputs; write concise, evidence-based rationales explaining preferences and reasoning; execute QA evaluation to validate labels, resolve edge cases, and flag inconsistencies; follow annotation guidelines compliance requirements and maintain high training data quality; conduct prompt evaluation and content safety labeling when required; validate tasks across German and English datasets, ensuring linguistic accuracy and policy adherence; escalate ambiguous cases, propose guideline improvements, and support calibration to improve inter-annotator consistency.",{"h2":34,"desc":35},"Basic Qualifications","Based in Germany and authorized to work as required for a remote role; fluent in German and English (reading, writing, and critical reasoning); strong analytical skills with the ability to compare answers and justify rankings; exceptional attention to detail and ability to follow annotation guidelines compliance standards; comfort working with structured rubrics, QA evaluation checklists, and high-throughput evaluation queues.",{"h2":37,"desc":38},"Preferred Qualifications","Experience with data labeling, RLHF, prompt evaluation, or LLM evaluation; familiarity with common LLM failure modes (hallucinations, instruction-following issues, safety\u002Fpolicy violations); prior work with annotation guidelines, calibration sessions, and training data quality programs; self-driven, reliable, and able to manage time effectively in a remote environment.",{"h2":40,"desc":41},"How to Apply","Apply through Rexzone with a brief summary of your bilingual (German\u002FEnglish) experience and any relevant evaluation, data labeling, or QA evaluation background. Selected candidates may complete a short skills assessment focused on ranking, reasoning, and annotation guidelines compliance.","AI Data Operations"]