[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"job-ai-generalist-trainer-remote-hamburg-germany":3},{"Ques":4,"Slug":22,"Header":23,"job_category":45},{"title":5,"content":6},"Frequently Asked Questions",[7,10,13,16,19],{"A":8,"Q":9},"Yes. This is a remote, full-time role, but you must be based in Germany.","Is this role remote?",{"A":11,"Q":12},"You will perform large language model evaluation, including evaluation and ranking of model outputs, prompt evaluation, QA evaluation, content safety labeling, validation checks, and writing rationales that reflect clear reasoning and annotation guidelines compliance.","What tasks will I do?",{"A":14,"Q":15},"AI experience is preferred but not required. You do need strong analytical skills, attention to detail, and the ability to follow guidelines to produce consistent training data quality.","Do I need AI experience?",{"A":17,"Q":18},"Fluency in English and German is required, including strong reading and writing skills in both languages.","What languages are required?",{"A":20,"Q":21},"You will evaluate general-domain prompts and responses across areas such as everyday knowledge, writing quality, helpfulness, factuality checks, and content safety, all focused on model performance improvement within RLHF workflows.","What domains are covered?","ai-generalist-trainer-remote-hamburg-germany",{"desc":24,"title":25,"content":26},"Rexzone is hiring Germany-based AI Generalist Trainers to support large language model evaluation through RLHF workflows, data labeling, and QA evaluation. You will assess, rank, and validate model outputs in English and German, write clear rationales, and follow annotation guidelines compliance to improve training data quality and drive model performance improvement.","Germany-Based English & German AI Generalist Trainer 2026 May",[27,30,33,36,39,42],{"h2":28,"desc":29},"About the Role","As a Germany-based English & German AI Generalist Trainer at Rexzone, you will contribute to large language model evaluation by reviewing and ranking model-generated responses, performing prompt evaluation, and documenting reasoning. Your work directly supports RLHF pipelines and training data quality, helping enable model performance improvement across multilingual use cases.",{"h2":31,"desc":32},"Key Responsibilities","Perform large language model evaluation by comparing and ranking outputs; execute QA evaluation to validate accuracy, relevance, and safety; write concise rationales explaining reasoning and decision criteria; apply annotation guidelines compliance for consistent labeling; conduct content safety labeling and policy-based assessments; identify edge cases, ambiguities, and failure modes and flag them with evidence; validate tasks through spot checks and cross-review to maintain training data quality; track issues and propose rubric updates that improve model performance improvement.",{"h2":34,"desc":35},"Basic Qualifications","Must be based in Germany and eligible to work remotely from Germany. Fluency in English and German (reading and writing) is required. Strong analytical skills, attention to detail, and the ability to explain reasoning clearly. Comfortable following detailed annotation guidelines compliance, meeting quality targets, and working independently in structured evaluation queues.",{"h2":37,"desc":38},"Preferred Qualifications","Experience with RLHF, LLM evaluation, data labeling, or QA evaluation in an annotation environment. Familiarity with prompt evaluation, rubric-based ranking, and writing high-quality rationales. Knowledge of content safety labeling and training data quality practices. Self-driven, organized, and able to learn evolving guidelines quickly.",{"h2":40,"desc":41},"Compensation","USD $35–$40 per hour (hourly), full-time, remote. Exact rate is based on demonstrated evaluation quality, language proficiency, and role fit.",{"h2":43,"desc":44},"How to Apply","Apply to Rexzone with a resume\u002FCV highlighting bilingual English\u002FGerman experience and any AI, annotation, evaluation, or QA work. Shortlisted candidates may complete a timed assessment covering ranking, reasoning, and annotation guidelines compliance.","AI Data Operations"]