Data Annotator
<div class="show-more-less-html__markup show-more-less-html__markup--clamp-after-5 relative overflow-hidden"> <p><strong>Key Qualifications</strong></p><ul><li>German Proficiency: Ability to read and write in German with a high degree of comp, as German is the focus language for this project.</li><li><strong>Personal Account Usage: Willingness to use your primary personal Google account (not a testing account) and enable personal data sources for a genuine assessment.</strong></li><li><strong>Schedule Flexibility: Full-time availability in your local time zone is required. We are staffing a global, 24-hour operations team.</strong></li><li><strong>Exceptional Analytical Thinking: Demonstrate ability to evaluate nuanced and ambiguous AI responses, specifically assessing personalization quality.</strong></li><li><strong>Creative Prompt Engineering: Experience in designing creative, multi-turn starting prompts based on personal context to thoroughly test the model's capabilities.</strong></li><li><strong>Strong Evaluation Acumen: Understanding of personalization concepts, including the ability to identify incorrect personalization, poor inferences, and forced connections.</strong></li><li><strong>Meticulous Attention to Detail: The ability to review Side-by-Side (SxS) model responses and spot subtle differences in naturalness and overnarrating.</strong></li><li><strong>Excellent Written Communication: Superior ability to write clear, concise, and structured rationales for model rankings, explicitly referencing specific turn numbers.</strong></li><li><strong>Feedback: Ability to provide constructive feedback and detailed annotations.</strong></li><li><strong>Communication: Excellent communication and collaboration skills.</strong></li><li><strong>Independence: Self-motivated and able to work independently in a remote setting.</strong></li><li><strong>Technical Setup: Desktop/Laptop set up with a good internet connection.</strong></li></ul><p><strong>Description:</strong></p><ul><li><strong>In this role, you will be part of a dynamic team focused on evaluating the quality of personalized AI interactions. Your day-to-day work will involve:</strong></li><li><strong>Designing and executing multi-turn conversational prompts (typically 1-5 turns) that require the AI to utilize your personal information and experiences.</strong></li><li><strong>Evaluating model responses based on your intent from the starting prompt, checking if the personalization was appropriately applied.</strong></li><li><strong>Analyzing responses for Grounding issues, ensuring claims about you are supported by evidence and not flawed inferences or hallucinations.</strong></li><li><strong>Assessing Integration quality to ensure personal data is woven naturally into the response without robotic "overnarrating".</strong></li></ul><p></p> </div>