
Quality Assurance Tester
Curvia AI • Greece
-
Role & seniority
- AI Quality Analyst (evaluation-focused, likely entry–mid level; seniority not explicitly stated)
-
Stack/tools
-
Gemini (personalisation feature evaluation)
-
User data sources: past Gemini conversations, Gmail, Google Search, YouTube activity
-
Prompt design for multi-turn conversations (typically 1–5 turns)
-
Analytical rubric dimensions: Grounding, Integration, Helpfulness
-
Desktop/laptop + reliable internet (for testing workflow)
-
-
Top 3 responsibilities
-
Design multi-turn prompts that require the model to use personal information/experience.
-
Evaluate response quality vs. the prompt intent, checking whether personalisation was applied correctly.
-
Assess Grounding: verify claims are supported by evidence and flag hallucinations/flawed inferences.
-
-
Must-have skills
-
Strong analytical reasoning for nuanced/ambiguous outputs
-
Ability to judge personalisation relevance and integration
-
Focused evaluation of grounding/helpfulness and evidence support
-
Comfortable generating prompts grounded in personal experience
-
-
Nice-to-haves
-
Experience in data annotation, AI quality evaluation, content moderation, or related work
-
BS/BA (or equivalent) in a relevant field
-
-
Location & work type
-
London, UK
-
Work type: not specified (appears remote/independent testing supported by local device + internet)
-
Full Description
Company Description
Company Description: Curvia AI is a London, UK-based technology and AI solutions company that transforms complexity into opportunity through enterprise solutions, AI-powered automation, and data-driven insights. We build practical, scalable, and human-centred technology grounded in integrity, security, and responsible innovation.
Role Overview
- As an AI Quality Analyst, you will evaluate a new personalisation feature for Gemini. You will assess how well the model uses information from your past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful. This role requires a unique blend of creativity and analytical rigour. You will actively design prompts from the perspective of your own personal experiences. You will then use your analytical skills to assess the quality of the model's personalised responses, evaluating dimensions like Grounding, Integration, and Helpfulness.
Key Qualifications
Exceptional Analytical Thinking: Demonstrate ability to evaluate nuanced and ambiguous AI responses, specifically assessing personalisation quality.
Technical Setup: Desktop/Laptop set up with a good internet connection.
Description
In this role, you will be part of a dynamic team focused on evaluating the quality of personalised AI interactions. Your day-to-day work will involve
- Designing and executing multi-turn conversational prompts (typically 1-5 turns) that require the AI to utilise your personal information and experiences.
- Evaluating model responses based on your intent from the starting prompt, checking if the personalisation was appropriately applied.
- Analysing responses for Grounding issues, ensuring claims about you are supported by evidence and not flawed inferences or hallucinations.
Education & Experience BS/BA degree or equivalent experience in a relevant field Experience in data annotation, AI quality evaluation, content moderation, or a related role is strongly preferred.
Evaluation Process: Shortlisted candidates will be sent a Job Interest Form.