Skip to content
NewEligibility scores are liveSee if you qualify
AI & dataOpen · checked today

AI Quality Analyst (Personalization) - Hindi

We are looking for AI Quality Analyst (Personalization) - Hindi candidates for a project delivered through Turing.

What you'll do

  • Evaluate a new personalization feature for Gemini.
  • Assess how well the model uses information from past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful.
  • Design prompts from the perspective of personal experiences.
  • Assess the quality of the model's personalized responses, evaluating dimensions like Grounding, Integration, and Helpfulness.
  • Design and execute multi-turn conversational prompts (typically 1-5 turns) that require the AI to utilize personal information and experiences.
  • Evaluate model responses based on intent from the starting prompt, checking if personalization was appropriately applied.
  • Analyze responses for Grounding issues, ensuring claims about the user are supported by evidence and not flawed inferences or hallucinations.
  • Assess Integration quality to ensure personal data is woven naturally into the response without robotic "overnarrating".
  • Rigorously evaluate and stack-rank two model responses side-by-side (SxS) to determine which is overall more helpful, easy to use, and enjoyable.
  • Write clear, defensible rationales for comparisons, explicitly referencing where issues or positive aspects occurred in the conversation.
  • Extract and verify "Debug Info" from the model to confirm that chat summaries and data sources were properly utilized.
  • Maintain strict data hygiene by deleting evaluation conversations to prevent them from polluting future chat history.

What you need

  • Ability to read and write in Hindi with a high degree of competence.
  • Willingness to use a primary personal Google account and enable personal data sources for a genuine assessment.
  • Full-time availability in the local time zone.
  • Ability to evaluate nuanced and ambiguous AI responses, specifically assessing personalization quality.
  • Experience in designing creative, multi-turn starting prompts based on personal context to thoroughly test the model's capabilities.
  • Understanding of personalization concepts, including the ability to identify incorrect personalization, poor inferences, and forced connections.
  • Ability to review Side-by-Side (SxS) model responses and spot subtle differences in naturalness and overnarrating.
  • Superior ability to write clear, concise, and structured rationales for model rankings, explicitly referencing specific turn numbers.
  • Ability to provide constructive feedback and detailed annotations.
  • Excellent communication and collaboration skills.
  • Self-motivated and able to work independently in a remote setting.
  • Desktop/Laptop setup with a good internet connection.
  • BS/BA degree or equivalent experience in a relevant field (e.g., Policy, Law, Ethics, Linguistics, Journalism, Computer Science, or a related analytical field).

Nice to have

  • Experience in data annotation, AI quality evaluation, content moderation, or a related role.

Expertise

RequiredAlso useful

Also listed: analytical thinking, attention to detail, written communication.

Who you work with

Project and contracting process: Turing. Applications continue on the provider's website.

More ai & data roles

See all

$15/hr

Turing · Worldwide

Apply