Thriveth

The AI job market, made clearer

Find work worth
your expertise.

Back to all jobs

AI Quality Analyst (Gemini) - Chinese

As an AI Quality Analyst, you will evaluate a new personalization feature for Gemini.

Remote
Hiring companyTuring
Application processN/ANo approved reviews yet
Work experienceN/ANo approved reviews yet

About the role

You will assess how well the model uses information from your past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful. This role requires a unique blend of creativity and analytical rigor. You will actively design prompts from the perspective of your own personal experiences. You will then use your analytical skills to assess the quality of the model's personalized responses, evaluating dimensions like Grounding, Integration, and Helpfulness.

Engagement type: Contractor.

Responsibilities

  • Designing and executing multi-turn conversational prompts (typically 1-5 turns) that require the AI to utilize your personal information and experiences.
  • Analyzing responses for Grounding issues, ensuring claims about you are supported by evidence and not flawed inferences or hallucinations.
  • Rigorously evaluating and stack-ranking two model responses side-by-side (SxS) to determine which is overall more helpful, easy to use, and enjoyable.
  • Writing clear, defensible rationales for your comparisons, explicitly referencing where issues or positive aspects occurred in the conversation.
  • Extracting and verifying "Debug Info" from the model to confirm that chat summaries and data sources were properly utilized.
  • Maintaining strict data hygiene by deleting evaluation conversations to prevent them from polluting your future chat history.

Experience and education

  • Creative Prompt Engineering: Experience in designing creative, multi-turn starting prompts based on personal context to thoroughly test the model's capabilities.
  • Strong Evaluation Acumen: Understanding of personalization concepts, including the ability to identify incorrect personalization, poor inferences, and forced connections.
  • Meticulous Attention to Detail: The ability to review Side-by-Side (SxS) model responses and spot subtle differences in naturalness and overnarrating.
  • Excellent Written Communication: Superior ability to write clear, concise, and structured rationales for model rankings, explicitly referencing specific turn numbers.
  • Feedback: Ability to provide constructive feedback and detailed annotations.
  • BS/BA degree or equivalent experience in a relevant field (e.g., Policy, Law, Ethics, Linguistics, Journalism, Computer Science, or a related analytical field).
  • Experience in data annotation, AI quality evaluation, content moderation, or a related role is strongly preferred.

Language requirements

  • Chinese Proficiency: Ability to read and write in Chinese with a high degree of comp, as Chinese is the focus language for this project.

Assessment requirements

Shortlisted candidates will be sent a Job Interest Form.
After the profile review, an assessment will be shared, which must be completed within 24 hours.
Based on the assessment outcomes, shortlisted candidates will be contacted to discuss the pre‑onboarding requirements.

Schedule details

Commitments : Required: at least 4 hours per day and up to 40 hours per week with 4 hours of overlap with PST.

Engagement Length: 3 months.

Eligibility

Work arrangement: Fully remote

Compensation details

Our offered rate for this project is $15 per hour.

More about the hiring companyTuring
Find AI-training work. Know what to expect.

Thriveth makes AI data-training work easier to find, understand, and navigate. We replace uncertainty with clear opportunities, realistic expectations, and insights from real application journeys.