
AI Quality Analyst (Personalization) - Indonesian
Overview
This role sits on a global team that checks how well Gemini tailors responses using a user's own history, including past chats, Gmail, Google Search, and YouTube activity. As an AI Quality Analyst, you craft personal, multi-turn prompts in Indonesian and then judge whether the model grounded its replies in your stated context. You also rank two model responses side by side to decide which is more helpful and natural. The work is remote, but you must keep a set schedule that overlaps with PST hours.
What You'll Do6
- 1Design 1-5 turn conversational prompts that force the model to use personal context from your own Google account.
- 2Evaluate whether the model's responses match the intent of your starting prompt and use personal information as expected.
- 3Check model output for grounding problems, such as unsupported claims about you, flawed inferences, or hallucinations.
- 4Rank two side-by-side responses on helpfulness, ease of use, and enjoyment while writing clear rationales that reference specific turn numbers.
- 5Extract and verify the model's debug info to confirm that chat summaries and data sources were applied correctly.
- 6Delete evaluation conversations after each session to keep your personal chat history clean.
Requirements9
- 1Native or professional proficiency in written and spoken Indonesian.
- 2Willingness to use a personal Google account (not a test account) and enable personal data sources for a realistic assessment.
- 3Ability to work at least 4 hours per day, 30 or 40 hours per week, with 4 hours of overlap with PST scheduling.
- 4Experience designing creative, multi-turn prompts based on personal context.
- 5Skill in evaluating ambiguous AI responses and identifying incorrect personalization, poor inferences, or forced connections.
- 6Attention to detail when reviewing side-by-side responses and spotting subtle differences in naturalness or overnarrating.
- 7Strong written communication for producing concise, structured rationales that cite specific turn numbers.
- 8A BS/BA degree or equivalent experience in a field like policy, law, ethics, linguistics, journalism, or computer science.
- 9Prior experience in data annotation, AI quality evaluation, or content moderation.
Who Should Apply
The ideal candidate has sharp analytical instincts and enjoys reading model output line by line to find hidden flaws. This role suits people who can write creative personal prompts in Indonesian and then defend their rankings with clear written arguments. People who are uncomfortable connecting their personal Google account to an AI evaluation task, or who prefer open-ended schedules, should look elsewhere. Candidates often get rejected when they submit vague rationales that do not reference exact conversation turns, or when they cannot demonstrate experience with data annotation or AI quality review.
Salary Insight
The offered rate is $15 per hour for a 3-month contractor engagement.
Location
Required Skills
Application Tip
Include a short sample of a side-by-side evaluation you have written, showing how you reference specific turn numbers and justify your ranking.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Turing
VerifiedAI Quality Analyst (Personalization) - Thai
Turing is assembling a global team of Thai-speaking quality analysts to evaluate a new Gemini personalization feature. Analysts design 1-5 turn prompts from their own experience and test how well the model uses data from Gmail, Google Search, and YouTube. Each day involves stack-ranking two model responses side-by-side, scoring for grounding, integration, and helpfulness, and writing rationales that reference specific turn numbers. The role requires using a primary personal Google account, full-time availability in your local time zone, and a 4-hour overlap with PST.

Turing
VerifiedAI Quality Analyst (Personalization) - Turkish
In this contract role, you will evaluate a new personalization feature for Gemini. You will measure how well the model incorporates signals from your past conversations, Gmail, Google Search, and YouTube activity to make responses more relevant. You will design multi-turn prompts based on your own experiences and then judge the model's output on dimensions like Grounding, Integration, and Helpfulness. The work demands both creative prompt design and disciplined analytical review of model responses.

Turing
VerifiedAI Quality Analyst (Personalization) - Vietnamese
This role puts you inside the evaluation loop for a new personalization feature in Gemini. You will design multi-turn prompts (1-5 turns) that draw on your own Google account data, including past chats, Gmail, Search, and YouTube history. Your core job is to judge whether the model's personalized responses are relevant, grounded in real evidence, and natural, using dimensions like Grounding and Integration. Work is conducted in Vietnamese, comparing two model outputs side-by-side and writing rationales that cite exact conversation turns. The engagement lasts 3 months and pays $15 per hour, with a daily commitment that includes PST overlap.

Turing
VerifiedAI Quality Analyst (Personalization) - Polish
Evaluate a new Gemini personalization feature by testing how well the model draws on past conversations, Gmail, Google Search, and YouTube activity to tailor responses. You will design multi-turn prompts from your own personal experiences, then score outputs on Grounding, Integration, and Helpfulness. The position is a contractor role that requires Polish reading and writing proficiency and at least 4 hours of daily overlap with Pacific Time. Work happens remotely on your own device, and every evaluation conversation must be deleted afterward to keep your personal history clean.

Turing
VerifiedAI Quality Analyst - English
This role puts you inside Gemini's personalization quality loop. You will design multi-turn prompts that draw on your own Google activity, including Gmail, Search, and YouTube history, then judge whether the model's responses are grounded, integrated, and helpful. The work involves side-by-side SxS evaluations, writing rationales that reference specific turns, and verifying debug info to confirm the model used your data correctly. This 3-month contractor engagement requires at least 4 hours per day with a 4-hour overlap with PST.

Turing
VerifiedAI Quality Analyst (Personalization) - Hindi
Turing runs a global evaluation team for Gemini's personalization feature, and this contract role focuses on Hindi responses. You will create multi-turn prompts from your own Google history, including Gmail, Search, YouTube, and past conversations, then rate how well the model uses that context. Scoring centers on Grounding, Integration, and Helpfulness, with side-by-side response comparisons and written justifications. The role requires at least 4 hours per day with PST overlap, up to 40 hours per week, for a 3-month contract.

