
AI Quality Analyst (Personalization) - Thai
Overview
Turing is assembling a global team of Thai-speaking quality analysts to evaluate a new Gemini personalization feature. Analysts design 1-5 turn prompts from their own experience and test how well the model uses data from Gmail, Google Search, and YouTube. Each day involves stack-ranking two model responses side-by-side, scoring for grounding, integration, and helpfulness, and writing rationales that reference specific turn numbers. The role requires using a primary personal Google account, full-time availability in your local time zone, and a 4-hour overlap with PST.
What You'll Do8
- 1Design multi-turn prompts (1-5 turns) that force the model to rely on your personal data and history.
- 2Judge whether each response stays true to the intent of your starting prompt and applies personalization without overreaching.
- 3Check responses for grounding issues, flagging claims that lack support from your personal data or that come from flawed inference.
- 4Assess the natural integration of your personal information, watching for robotic overnarrating.
- 5Compare two model outputs side-by-side and rank them by overall helpfulness, ease of use, and enjoyment.
- 6Write concise rationales for each ranking, citing the exact turn numbers where strengths or issues appear.
- 7Extract and verify Debug Info to confirm chat summaries and data sources are used as expected.
- 8Delete each evaluation conversation after scoring to keep your future chat history clean.
Requirements12
- 1Professional proficiency in reading and writing Thai, since Thai is the project's focus language.
- 2Willingness to use your primary personal Google account and enable personal data sources for genuine assessment.
- 3Full-time schedule availability in your local time zone with at least 4 hours of overlap with PST.
- 4Strong analytical thinking to evaluate nuanced and ambiguous AI responses, especially around personalization.
- 5Experience designing creative, multi-turn starting prompts from personal context.
- 6Understanding of personalization quality signals, including incorrect personalization, poor inferences, and forced connections.
- 7Ability to spot subtle differences in naturalness and overnarrating when reviewing side-by-side responses.
- 8Excellent written communication to produce concise, structured rationales that cite specific turn numbers.
- 9Self-motivated working style, with the discipline to complete tasks without supervision in a remote setting.
- 10A reliable desktop or laptop with a good internet connection.
- 11BS/BA degree or equivalent in a relevant analytical field such as linguistics, journalism, policy, or computer science.
- 12Preferred experience in data annotation, AI quality evaluation, or content moderation.
Who Should Apply
The ideal candidate for this role is a Thai-speaking evaluator who enjoys creative prompt design and rigorous analytical review. You should be comfortable letting Gemini access your personal Google data, because the entire evaluation depends on real personal context. Someone who is unwilling to use their primary Google account or cannot commit to a 4-hour PST overlap will struggle with the core requirements. Candidates often get screened out for not passing the Thai proficiency check, or for submitting the assessment after the 24-hour window. A background in linguistics, data annotation, or content moderation helps but is not a hard requirement.
Salary Insight
The offered rate is $15 per hour.
Location
Required Skills
Application Tip
During the assessment, write rationales that cite exact turn numbers and explain why one response is more grounded and helpful, because the evaluation rubric puts high weight on that skill.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Turing
VerifiedAI Quality Analyst (Personalization) - Vietnamese
This role puts you inside the evaluation loop for a new personalization feature in Gemini. You will design multi-turn prompts (1-5 turns) that draw on your own Google account data, including past chats, Gmail, Search, and YouTube history. Your core job is to judge whether the model's personalized responses are relevant, grounded in real evidence, and natural, using dimensions like Grounding and Integration. Work is conducted in Vietnamese, comparing two model outputs side-by-side and writing rationales that cite exact conversation turns. The engagement lasts 3 months and pays $15 per hour, with a daily commitment that includes PST overlap.

Turing
VerifiedAI Quality Analyst (Personalization) - Indonesian
This role sits on a global team that checks how well Gemini tailors responses using a user's own history, including past chats, Gmail, Google Search, and YouTube activity. As an AI Quality Analyst, you craft personal, multi-turn prompts in Indonesian and then judge whether the model grounded its replies in your stated context. You also rank two model responses side by side to decide which is more helpful and natural. The work is remote, but you must keep a set schedule that overlaps with PST hours.

Turing
VerifiedAI Quality Analyst (Personalization) - Korean
This role evaluates a new personalization feature for Gemini that draws on past conversations, Gmail, Google Search, and YouTube activity. You will design multi-turn prompts from your own personal context and then judge how well the model grounds responses in real evidence rather than inferences. The work includes side-by-side (SxS) ranking of two model outputs, writing structured rationales that cite specific turn numbers, and checking Debug Info to verify chat summaries and data sources. Korean is the focus language, so reading and writing fluency in Korean is required. The position is remote, full-time with at least 4 hours of daily overlap with PST, and runs as a 3-month contractor engagement.

Turing
VerifiedAI Quality Analyst - English
This role puts you inside Gemini's personalization quality loop. You will design multi-turn prompts that draw on your own Google activity, including Gmail, Search, and YouTube history, then judge whether the model's responses are grounded, integrated, and helpful. The work involves side-by-side SxS evaluations, writing rationales that reference specific turns, and verifying debug info to confirm the model used your data correctly. This 3-month contractor engagement requires at least 4 hours per day with a 4-hour overlap with PST.

Turing
VerifiedAI Quality Analyst (Personalization) - Turkish
In this contract role, you will evaluate a new personalization feature for Gemini. You will measure how well the model incorporates signals from your past conversations, Gmail, Google Search, and YouTube activity to make responses more relevant. You will design multi-turn prompts based on your own experiences and then judge the model's output on dimensions like Grounding, Integration, and Helpfulness. The work demands both creative prompt design and disciplined analytical review of model responses.

Turing
VerifiedAI Quality Analyst (Personalization) - Spanish
Evaluators in this role test Gemini's new personalization feature by measuring how well the model uses past conversations, Gmail, Google Search, and YouTube activity to tailor responses. The work blends creative prompt writing with structured quality review. Each evaluation prompt starts from the evaluator's own experiences and runs 1-5 turns. You will score responses on Grounding, Integration, and Helpfulness, then compare them side-by-side and justify each ranking. The project centers on Spanish, and you must use your personal Google account for authentic data access.

