Remote AI Quality Analyst (Thai)
ประกาศจากแหล่งภายนอกคุณสมัครได้โดยตรง — เราจะพาคุณไปยังหน้าสมัครงานของบริษัท ไม่ต้องสมัครสมาชิก ไม่มีคนกลาง ไม่ต้องล็อกอิน ThaiJobz
รายละเอียดงาน
About the Company
Based in San Francisco, California, Turing is a research accelerator for frontier AI labs and a partner for enterprises deploying advanced AI systems. Turing accelerates research with high-quality data, training pipelines, and AI researchers, and helps enterprises transform AI from proof of concept into production intelligence.
About the Role
As an AI Quality Analyst, you will evaluate a personalization feature for Gemini by assessing how well the model uses information from past Gemini conversations, Gmail, Google Search, and YouTube activity to make responses more relevant and helpful. The role combines creative prompt design with analytical evaluation of personalized AI responses.
Responsibilities
- Design and execute multi-turn conversational prompts (typically 1–5 turns) that require the AI to utilize your personal information and experiences.
- Evaluate model responses against the intent of the starting prompt and check whether personalization was appropriately applied.
- Analyze responses for grounding issues, poor inferences, incorrect personalization, and hallucinations.
- Assess integration quality to ensure personal data is woven naturally without robotic overnarrating.
- Stack-rank two model responses side-by-side (SxS) to determine which is more helpful, usable, and enjoyable.
- Write clear, defensible rationales for comparisons, explicitly referencing specific turn numbers and providing detailed annotations and feedback.
- Extract and verify debug info to confirm chat summaries and data sources were used properly.
- Maintain strict data hygiene by deleting evaluation conversations to avoid polluting future chat history.
Qualifications
- Thai proficiency: ability to read and write Thai at a high degree of competency (Thai is the focus language for this project).
- Willingness to use your primary personal Google account and enable personal data sources for a genuine assessment.
- Full-time availability in your local time zone with schedule flexibility; part of a global, 24-hour operations team.
- Exceptional analytical thinking with experience evaluating nuanced and ambiguous AI responses and personalization quality.
- Creative prompt engineering experience, able to design multi-turn starting prompts based on personal context.
- Meticulous attention to detail and strong evaluation acumen to spot subtle differences in naturalness and overnarrating.
- Excellent written communication: able to write clear, concise, and structured rationales referencing specific turns.
- Self-motivated and able to work independently in a remote setting; good collaboration and communication skills.
- Technical setup: desktop/laptop with a good internet connection.
- BS/BA degree or equivalent experience in a relevant field (e.g., Policy, Law, Ethics, Linguistics, Journalism, Computer Science) or equivalent experience; experience in data annotation, AI quality evaluation, content moderation, or related roles strongly preferred.
Additional Information
- Commitment required: at least 4 hours per day and minimum 30 hours per week with 4 hours overlap with PST (options: 30 hrs/week or 40 hrs/week).
- Engagement type: Contractor.
- Engagement length: 3 months.
- After applying, you will receive an email with a login link to access the portal and complete your profile.
คุณสมบัติผู้สมัคร
- ประสบการณ์
- 1-2 ปี
- การศึกษา
- ไม่ระบุ
