
Domain Expert- Sports
Overview
This role focuses on improving Large Language Models (LLMs) by applying deep sports knowledge and analytical rigor. The expert will craft challenging prompts across global sports, leagues, athletes, tournaments, rules, statistics, and analytics. Each response gets checked for factual accuracy, reasoning quality, completeness, and nuance, with evidence-backed feedback. The work also involves building benchmark datasets and adversarial test cases to expose model gaps. This is a remote contractor assignment lasting 8 weeks with a 40-hour weekly commitment.
What You'll Do7
- 1Design challenging prompts that cover worldwide sports, major leagues, teams, athletes, tournaments, rules, statistics, and analytics.
- 2Review AI-generated responses and score them for factual accuracy, reasoning quality, completeness, and nuance.
- 3Identify hallucinations, logical inconsistencies, outdated information, and edge cases in model output.
- 4Create benchmark datasets and adversarial test cases to measure model performance and uncover weaknesses.
- 5Deliver evidence-based feedback backed by reliable references for every evaluation.
- 6Work alongside AI researchers to refine and improve model responses.
- 7Keep detailed documentation of annotations and quality decisions.
Requirements6
- 1Master's degree or higher in any field; a degree in sports management, sports science, journalism, communications, media studies, or a sports-related discipline is preferred.
- 2Strong command of major sports, leagues, teams, athletes, tournaments, rules, statistics, and current developments across one or more sports.
- 3At least 3 years of professional experience in sports journalism, sports media, sports content, research, analysis, reporting, or a related area.
- 4Excellent written English, research, and analytical skills for assessing sports-related information accurately.
- 5High attention to detail and a consistent approach to verifying facts.
- 6Experience with LLMs, generative AI, prompt engineering, or AI evaluation is a plus, as is published research, industry recognition, or teaching experience.
Who Should Apply
Sports professionals with a Master's degree and at least 3 years of hands-on experience in sports journalism, media, research, or analysis will fit this role well. The work suits people who enjoy scrutinizing AI output for factual errors and can back every judgment with reliable sources. It is less suitable for casual sports fans or those without strong writing and research skills. Candidates often get rejected when they rely on general impressions instead of concrete evidence, or when their assessment submission misses the depth and reference-backed detail required. Those who lack familiarity with LLMs and AI evaluation should also expect a harder fit review.
Location
Required Skills
Application Tip
During the assessment, include specific references and cite sources for every factual judgment you make, because the role centers on evidence-based feedback and factual accuracy.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Turing
VerifiedDomain Expert - Art
The role centers on evaluating and improving Large Language Models (LLMs) using deep knowledge of art history, visual arts, architecture, and design. You will design prompts that range from simple to advanced and assess model answers for factual accuracy, reasoning, completeness, and nuance. Each day involves comparing multiple AI outputs, ranking them, and explaining correct or incorrect responses with reliable references. You will also create benchmark datasets and adversarial test cases to expose hallucinations and logical gaps. The position is an 8-week contractor role requiring 40 hours per week with at least 4 hours of PST overlap.

Turing
VerifiedDomain Expert- Politics
Turing seeks a Politics Domain Expert to evaluate and improve Large Language Models with a focus on political science, governance, elections, public policy, and international relations. The role involves designing challenging prompts that test factual accuracy, reasoning, completeness, and nuance in AI-generated responses. You will build benchmark datasets and adversarial test cases to expose hallucinations, logical inconsistencies, outdated information, and edge cases. Work happens remotely on a contractor basis for 8 weeks at 40 hours per week with a 4-hour PST overlap.

Turing
VerifiedDomain Expert - TV Show & Movies
Turing is hiring a TV Shows & Movies Domain Expert for an 8-week contractor assignment focused on evaluating and improving large language models. The day-to-day work centers on creating prompts of varying difficulty across entertainment topics, then comparing and ranking AI responses. The expert will judge responses for factual accuracy, reasoning, completeness, and nuance, and will flag hallucinations, outdated information, and edge cases. This remote role is open to US-based candidates who can commit 40 hours per week with 4 hours of daily overlap with PST.

Micro1
VerifiedData Science Domain Expert for AI Evaluation and Prompting
Remote contractor role focusing on AI data science with a domain expert lens. You’ll assess AI outputs, refine prompts, and annotate data to support high-quality model training. The work centers on applying deep domain knowledge to review research-style documents, technical reports, and experiment notes, shaping how models learn and reason. Strong writing, precise attention to detail, and independent research are essential as you contribute to rubric-based evaluations and content quality.

Micro1
VerifiedAi Domain Expert Focused on Domain Knowledge and Evaluation
Remote, part-time contractor role focused on guiding AI systems through real-world input. You’ll review AI outputs for accuracy, craft prompts that test reasoning, and provide precise written feedback to improve model performance. Your domain knowledge matters most, even if you’re not required to have AI prior experience. You’ll join a global, distributed team and help shape how models learn and reason. Bolded areas reflect core competencies like data annotation, prompt engineering, and ethical awareness in AI.

Micro1
VerifiedAi Trainer for Domain Experts Remote Contract
A remote contractor role focused on shaping AI performance through real-world input. You’ll review and create examples for model outputs, collaborate with a distributed team, and provide clear written rationales to support learning objectives. Your domain knowledge in a specialized field drives the quality of prompts, tasks, and data annotations that guide AI reasoning. This project relies on precise feedback, adaptable methods, and thorough documentation of progress.

