Turing
TuringVerified listing
Remote

Data Scientist/Analyst

Remote
Posted August 11, 2026
contract

Overview

This contract role focuses on improving AI model performance through hands-on Python development and rigorous data analysis. You will build and maintain code for model training, run evaluations, and rank model responses across diverse domains. The work includes creating high-quality datasets for supervised fine-tuning and collaborating with researchers on RLHF efforts. The position is fully remote and requires a minimum of 20 hours per week with 4 hours of overlap with Pacific Time.

What You'll Do11

  • 1Build and maintain high-quality Python code used to train and improve AI models.
  • 2Run model evaluations to benchmark performance and convert results into actionable changes.
  • 3Assess and rank AI responses to user queries across various domains using predefined criteria.
  • 4Write clear explanations and rationales for each evaluation, showing strong technical reasoning.
  • 5Lead supervised fine-tuning efforts, including building and maintaining task-specific datasets.
  • 6Collaborate with researchers and annotators on reinforcement learning from human feedback and reward model refinement.
  • 7Design new evaluation strategies that keep model outputs aligned with user needs and ethical guidelines.
  • 8Create and refine model responses to improve clarity, relevance, and technical accuracy.
  • 9Review peer code and documentation, provide constructive feedback, and point out improvement areas.
  • 10Partner with cross-functional teams to improve model performance and support product upgrades.
  • 11Test and integrate new tools, techniques, and methodologies into AI training processes.

Requirements6

  • 1A bachelor's or master's degree in engineering, computer science, or equivalent practical experience.
  • 2Strong data analysis skills and business sense to draw conclusions from datasets, act on findings, and explain them clearly.
  • 3Proven problem-solving and analytical skills.
  • 4Clear communication skills for working with stakeholders and researchers.
  • 5Professional fluency in conversational and written English.
  • 6A strong interest in having a measurable impact on artificial intelligence.

Who Should Apply

Ideal candidates are those who enjoy turning messy datasets into clear conclusions and can defend those conclusions with code. The role suits someone comfortable writing Python for model evaluation and fine-tuning, and who can explain technical reasoning to both researchers and business stakeholders. This position is less suitable for people who prefer fixed schedules or want a permanent employment package, since it is a short-term contractor assignment with no paid leave. Common reasons candidates fall short include weak English communication during interviews or an inability to show structured analytical thinking when ranking model responses. Another frequent gap is lack of hands-on experience with evaluation or supervised fine-tuning workflows.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

pythondata analysisdata sciencemachine learningartificial intelligencesupervised fine-tuningrlhfmodel evaluationbenchmarkingllmdataset creationreward modeling

Application Tip

Before the technical interview, prepare a one-page write-up of a past model evaluation or data analysis project that covers the metrics you used, the Python code you wrote, and how your findings led to a concrete model or dataset change.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Micro1

Micro1

8d agoRemotecontract
Hot

Data Scientist for AI Model Evaluation and Quality Assurance

A remote contractor role with micro1 in partnership with a leading AI lab. You’ll apply rigorous statistics, machine learning, and quantitative reasoning to evaluate AI model outputs and improve performance. Expect to develop expert prompts, datasets, and reference solutions, while spotting methodological flaws and statistical errors. The work centers on enhancing model capabilities through structured, actionable feedback. Python, Machine Learning, and Data Analysis are core to the daily tasks and evaluation.”

245–280/hr
· 5 openings
PythonMachine LearningStatistics+4 more
Micro1

Micro1

30d agoRemotecontract
Hot

Data Scientist

This contractor role is perfect for data scientists who want to directly improve how AI models learn and reason. You'll apply your analytical skills to collect, clean, and model real-world data, helping train next-generation systems without needing prior AI experience—your domain expertise is the key ingredient. Work remotely and contribute to cutting-edge projects that turn raw data into smarter, more capable AI.

Competitive salary
· 10 openings
Statistics & MathematicsData HandlingData Collecting+4 more
SME Careers

SME Careers

6h agoRemotecontract

Python Data Scientist for AI Reasoning Review

Work remotely as a contract Data Scientist who reviews AI-generated analytical reasoning, code, and model outputs, and crafts precise reference solutions for data problems. You will assess prompts for accuracy and clarity, then write step-by-step explanations that demonstrate correct methods. Rate and compare AI responses on correctness and reasoning quality to guide improvements in machine learning workflows. This remote, hourly role supports projects in data science and analytics for a fast-growing AI data services company.

Up to 100/hr
PythonPandasNumPy+23 more
Turing

Turing

14d agoRemotecontract

Senior Python Developer

This contract role supports a foundational LLM company by producing high-quality data used to fine-tune and benchmark their models. You will write Python solutions to code-based prompts, evaluate responses from two model versions, and create detailed rationale for ranking decisions. Day-to-day work involves SFT and RLHF data generation, designing evaluation strategies, and reviewing code with a small team. The position does not involve building or fine-tuning models directly, but your outputs directly inform their improvement.

Competitive salary
PythonLLMSft+8 more
Micro1

Micro1

1mo agoRemotecontract
Hot

Data Science Domain Expert for AI Evaluation and Prompting

Remote contractor role focusing on AI data science with a domain expert lens. You’ll assess AI outputs, refine prompts, and annotate data to support high-quality model training. The work centers on applying deep domain knowledge to review research-style documents, technical reports, and experiment notes, shaping how models learn and reason. Strong writing, precise attention to detail, and independent research are essential as you contribute to rubric-based evaluations and content quality.

100–200/hr
· 15 openings
Critical ThinkingAnalytical ReasoningAttention To Detail+22 more
SME Careers

SME Careers

7h agoRemotecontract

R Engineer for Data Analysis and AI Evaluation

Remote, hourly contract role for an experienced R engineer focused on data analysis tasks in an AI context. You will review AI-generated responses and generate high-quality R content that demonstrates correct reasoning using reproducible workflows such as R Markdown. You’ll assess accuracy, clarity, and adherence to prompts, and identify missteps in statistical methods or data-wrangling within analyses. Document expert explanations and model solutions that showcase proper R usage, then rate and compare AI outputs to guide future work. This engagement has no immediate project and offers opportunities to join SME Careers for future roles within the SuperAnnotate network.

Up to 55/hr
RData AnalysisData Cleaning+30 more