Mercor
MercorVerified listing
Remote

Bilingual Italian-English AI Safety Evaluator

40–44/hr
Remote · Remote — Western Europe preferred
Posted September 4, 2026
part-time

Overview

Help advance AI safety by evaluating how models respond to sensitive topics in Italian. Your linguistic and cultural judgment will shape which prompts seem adversarial and which conversations escalate. Since no AI or ML expertise is needed, we'll train you on our guidelines and workflow. This remote, part-time role suits someone with excellent written reasoning who can document their decisions clearly. While Western Europe is preferred, we'll consider applicants from other regions.

What You'll Do4

  • 1Craft high-quality Italian prompts on a variety of sensitive topics to test AI responses.
  • 2Apply structured classification systems to sort prompts and interactions according to risk.
  • 3Identify tricky phrasings that might bypass safeguards and spot potential escalation patterns.
  • 4Record the rationale for your assessments in clear written English.

Requirements6

  • 1Native or near-native Italian fluency, with business-level written English.
  • 2Pursuing or holding a bachelor's degree in any field.
  • 3Strong attention to detail and ability to explain reasoning concisely.
  • 4Sound judgment when dealing with sensitive or dual-use information.
  • 5Prior experience in content review, grading, or red-teaming is a plus.
  • 6Background in trust and safety or policy evaluation is advantageous.

Who Should Apply

The ideal candidate is fluent in both Italian and English, comfortable writing about difficult topics, and able to think critically about how language can provoke AI. People with a background in moderation, quality assurance, or adversarial testing will find this work intuitive. This role isn't for someone who shies away from sensitive material or struggles to articulate nuanced decisions. Common rejection reasons include insufficient Italian proficiency or failing to demonstrate judgment in the application.

Salary Insight

The rate is $40.00 to $44.00 per hour, depending on experience. That sits at the higher end for remote language evaluation work, reflecting the need for cultural sensitivity and careful reasoning. Pay is discussed later in the process.

Location

Typeremote
LocationRemote — Western Europe preferred
This is a remote position

Required Skills

italianenglishcontent moderationprompt writingred-teamingtrust and safetyrisk assessmentquality assurance

Application Tip

If you have prior experience with content moderation or safety reviews, describe a specific case you handled in Italian. Alternatively, include a short writing sample that shows how you'd assess a risky prompt.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

11h agoRemotepart-time

Bilingual Italian STEM Expert (PhD) — AI Safety

Apply your PhD-level grasp of chemistry or biology to improve AI safety in Italian. You'll write prompts that probe specialized knowledge, then evaluate model responses for correctness and responsible content. No prior AI/ML experience is necessary; training covers the workflow. Expect a part-time, remote schedule of about seven hours weekly, with an immediate start. Candidates based in Western Europe receive preference, but all locations are considered.

50–54/hr
ItalianEnglishChemistry+6 more
Mercor

Mercor

11h agoRemotepart-time

Bilingual German Generalist Expert — AI Safety

This remote, part-time role puts your German language skills to work making advanced AI models safer. You'll pair with English fluency to craft, classify, and judge how AI handles delicate topics in German, no prior machine learning background required. The project starts immediately and demands about 7 hours per week, with training provided on their evaluation workflow from day one. Your cultural instincts and written judgment matter more than technical credentials here.

48–52/hr
· 2 openings
GermanEnglishGerman Fluency+10 more
SME Careers

SME Careers

6h agoRemotecontract

Italian Language Expert for AI Content QA Remote Contract

This hourly remote contractor role leverages your mastery of Italian and linguistics to review AI-generated Italian responses and craft expert language content. You'll evaluate reasoning quality and deliver step-by-step edits with precise, written feedback to guide model improvements. You will apply localization principles and style guides to ensure consistent terminology across audiences. There is no immediate project, but Italy-based experts will be contacted first when relevant opportunities arise.

Up to 37/hr
ItalianLinguisticsLocalization+21 more
Micro1

Micro1

21d agoRemotecontract
Hot

Italian Bilingual Expert for AI Training and Audio Evaluation

A remote contractor role for native Italian speakers who can help train AI by evaluating Italian audio. You’ll listen to Italian clips, judge nativeness and fluency, and provide written feedback in English. Your domain knowledge takes center stage, and no previous AI work is required. You’ll work with project coordinators to clarify needs and contribute ideas to improve Italian language understanding in models. Expect clear guidelines and timely task completion that upholds high quality standards. Bold Italian and linguistic assessment as you apply your expertise.

30–65/hr
· 100 openings
English (b2+)Attention To DetailWritten Communication+2 more
Mercor

Mercor

11h agoRemotepart-time

Bilingual Polish Generalist Expert — AI Safety

Your Polish and English fluency drives this part-time remote role, where you help make AI models safer. You'll write detailed prompts in Polish and apply structured rules to judge how models respond to sensitive or adversarial topics. The job doesn't need a machine learning background; the team provides complete training. You can work from anywhere, though being in Poland or Eastern Europe is a plus, and they want a quick start.

40–44/hr
· 6 openings
PolishEnglishPrompt Engineering+5 more
Mercor

Mercor

11h agoRemotepart-time

Bilingual Japanese Generalist Expert — AI Safety

Join a remote team improving AI safety by testing how models respond to sensitive prompts in Japanese. Your linguistic and cultural expertise will shape safer outputs. Mercor provides full training, so no machine learning experience is necessary. Work begins right away at about seven hours per week.

48–52/hr
· 3 openings
JapaneseEnglishAI Safety+7 more