Mercor
MercorVerified listing
RemoteAI & Machine LearningCybersecurity

AI Safety Experts — English & Indonesian | $17-$25/hr Remote

17–25/hr
Remote
Posted August 5, 2026
hourly
5 openings

Overview

Neon is looking for bilingual AI safety experts fluent in English and Indonesian to join a human-driven red team. You'll probe advanced conversational AI models for vulnerabilities like bias, misinformation, and harmful behavior — all through text-based work. This remote, hourly role pays $17–$25/hr and gives you a direct hand in making AI systems safer before they reach the public.

What You'll Do4

  • 1Red team conversational AI agents by designing adversarial inputs — jailbreaks, prompt injections, misuse cases, and multi-turn manipulation.
  • 2Generate high-quality human data by annotating model failures, classifying vulnerability types, and flagging systemic risks.
  • 3Follow structured taxonomies, benchmarks, and playbooks to ensure consistent and repeatable testing.
  • 4Document attack cases, produce reproducible reports, and deliver datasets that customers can act on to strengthen their models.

Requirements3

  • 1Native-level fluency in both English and Indonesian (written), with the ability to produce nuanced text.
  • 2Prior experience in AI red teaming, adversarial ML, or cybersecurity — you've probed systems and documented vulnerabilities.
  • 3A structured mindset: you use frameworks or benchmarks rather than random hacks, and you can explain risks clearly to technical and non-technical audiences.

Who Should Apply

We're looking for curious, adversarial thinkers who instinctively push AI systems to their limits. You should be comfortable with sensitive content (bias, misinformation, harmful outputs) and have a disciplined approach to testing — you don't just break things, you catalog exactly how and why. A background in psychology, creative writing, or socio-technical risk analysis is a plus, but the core trait is an irresistible urge to find the edge case.

Salary Insight

The role pays $17.00–$25.00 per hour on a remote, contract basis.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

adversarial machine learningred teamingprompt injectionjailbreak techniquesrlhf attackdpo attackmodel extractionpenetration testingexploit developmentreverse engineeringdata annotationrisk classificationconversational aisocio-technical analysisbenchmarkingtaxonomies

Application Tip

In your application, include a short write-up of one specific adversarial test you've run — describe the input, the model's response, and what vulnerability it revealed. This will immediately show you understand the structured, reproducible approach we need.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Thai

Neon is assembling a remote team of bilingual (English & Thai) AI safety specialists to stress-test conversational AI systems. You'll probe models for vulnerabilities like bias, misinformation, and prompt injection, then document and classify findings to help make AI safer. This text-based red teaming role offers flexible hourly work and the chance to shape AI safety at the frontier.

24–35/hr
· 5 openings
Red TeamingAdversarial MLPrompt Injection+10 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Malay

This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

17–25/hr
· 20 openings
Red TeamingAdversarial Machine LearningPrompt Injection+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Portuguese (global)

Neon is assembling a distributed red team to stress-test conversational AI models by probing for vulnerabilities like jailbreaks, prompt injections, and bias exploits. This remote, hourly role invites bilingual experts (fluent in English and global Portuguese, excluding Brazilian variants) to generate high-quality adversarial data that strengthens AI safety. Work is text-based, with optional exposure to sensitive topics supported by clear guidelines and wellness resources.

29–45/hr
· 5 openings
Red TeamingAdversarial Machine LearningCybersecurity+15 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Danish

We're recruiting bilingual AI safety specialists (English & Danish) to join a red team that systematically probes conversational AI models for vulnerabilities. You'll generate adversarial data, surface hidden risks, and create actionable documentation to strengthen AI systems. This remote, hourly contract pays $48–$62/hr based on experience.

48–62/hr
· 5 openings
Red TeamingAdversarial MLPrompt Injection+12 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Dutch

We're looking for bilingual English-Dutch AI red teamers to stress-test conversational AI models from every angle. In this remote, hourly role, you'll design and execute adversarial attacks—from jailbreaks to bias exploitation—and generate structured data that helps make AI systems safer. You'll work with clear guidelines and optional high-sensitivity projects, with wellness support available.

48–62/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+12 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Vietnamese

We're hiring bilingual (English & Vietnamese) red teamers to stress-test AI models by probing for vulnerabilities such as jailbreaks, bias, and misinformation. In this remote, contract role, you'll generate adversarial data and document findings to help make AI systems safer. It's a fit for anyone with a knack for pushing systems to their limits and a structured approach to testing.

17–25/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+13 more