
AI Safety Experts — English & Portuguese (global) | $29-$45/hr Remote
Overview
Neon is assembling a distributed red team to stress-test conversational AI models by probing for vulnerabilities like jailbreaks, prompt injections, and bias exploits. This remote, hourly role invites bilingual experts (fluent in English and global Portuguese, excluding Brazilian variants) to generate high-quality adversarial data that strengthens AI safety. Work is text-based, with optional exposure to sensitive topics supported by clear guidelines and wellness resources.
What You'll Do5
- 1Stress-test conversational AI agents using adversarial techniques such as jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation.
- 2Produce high-quality human data by annotating model failures, classifying vulnerability types, and flagging systemic risks.
- 3Apply structured taxonomies, benchmarks, and playbooks to keep red-teaming efforts consistent and reproducible.
- 4Document findings in reports and datasets that customers can directly use to harden their AI systems.
- 5Surface vulnerabilities that automated tests overlook, delivering reproducible artifacts for safer deployments.
Requirements5
- 1Prior experience in AI red teaming, adversarial probing, or cybersecurity penetration testing.
- 2Native-level fluency in both English and Portuguese (global, excluding Brazilian Portuguese) — all work is text-based.
- 3Strong analytical mindset with the ability to apply frameworks or benchmarks (not just random hacking) to evaluate model safety.
- 4Excellent communication skills to explain risks clearly to both technical and non-technical stakeholders.
- 5Adaptability to move across different projects and customer environments, plus comfort with potentially sensitive topics (bias, misinformation, harmful behaviors).
Who Should Apply
This role is for curious, adversarial thinkers who instinctively want to break systems to make them safer. Ideal candidates bring hands-on experience in red teaming, adversarial ML, or cybersecurity, and are comfortable working with structured taxonomies and reproducible documentation. You should thrive on varied challenges, communicate clearly, and be eager to work at the frontier of AI safety — all while being fully fluent in English and global Portuguese.
Salary Insight
The role pays $29.00 – $45.00 per hour, based on experience and expertise.
Location
Required Skills
Application Tip
When applying, include a brief example of a time you uncovered a subtle vulnerability in an AI system or software — even if it was informal. Demonstrating your adversarial mindset with a concrete story will set you apart.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedAI Safety Experts — English & Indonesian
Neon is looking for bilingual AI safety experts fluent in English and Indonesian to join a human-driven red team. You'll probe advanced conversational AI models for vulnerabilities like bias, misinformation, and harmful behavior — all through text-based work. This remote, hourly role pays $17–$25/hr and gives you a direct hand in making AI systems safer before they reach the public.

Mercor
VerifiedAI Safety Experts — English & Odia
We're assembling a red team of human experts to stress-test conversational AI models for safety vulnerabilities. This remote, text-based position requires native fluency in English and Odia. You'll attack AI systems with adversarial inputs, uncover weaknesses, and generate the high-quality data needed to make AI safer. The work covers sensitive topics like bias and misinformation, with clear guidelines and wellness support available.

Mercor
VerifiedAI Safety Experts — English & Norwegian
This remote role invites fluent English and Norwegian speakers to join an elite red team probing AI models for vulnerabilities. You'll simulate adversarial attacks, document exploits, and generate high-quality human data that directly strengthens AI safety. The work is text-based and focuses on sensitive topics like bias and misinformation, with optional high-sensitivity projects supported by clear guidelines.

Mercor
VerifiedAI Safety Experts — English & Thai
Neon is assembling a remote team of bilingual (English & Thai) AI safety specialists to stress-test conversational AI systems. You'll probe models for vulnerabilities like bias, misinformation, and prompt injection, then document and classify findings to help make AI safer. This text-based red teaming role offers flexible hourly work and the chance to shape AI safety at the frontier.

Mercor
VerifiedAI Safety Experts — English & Bengali
This role brings fluent English and Bengali speakers into a remote AI red team. You'll probe conversational AI systems with adversarial inputs to uncover biases, misinformation, and manipulation risks. All work is text-based, and you'll produce structured red team data that improves AI safety. You can opt out of higher-sensitivity topics, and clear guidelines plus wellness resources are provided.

Mercor
VerifiedAI Safety Experts — English & Malay
This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

