Mercor
MercorVerified listing
RemoteAI & Machine LearningCybersecurity

AI Safety Experts — English & Danish | $48-$62/hr Remote

48–62/hr
Remote
Posted August 5, 2026
hourly
5 openings

Overview

We're recruiting bilingual AI safety specialists (English & Danish) to join a red team that systematically probes conversational AI models for vulnerabilities. You'll generate adversarial data, surface hidden risks, and create actionable documentation to strengthen AI systems. This remote, hourly contract pays $48–$62/hr based on experience.

What You'll Do5

  • 1Design and execute adversarial attacks on conversational AI models, including jailbreaks, prompt injections, and multi-turn manipulation scenarios.
  • 2Annotate and classify model failures, flagging systemic vulnerabilities and bias-related risks.
  • 3Follow structured taxonomies, benchmarks, and playbooks to maintain consistent, repeatable testing.
  • 4Produce detailed reports and datasets that document attack methods and findings for customer remediation.
  • 5Collaborate across projects to expand evaluation coverage and uncover edge cases missed by automated tests.

Requirements5

  • 1Native-level fluency in English and Danish — both written and spoken.
  • 2Proven experience in AI red teaming, adversarial testing, or cybersecurity (e.g., penetration testing, exploit development).
  • 3Strong adversarial mindset with the ability to think like an attacker and systematically probe for weaknesses.
  • 4Ability to document findings clearly for both technical and non-technical stakeholders.
  • 5Familiarity with RLHF / DPO attacks, prompt injection, or socio-technical risk probing is a plus.

Who Should Apply

This role is perfect for security researchers, AI safety analysts, or penetration testers who enjoy stress-testing models in creative ways. You should be methodical, curious, and comfortable handling sensitive topics. Bilingual proficiency in English and Danish is non-negotiable.

Salary Insight

Hourly pay ranges from $48.00 to $62.00, depending on experience and project complexity.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

red teamingadversarial mlprompt injectionjailbreakrlhfdpomodel extractionpenetration testingexploit developmentreverse engineeringabuse analysisconversational aisocio-technical riskdanishenglish

Application Tip

In your application, emphasize specific examples of adversarial attacks you've conducted — especially any involving multilingual or Danish-language AI systems — and how you documented reproducible findings.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Norwegian

This remote role invites fluent English and Norwegian speakers to join an elite red team probing AI models for vulnerabilities. You'll simulate adversarial attacks, document exploits, and generate high-quality human data that directly strengthens AI safety. The work is text-based and focuses on sensitive topics like bias and misinformation, with optional high-sensitivity projects supported by clear guidelines.

48–62/hr
· 5 openings
EnglishNorwegianRed Teaming+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Finnish

We're assembling a specialized team to stress-test AI systems by simulating adversarial attacks. As an AI Safety Expert, you'll probe conversational models for vulnerabilities like bias, misinformation, and harmful behaviors — all while working remotely on an hourly basis. Native fluency in both English and Finnish is essential for this role.

48–62/hr
· 5 openings
Adversarial Machine LearningPrompt InjectionJailbreaking+8 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Indonesian

Neon is looking for bilingual AI safety experts fluent in English and Indonesian to join a human-driven red team. You'll probe advanced conversational AI models for vulnerabilities like bias, misinformation, and harmful behavior — all through text-based work. This remote, hourly role pays $17–$25/hr and gives you a direct hand in making AI systems safer before they reach the public.

17–25/hr
· 5 openings
Adversarial Machine LearningRed TeamingPrompt Injection+13 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Dutch

We're looking for bilingual English-Dutch AI red teamers to stress-test conversational AI models from every angle. In this remote, hourly role, you'll design and execute adversarial attacks—from jailbreaks to bias exploitation—and generate structured data that helps make AI systems safer. You'll work with clear guidelines and optional high-sensitivity projects, with wellness support available.

48–62/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+12 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Swedish

This role involves stress-testing AI systems by simulating adversarial attacks in English and Swedish. You'll generate critical safety data by probing models for vulnerabilities like bias, misinformation, and security flaws. As part of a human red team, you'll help ensure AI behaves safely before it reaches users.

48–62/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+12 more
Mercor

Mercor

13h agoRemotepart-time

Bilingual Danish Generalist Expert — AI Safety

Danish-English bilinguals can apply their language expertise to improve AI safety in this remote, part-time role. You'll help evaluate and strengthen how advanced models respond to sensitive or dual-use topics in Danish, using your cultural judgment to craft prompts and classify conversations. No prior AI or machine learning experience is needed; you'll receive training on the workflow before the project begins. The assignment starts immediately, with pay ranging from $48 to $52 per hour.

48–52/hr
· 6 openings
DanishEnglishAI Safety+5 more