Micro1Verified Source
RemoteHotAI & Machine LearningCybersecurity

AI Jailbreak & Prompt-Injection Security Expert | $50-$90/hr Remote

$50–$90/hr
Posted June 30, 2026
contract

Overview

micro1 is looking for a contractor to join a forward-thinking client project focused on hardening AI systems against exploitation. Your mission: design creative adversarial tests—think ethical jailbreaks, prompt injection, and tool-use abuse—to uncover weaknesses in modern LLMs. No deep AI background is required; your hands-on security expertise and adversarial mindset are what count.

What You'll Do6

  • 1Build and execute advanced testing strategies for AI safety, including multi-turn jailbreak attempts, prompt injection attacks, and abuse of model tool-calling capabilities.
  • 2Develop cross-domain elicitation methods that probe complex, chained adversarial behaviors across different contexts.
  • 3Create and maintain regression test suites that systematically check models for known and novel jailbreak or injection vulnerabilities.
  • 4Construct evaluation frameworks that put LLMs under realistic stress scenarios to measure and improve their robustness.
  • 5Work with technical stakeholders to turn your findings into concrete safety improvements and mitigations.
  • 6Document your processes, results, and recommended practices in clear reports and presentations for both engineering and non-technical audiences.

Requirements7

  • 1At least 2 years of experience in adversarial machine learning, LLM red teaming, AI safety assessment, or a closely related security specialty.
  • 2A proven track record of discovering or testing vulnerabilities linked to ethical jailbreaks, prompt injection, tool-use abuse, or other adversarial AI attacks.
  • 3An advanced degree (PhD or MS) in computer science, cybersecurity, machine learning, or a similar field—or equivalent professional credentials.
  • 4Credibility within the AI security community, evidenced by published research, open-source tools, talks at conferences, or recognized participation in bug bounty programs.
  • 5Strong written and verbal communication skills, with an emphasis on precise documentation and collaborative problem-solving.
  • 6Experience working on multi-disciplinary or cross-functional safety initiatives is a plus.
  • 7Familiarity with current LLM architectures, prompt engineering techniques, and security testing tools is highly desirable.

Who Should Apply

This role is ideal for a security researcher or red-teamer who thrives on breaking things to make them safer. You have a hacker’s curiosity and a methodical approach to finding edge cases in AI behavior. You enjoy documenting your exploits clearly and collaborating with engineers to turn insights into defenses.

Salary Insight

$50–$90 per hour, based on experience and expertise.

Required Skills

Ethical JailbreaksLLM Red TeamingPrompt InjectionTool-Use Abuse

Application Tip

Include concrete examples of jailbreaks or prompt injections you’ve successfully executed—describe the technique, the model, and what you uncovered. A short write-up or a link to a published proof of concept will set you apart.

Share:

Similar open positions

Explore active roles that match your skills and interests.

Mercor

21d agoRemotehourly

AI Safety Experts — English & Assamese | $20-$22/hr Remote

We're assembling a remote team of AI red teamers to stress-test conversational models and uncover hidden vulnerabilities. In this role, you'll probe AI systems with adversarial inputs—jailbreaks, prompt injections, bias exploitation—and generate the high-quality human data that makes AI safer for real-world use. Fluency in both English and Assamese is essential, as you'll review sensitive outputs and document reproducible attack cases. This project is ideal for structured, curious thinkers who enjoy pushing systems to their limits.

$20–$22/hr
· 5 openings

Micro1

4d agoRemote
Hot

LLM Red-Teamer | $40-$65/hr Remote

micro1 is looking for sharp critical thinkers to help push the limits of frontier language models. In this role, you'll design tricky, adversarial multi-turn conversations and evaluate how well AI systems handle them. Your unique domain knowledge matters more than prior AI experience — we want people who can think like a hacker and write with precision.

$40–$65/hr
· 100 openings
Adversarial prompt constructionPrecision in written EnglishRubric design+1 more

Mercor

21d agoRemotehourly

AI Safety Experts — English & Odia | $20-$22/hr Remote

Mercor is looking for native English and Odia speakers to join a remote red team that stress-tests AI conversational models. You'll probe for vulnerabilities using adversarial techniques like jailbreaks and prompt injections, generating critical data that makes AI systems safer. This text-based project covers sensitive topics with clear guidelines and wellness support, allowing you to work on high-impact safety work from anywhere.

$20–$22/hr
· 5 openings

Micro1

19d agoRemote
Hot

Red Team Lead (Offensive Cybersecurity) | $50-$90/hr Remote

This remote contractor role with micro1 puts your offensive cybersecurity expertise to work in a unique way: you'll help train next-generation AI systems by designing evaluation frameworks and attack simulations. No prior AI experience is required—your hands-on knowledge of exploit chains, cloud/appsec, and social engineering is what makes you the ideal candidate. You'll shape how models learn about real-world threats while operating under strict ethical boundaries.

$50–$90/hr
Cyber SecurityExploit ChainsCloud/Appsec+1 more

Micro1

1mo agoRemotefull-time

Penetration Tester | $25-$80/hr Remote

micro1 is looking for an experienced penetration tester to contribute your security expertise to the development of next-generation AI systems. You won't need an AI background — your hands-on knowledge of ethical hacking and vulnerability assessment is exactly what's needed to help train smarter models. This is a remote contract opportunity paying $25–$80 per hour.

$25–$80/hr
· 7 openings
Ethical HackingVulnerability AssessmentPenetration Testing

Mercor

21d agoRemotehourly

AI Safety Experts — English & Bengali | $20-$22/hr Remote

Neon is seeking bilingual (English & Bengali) AI safety experts to join a remote red teaming project. In this role, you'll proactively probe conversational AI models using adversarial techniques like jailbreaks and prompt injections to uncover vulnerabilities before they reach production. Your work directly contributes to making AI systems more robust and trustworthy.

$20–$22/hr
· 5 openings