AI Jailbreak & Prompt-Injection Security Expert | $50-$90/hr Remote
Overview
micro1 is looking for a contractor to join a forward-thinking client project focused on hardening AI systems against exploitation. Your mission: design creative adversarial tests—think ethical jailbreaks, prompt injection, and tool-use abuse—to uncover weaknesses in modern LLMs. No deep AI background is required; your hands-on security expertise and adversarial mindset are what count.
What You'll Do6
- 1Build and execute advanced testing strategies for AI safety, including multi-turn jailbreak attempts, prompt injection attacks, and abuse of model tool-calling capabilities.
- 2Develop cross-domain elicitation methods that probe complex, chained adversarial behaviors across different contexts.
- 3Create and maintain regression test suites that systematically check models for known and novel jailbreak or injection vulnerabilities.
- 4Construct evaluation frameworks that put LLMs under realistic stress scenarios to measure and improve their robustness.
- 5Work with technical stakeholders to turn your findings into concrete safety improvements and mitigations.
- 6Document your processes, results, and recommended practices in clear reports and presentations for both engineering and non-technical audiences.
Requirements7
- 1At least 2 years of experience in adversarial machine learning, LLM red teaming, AI safety assessment, or a closely related security specialty.
- 2A proven track record of discovering or testing vulnerabilities linked to ethical jailbreaks, prompt injection, tool-use abuse, or other adversarial AI attacks.
- 3An advanced degree (PhD or MS) in computer science, cybersecurity, machine learning, or a similar field—or equivalent professional credentials.
- 4Credibility within the AI security community, evidenced by published research, open-source tools, talks at conferences, or recognized participation in bug bounty programs.
- 5Strong written and verbal communication skills, with an emphasis on precise documentation and collaborative problem-solving.
- 6Experience working on multi-disciplinary or cross-functional safety initiatives is a plus.
- 7Familiarity with current LLM architectures, prompt engineering techniques, and security testing tools is highly desirable.
Who Should Apply
This role is ideal for a security researcher or red-teamer who thrives on breaking things to make them safer. You have a hacker’s curiosity and a methodical approach to finding edge cases in AI behavior. You enjoy documenting your exploits clearly and collaborating with engineers to turn insights into defenses.
Salary Insight
$50–$90 per hour, based on experience and expertise.
Required Skills
Application Tip
Include concrete examples of jailbreaks or prompt injections you’ve successfully executed—describe the technique, the model, and what you uncovered. A short write-up or a link to a published proof of concept will set you apart.
Similar open positions
Explore active roles that match your skills and interests.
Mercor
VerifiedAI Safety Experts — English & Assamese | $20-$22/hr Remote
We're assembling a remote team of AI red teamers to stress-test conversational models and uncover hidden vulnerabilities. In this role, you'll probe AI systems with adversarial inputs—jailbreaks, prompt injections, bias exploitation—and generate the high-quality human data that makes AI safer for real-world use. Fluency in both English and Assamese is essential, as you'll review sensitive outputs and document reproducible attack cases. This project is ideal for structured, curious thinkers who enjoy pushing systems to their limits.
Micro1
VerifiedLLM Red-Teamer | $40-$65/hr Remote
micro1 is looking for sharp critical thinkers to help push the limits of frontier language models. In this role, you'll design tricky, adversarial multi-turn conversations and evaluate how well AI systems handle them. Your unique domain knowledge matters more than prior AI experience — we want people who can think like a hacker and write with precision.
Mercor
VerifiedAI Safety Experts — English & Odia | $20-$22/hr Remote
Mercor is looking for native English and Odia speakers to join a remote red team that stress-tests AI conversational models. You'll probe for vulnerabilities using adversarial techniques like jailbreaks and prompt injections, generating critical data that makes AI systems safer. This text-based project covers sensitive topics with clear guidelines and wellness support, allowing you to work on high-impact safety work from anywhere.
Micro1
VerifiedRed Team Lead (Offensive Cybersecurity) | $50-$90/hr Remote
This remote contractor role with micro1 puts your offensive cybersecurity expertise to work in a unique way: you'll help train next-generation AI systems by designing evaluation frameworks and attack simulations. No prior AI experience is required—your hands-on knowledge of exploit chains, cloud/appsec, and social engineering is what makes you the ideal candidate. You'll shape how models learn about real-world threats while operating under strict ethical boundaries.
Micro1
VerifiedPenetration Tester | $25-$80/hr Remote
micro1 is looking for an experienced penetration tester to contribute your security expertise to the development of next-generation AI systems. You won't need an AI background — your hands-on knowledge of ethical hacking and vulnerability assessment is exactly what's needed to help train smarter models. This is a remote contract opportunity paying $25–$80 per hour.
Mercor
VerifiedAI Safety Experts — English & Bengali | $20-$22/hr Remote
Neon is seeking bilingual (English & Bengali) AI safety experts to join a remote red teaming project. In this role, you'll proactively probe conversational AI models using adversarial techniques like jailbreaks and prompt injections to uncover vulnerabilities before they reach production. Your work directly contributes to making AI systems more robust and trustworthy.