This one is closed
Live roles like this one
-
B
1w ago
AI Safety Expert - Red Teaming - AI Trainer
mercor Denmark $48 - $62/hr
-
B
1w ago
Bright Vision Technologies United States $100k - $150k/yr
-
B
1w ago
mercor India $20 - $22/hr
- D 2w ago
See every "AI Jailbreak Prompt" role →
Get new “AI Jailbreak Prompt” roles by email
One email a day with what is new in "AI Jailbreak Prompt". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Why this grade This listing scored 34/100, which is an F. It lost the most ground on pay transparency. See the breakdown
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Remote clarity 15 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Corroboration 5 / 10 Whether more than one source carries this listing.
- Freshness 4 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
- Role specificity 0 / 10 Whether the listing is tagged well enough to tell what the role actually is.
- Pay transparency 0 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
-10 Ghost-job penalty — Deducted for signals that this posting may not be a real, currently-open role — staleness, repeated relisting, or talent-pool language.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
Role Title: AI Jailbreak & Prompt-Injection Security Expert
Role Type: Contractor
Location: Remote
micro1 is engaging AI Jailbreak & Prompt-Injection Security Experts to contribute to a cutting-edge customer initiative focused on AI safety and robustness. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required — your domain knowledge is what matters.
Scope of Work
- Design and implement advanced methodologies for evaluating AI system safety, focusing on ethical jailbreaks, LLM red teaming, prompt injection, and tool-use abuse scenarios.
- Create comprehensive cross-domain elicitation strategies to uncover multi-turn and complex adversarial bypass patterns in AI models.
- Develop, maintain, and update regression test suites that systematically test for jailbreak susceptibility and prompt-injection vulnerabilities.
- Construct robust evaluation frameworks that stress-test AI models against real-world adversarial threats, aiming to enhance overall system robustness.
- Collaborate with technical stakeholders to translate security findings into actionable improvements for model safety and risk mitigation.
- Document methodologies, findings, and best practices in clear, well-structured written reports and presentations for both technical and non-technical audiences.
Preferred Qualifications
- 2+ years of expertise in adversarial machine learning, LLM red teaming, AI safety evaluation, or a closely related security domain
- Proven experience researching, testing, or uncovering vulnerabilities related to ethical jailbreaks, prompt injection, tool-use abuse, or adversarial AI attacks.
- Advanced degree (PhD, MS) in computer science, cybersecurity, machine learning, or a relevant discipline, or equivalent operational/professional background.
- High credibility and recognition within the AI security or adversarial ML community—such as published research, open-source tools, or conference presentations.
- Exceptional written and verbal communication skills, with a strong focus on clear documentation and collaborative problem-solving.
- Prior participation in multi-disciplinary projects or cross-functional AI safety initiatives is a plus.
- Familiarity with current LLM architectures, prompt engineering techniques, and security assessment tools is highly desirable.
Originally posted on Himalayas
Apply for this role Opens himalayas.app — the link as listed; we have not yet verified it is the employer's own page
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 04 Aug 2026 Himalayas first sighting
Seen on 1 board over 0 days.