AI Red Teamer (LLM Generalist) - Remote
Hiring in only
This employer appears to hire only in the region above. Confirm you're eligible to be hired there before applying.
Craft adversarial prompts to expose LLM vulnerabilities across risk categories
As an AI Red Teamer, you will stress-test large language models by designing creative, adversarial prompts that expose weaknesses such as unsafe content, bias, hallucinations, and prompt injection flaws. Your work involves probing models across content safety, CBRN, cybersecurity, persuasion, child safety, self-harm, over-companionship, and regulatory compliance. You will document findings, evaluate responses against harm taxonomies, and colla...
Why This Role?
Direct contribution to AI safety and model robustness for leading research labs
Key Responsibilities
- Craft creative prompts and multi-turn scenarios to stress-test AI guardrails
- Discover ways around safety filters using jailbreak and prompt injection techniques
- Explore edge cases to provoke disallowed, harmful, or incorrect model outputs
- Evaluate and score model responses against structured harm taxonomies and severity rubrics
- Document experiments clearly, including what was tried and what it revealed
- Review and refine adversarial prompts generated by other team members
Requirements
- Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models)
- Intuition for crafting adversarial prompts
- Familiarity with jailbreak or evasion techniques is a strong plus
- Creative, adversarial problem-solving skills
- Clear and thoughtful written communication
- Strong ethical judgment and ability to separate adversarial thinking from personal values
Required Skills
Indonesia Context
- Working Hours Overlap:
- Flexible — work your own hours
View Original Description from Ashby Job BoardsShow more
Original description from Ashby Job Boards
AI RED TEAMER (LLM GENERALIST) Location: Remote (USA) Type: Contract, 40 hours per week ABOUT THE ROLE As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking whether an answer is correct, you will design creative, adversarial prompts that expose vulnerabilities: unsafe content, bias, broken guardrails, hallucinations, prompt injection weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a generalist red teaming role. You will probe models across the full spectrum of risk categories, including content safety, CBRN (chemical, biological, radiological, nuclear), cybersecurity, persuasion and influence operations, child safety, self-harm, over-companionship, and regulatory compliance. Red teaming may span text, image, voice, and agentic model capabilities depending on project needs. This role requires creativity, curiosity, and an ability to think like an adversary while operating with strong ethical judgment. DAY-TO-DAY RESPONSIBILITIES - Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories - Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques - Explore edge cases to provoke disallowed, harmful, or incorrect outputs - Evaluate and score model responses against structured harm taxonomies and severity rubrics - Document experiments clearly, including what you tried, why you tried it, and what it revealed - Review and refine adversarial prompts generated by other team members - Contribute to harm taxonomy development, calibration exercises, and inter-rater reliability work - Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses - Work with potentially disturbing content on a regular basis (see Content Warning below) - Stay current on jailbreaks, attack methods, and evolving model behaviors DESIRED CAPABILITIES Core - Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.) - Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus - Creative, adversarial problem-solving skills - Clear and thoughtful written communication - Strong ethical judgment and the ability to separate adversarial thinking from personal values - Self-directed, collaborative, and comfortable in feedback-heavy environments - Curiosity, persistence, and comfort with frequent failure in experimentation Nice to Have - Familiarity with Python or other scripting languages - Experience working with LLM APIs or evaluation tooling - Comfort with structured data annotation and rubric-based scoring - Prior work in trust and safety, content moderation, QA, or security research - Subject matter expertise in any high-risk domain (cybersecurity, chemistry, biology, medicine, law, finance, etc.) YOU WILL THRIVE HERE IF - You treat every model response as a hypothesis to challenge - You can switch between creative free-association and rigorous documentation in the same session - You go deep into unusual interests (fandoms, niche internet cultures, gaming exploits, Wikipedia rabbit holes, etc.) - You come from a creative background: writing, visual art, improv, puzzle design, or similar - You are energized by finding the thing nobody else thought to try - You are genuinely passionate about AI and follow the space closely CONTENT WARNING This role involves regular and deliberate exposure to harmful content. You will encounter and intentionally generate content involving violence, self-harm, hate speech, sexually explicit material, child safety scenarios, and other categories of harmful output as part of structured adversarial testing. Candidates must be able to engage with this material professionally and sustainably. Support resources are available. ABOUT HANDSHAKE AI Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.
Salary Context
Similar AI roles on LokerDollar pay around $140k/yr (range $12k–242.4k/yr, n=81 active listings).
Hiring at Handshake
Handshake has 4 other active roles on LokerDollar and has been hiring here since May 31, 2026 — across AI, Engineering, Design.
View all Handshake openings →Frequently asked questions
- Is AI Red Teamer (LLM Generalist) - Remote at Handshake a remote job?
- Yes, AI Red Teamer (LLM Generalist) - Remote at Handshake is remote, but the employer hires only in certain regions. Check the listing for location requirements before applying.
- What is the salary for AI Red Teamer (LLM Generalist) - Remote at Handshake?
- The listed pay range for this role is $32–95/hr.
- What type of employment is AI Red Teamer (LLM Generalist) - Remote at Handshake?
- This is a full time position.
- How do I apply?
- Click the "Apply" button on this page to go to the official application at Handshake.
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Remote ≠ Remote: The Skills That Open Global Work to Indonesians (2026)12,891 remote listings: the highest-paid coding skills are the most geo-locked for Indonesia-based applicants. CC BY 4.0 aggregate dataset.
- The Compliance Layer of the AI Hiring Stack (2026)6,349 remote listings: 77.2% never state who may apply. Methodology and a CC BY 4.0 aggregate dataset.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesian Remote Work Salary & Demand IndexHow much of the global remote job corpus is open to Indonesia, and what it pays (USD) by role.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
From the blog
- The Truth Behind Viral Airbnb Remote JobsViral claims of $3,500/week remote jobs at Airbnb are circulating. We cut through the hype to see what's real and highlight 7 active remote roles.
- The AI Skill Premium: What Does it MeanSeptember 2026 remote job data reveals a surge in AI skill demand. But does this mean higher salaries for remote workers worldwide?
- The Truth About Clipboard OnboardingA candid look at the Clipboard Onboarding Documents Associate role, salary estimates, and how to navigate remote job postings that lack transparency.