AI Red Teamer, Cybersecurity (Remote)
Hiring in only
This employer appears to hire only in the region above. Confirm you're eligible to be hired there before applying.
Evaluate AI models for cybersecurity vulnerabilities through adversarial prompting
As a Cybersecurity Red Teamer, you will assess whether AI models can be manipulated into generating functional malware, exploit code, or attack guidance. Your work involves crafting adversarial prompts and multi-turn interaction chains to simulate real threat actor behavior. This role focuses on identifying gaps between model safety guardrails and what skilled adversaries can extract, supporting frontier AI labs in improving model robustness.
Why This Role?
Partner closely with world-class AI labs and help build a rapidly growing AI data business with billions in revenue
Key Responsibilities
- Evaluate AI models for potential to generate functional malware or exploit code
- Craft adversarial prompts to test model safety guardrails
- Develop multi-turn interaction chains simulating threat actor behavior
- Identify gaps between intended model protections and actual adversarial extraction
- Support frontier AI labs in improving AI model cybersecurity resilience
Requirements
- Ability to think like an attacker with access to highly capable AI assistants
- Experience crafting adversarial prompts for language models
- Understanding of cyberattack techniques and threat actor methodologies
- Familiarity with LLMs and their safety mechanisms
- Skill in simulating interactions from novice to advanced persistent threat actors
Required Skills
Indonesia Context
- Working Hours Overlap:
- Flexible — work your own hours
View Original Description from Ashby Job Boards
Original description from Ashby Job Boards
CYBERSECURITY RED TEAMER Location: Seattle, WA, or Remote within the United States Work arrangement: Flexible, including part-time availability Handshake was founded on a simple belief: everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we support 25 million job seekers, more than 1 million employers, and 1,600 educational institutions. In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with researchers at frontier AI labs to create evaluations, publish benchmarks, and push the boundaries of data. We have grown from $0 to approximately $1 billion in run rate and pay approximately $60 million to more than 30,000 individuals every month. WHY JOIN HANDSHAKE NOW - Shape how careers evolve in the AI economy at a global scale, with an impact your friends, family, and peers can see and feel - Partner closely with world-class AI labs, Fortune 500 companies, and leading educational institutions - Work with engineers, scientists, operators, and other professionals from organizations such as Palantir, Meta, and Scale AI, as well as former YC founders - Help build a massive, rapidly growing business with billions in revenue - Work from Seattle or remotely from anywhere within the United States ABOUT HANDSHAKE AI Human data is core infrastructure for AI advancement. Frontier AI labs currently improve model capabilities through data-intensive post-training techniques. We believe spending on AI training data will increase three to five times over the next few years and continue growing as models expand into new domains. Handshake AI supports frontier AI labs by working on their most complex data challenges at scale. ABOUT THE ROLE As a Cybersecurity Red Teamer, you will evaluate whether AI models can be manipulated into generating functional malware, viable exploit code, attack tooling, or step-by-step operational guidance that could give a threat actor meaningful assistance in carrying out cyberattacks. Your job is to find the gaps between what a model’s safety guardrails are intended to block and what a skilled adversary can actually extract. This role requires you to think like an attacker who has access to a highly capable AI assistant. You will craft adversarial prompts and multi-turn interaction chains that simulate how real threat actors, ranging from inexperienced attackers to advanced persistent threat operators, might use LLMs to accelerate reconnaissance, weaponization, exploitation, lateral movement, persistence, and exfiltration. You will then evaluate whether the model’s output is genuinely dangerous or merely surface-level noise. Deep cybersecurity expertise is essential. Your value will come from being able to examine a model-generated payload, exploit chain, or attack plan and determine whether it would actually work, how much refinement it would require, and what type of attacker it could meaningfully assist. This position may be performed from our Seattle location or remotely within the United States. Seattle-based and remote team members will collaborate closely through shared evaluation workflows, regular feedback, and virtual working sessions. WHAT YOU WILL DO - Design technically grounded adversarial prompts that test whether models provide meaningful assistance across the cyber kill chain, from reconnaissance through exfiltration and impact - Evaluate model-generated code and technical output for functional correctness, determining whether outputs represent real exploits, plausible attack tooling, or nonfunctional noise - Test model behavior across offensive categories, including malware generation, vulnerability exploitation, social engineering, credential harvesting, privilege escalation, command-and-control infrastructure, and data exfiltration - Probe dual-use boundaries by testing how models respond to queries that combine legitimate security research, penetration testing, and defensive operations with offensive applications - Simulate attacker personas at varying skill levels, including opportunistic, intermediate, and advanced or APT-level actors - Test multi-step and multi-turn attack chains, including scenarios in which early turns establish benign context before pivoting to malicious requests - Score model responses using structured harm taxonomies and severity rubrics calibrated to real-world exploitability - Document findings with clear technical reasoning, including what a response gets right, what it gets wrong, and what level of attacker it could realistically assist - Contribute to the development and refinement of cybersecurity-specific evaluation frameworks and threat models - Collaborate with red teamers, AI researchers, and policy teams to translate findings into actionable model improvements - Stay current on evolving tactics, techniques, and procedures, CVEs, jailbreak techniques, and the intersection of AI and offensive security CORE QUALIFICATIONS - Professional experience in offensive security, penetration testing, red teaming, vulnerability research, malware analysis, threat intelligence, or incident response - Ability to read, write, and evaluate code in languages commonly used for offensive tooling, such as Python, PowerShell, Bash, C/C++, or JavaScript - Understanding of common attack frameworks, techniques, and procedures, including MITRE ATT&CK and OWASP - Ability to assess the functional correctness and real-world exploitability of model-generated technical output - Strong hands-on experience using multiple LLMs, such as ChatGPT, Claude, Gemini, or open-source models - Creative and adversarial problem-solving skills - Clear and precise written communication, including the ability to explain technical risk to nonspecialist audiences - Strong ethical judgment and the ability to separate adversarial thinking from personal values - Ability to work independently while collaborating effectively in a feedback-heavy, distributed environment NICE TO HAVE - Relevant certifications, such as OSCP, OSCE, GPEN, GXPN, CRTO, CRTL, CEH, or similar - Active or previous security clearance - Experience with exploit development, reverse engineering, or binary analysis - Background in cloud security, container security, or infrastructure-as-code attack surfaces - Familiarity with AI and machine-learning attack surfaces, including prompt injection, model extraction, training-data poisoning, and adversarial examples - Experience building or operating command-and-control frameworks, custom implants, or offensive tooling - A bug-bounty track record or published CVEs - Previous work in trust and safety, content moderation, or AI evaluation - Familiarity with LLM APIs or evaluation tooling YOU MAY BE A STRONG FIT IF - You have spent years breaking into systems and want to apply that mindset to testing AI models - You can examine a model-generated reverse shell, phishing template, or privilege-escalation script and quickly determine whether it would work in a real environment - You think in kill chains and attack graphs, not just individual prompts - You understand that the difference between a useful coding assistant and a dangerous one often comes down to context, specificity, and operational detail - You closely follow the offensive-security community and stay current when new techniques emerge - You care about AI safety because you understand what can happen when powerful tools are used irresponsibly - You can collaborate effectively with a team whether you are working from Seattle or remotely CONTENT NOTICE This role involves regular and deliberate engagement with offensive cybersecurity content. You will create and evaluate scenarios involving malware, exploit code, social engineering, network-intrusion techniques, and other attack methodologies. All work is conducted within a structured evaluation framework with strict ethical guidelines. Candidates must be able to engage with this material professionally, responsibly, and sustainably.
Salary Context
Similar Engineering roles on LokerDollar pay around $170k/yr (range $855–1000k/yr, n=868 active listings).
Hiring at Handshake
Handshake has 3 other active roles on LokerDollar and has been hiring here since May 31, 2026 — across Engineering, Design, Marketing.
View all Handshake openings →Frequently asked questions
- Is AI Red Teamer, Cybersecurity (Remote) at Handshake a remote job?
- Yes. AI Red Teamer, Cybersecurity (Remote) at Handshake is a fully remote role open to candidates worldwide.
- What is the salary for AI Red Teamer, Cybersecurity (Remote) at Handshake?
- The listed pay range for this role is $65–125/hr.
- What type of employment is AI Red Teamer, Cybersecurity (Remote) at Handshake?
- This is a contract position.
- How do I apply?
- Click the "Apply" button on this page to go to the official application at Handshake.
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Remote ≠ Remote: The Skills That Open Global Work to Indonesians (2026)12,891 remote listings: the highest-paid coding skills are the most geo-locked for Indonesia-based applicants. CC BY 4.0 aggregate dataset.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesian Remote Work Salary & Demand IndexHow much of the global remote job corpus is open to Indonesia, and what it pays (USD) by role.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
From the blog
- Remote Work Realities: What You Earn in USDExplore remote job opportunities with USD salaries and understand the realities of pay, roles, and companies in 2026.
- Are You AI-Ready? Remote Jobs and Global PayIs the AI era passing you by? Explore the latest global remote job openings and learn how to benchmark your salary against international standards.
- Remote Jobs in Sept 2026: AI’s Impact & RealAI is shifting the remote work landscape. Which skills are actually in demand this September? Analyzing global trends and job market data.