Role Overview
As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them, designing creative, adversarial prompts that expose vulnerabilities such as unsafe content, bias, broken guardrails, hallucinations, and prompt injection weaknesses.
About Handshake ai
Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.
Key Responsibilities
- Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories
- Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques
- Explore edge cases to provoke disallowed, harmful, or incorrect outputs
- Evaluate and score model responses against structured harm taxonomies and severity rubrics
- Document experiments clearly, including what you tried, why you tried it, and what it revealed
Requirements & Eligibility
- Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.)
- Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus
- Creative, adversarial problem-solving skills
- Clear and thoughtful written communication
- Strong ethical judgment and the ability to separate adversarial thinking from personal values
- Self-directed, collaborative, and comfortable in feedback-heavy environments
Required Skills & Tech
LLMChatGPTClaudeGeminiPrompt Engineering