AI Red Teamer (LLM Generalist) - Remote

Added
12 minutes ago
Type
Contract
Salary
Upgrade to Premium to se...

Related skills

chatgpt llms gemini ai safety claude

πŸ“‹ Description

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk
  • Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and
  • Explore edge cases to provoke disallowed, harmful, or incorrect outputs
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics
  • Document experiments clearly, including what you tried, why you tried it, and what it revealed
  • Review and refine adversarial prompts generated by other team members

🎯 Requirements

  • Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.)
  • Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a
  • Creative, adversarial problem-solving skills
  • Clear and thoughtful written communication
  • Strong ethical judgment and the ability to separate adversarial thinking from personal values
  • Self-directed, collaborative, and comfortable in feedback-heavy environments

🎁 Benefits

  • Contract, 40 hours per week
  • Remote (USA) work
  • Support resources available for handling disturbing content
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Data Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Data Jobs

See more Data jobs β†’