Skip to main content
White Circle

AI Red Team Engineer

RemoteUnited States only
Published
Role
Security
Employment
Full-time
Company size
Startup
Salary not disclosed
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

AI Red Team Engineer to break LLM systems, automate attacks, and document findings. Requires Python scripting, API/web app testing, and understanding of LLM abuse vectors. Must be able to reason adversarially and write clear reports in English.

Core skills

PythonLLM Red TeamingAdversarial Testing

Required skills

LLMspromptssystem instructionsRAGagentstool/function callingAPIsweb appsbackendsSaaS products

Optional skills

Burp SuitePostmanPlaywrightpytestmodern LLM red-teaming automated agents and pipelinesLangChainLangGraphLlamaIndex

Required languages

English

What you'll do

  • Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
  • Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse.
  • Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
  • Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
  • Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
  • Convert successful attacks into regression tests and product requirements.
  • Track new red-team and safety techniques and fold the useful ones into our tests.
  • Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.

What they require

  • Genuinely love breaking things and reasoning adversarially.
  • Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty.
  • Have strong Python scripting skills.
  • Have experience testing APIs, web apps, backends, or SaaS products.
  • Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
  • Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion).
  • Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact.
  • Can separate real customer risk from low-impact prompt tricks.
  • Write clear, reproducible bug reports in clear English.
  • Can move fast without perfect requirements.
  • Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.

Benefits

  • Paid time off in line with your local regulations, no matter where you work from
  • Meaningful equity package
  • All the hardware, tools, and services you need
  • Covered subscriptions for AI agents
  • Team off-sites twice a year: we've recently been to the Alps and to Saint-Tropez

White Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of its platform are policies — natural-language rules defining what an AI model should and shouldn’t do — which it automatically tests, enforces, and improves at scale.

🇺🇸 United StatesAI SafetyStartup
Salary not disclosed