White Circle

White Circle

AI Red Team Engineer

Remote • US

Sponsorship not specifiedDetected 16 days ago
PythonDatadogLLMsRAGAgentic AILangGraphPenetration TestingPlaywrightPostmanTest Automation

About the role

  • If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built - you're the one we need.
  • Final call with CEO (45 min) Please submit your application in English

Responsibilities

  • Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.
  • Experience building eval pipelines, regression suites, dashboards, or CI-friendly security tests.

Requirements

  • Experience with Burp Suite, Postman, Playwright, pytest.
  • Experience with modern LLM red-teaming automated agents and pipelines.
  • Experience with trust & safety, abuse prevention, fraud, moderation, or platform security.

Skills

  • Genuinely love breaking things and reasoning adversarially.
  • Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty.
  • Have strong Python scripting skills.
  • Have experience testing APIs, web apps, backends, or SaaS products.
  • Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
  • Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion).
  • Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact.
  • Can separate real customer risk from low-impact prompt tricks.
  • Write clear, reproducible bug reports in clear English.
  • Can move fast without perfect requirements.

Benefits

  • Paid time off in line with your local regulations, no matter where you work from
  • Best medical insurance in France

Company info

  • White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies - simple natural-language rules that define what an AI model should and shouldn't do. We automatically test, enforce, and continuously improve these policies at scale.
  • We've raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others
  • We process over one hundred million API calls every month
  • We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model
  • White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems.
  • At the core of our platform are policies - simple natural-language rules that define what an AI model should and shouldn't do.
  • We automatically test, enforce, and continuously improve these policies at scale.
  • We're a small, highly focused team.
  • Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
  • Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse.
  • Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
  • Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
  • Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
  • Convert successful attacks into regression tests and product requirements.
  • Track new red-team and safety techniques and fold the useful ones into our tests.
  • You'll fit right in if you:
  • Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.

This listing is sourced directly from White Circle's careers page and normalized into a canonical job model.