White Circle
AI Red Team Engineer
Remote • US
Sponsorship not specifiedDetected 16 days ago
PythonDatadogLLMsRAGAgentic AILangGraphPenetration TestingPlaywrightPostmanTest Automation
About the role
- If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built - you're the one we need.
- Final call with CEO (45 min) Please submit your application in English
Responsibilities
- Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.
- Experience building eval pipelines, regression suites, dashboards, or CI-friendly security tests.
Requirements
- Experience with Burp Suite, Postman, Playwright, pytest.
- Experience with modern LLM red-teaming automated agents and pipelines.
- Experience with trust & safety, abuse prevention, fraud, moderation, or platform security.
Skills
- Genuinely love breaking things and reasoning adversarially.
- Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty.
- Have strong Python scripting skills.
- Have experience testing APIs, web apps, backends, or SaaS products.
- Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
- Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion).
- Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact.
- Can separate real customer risk from low-impact prompt tricks.
- Write clear, reproducible bug reports in clear English.
- Can move fast without perfect requirements.
Benefits
- Paid time off in line with your local regulations, no matter where you work from
- Best medical insurance in France
Company info
- White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies - simple natural-language rules that define what an AI model should and shouldn't do. We automatically test, enforce, and continuously improve these policies at scale.
- We've raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others
- We process over one hundred million API calls every month
- We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model
- White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems.
- At the core of our platform are policies - simple natural-language rules that define what an AI model should and shouldn't do.
- We automatically test, enforce, and continuously improve these policies at scale.
- We're a small, highly focused team.
- Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
- Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse.
- Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
- Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
- Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
- Convert successful attacks into regression tests and product requirements.
- Track new red-team and safety techniques and fold the useful ones into our tests.
- You'll fit right in if you:
- Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.
Apply directly at White Circle →Create a free account for alerts like thisView White Circle immigration profile
This listing is sourced directly from White Circle's careers page and normalized into a canonical job model.