--- title: 'AI Red Team Engineer at White Circle' canonical: 'https://feeny.ai/job/ai-red-team-engineer-white-circle-j1xbb74qxv2e' type: 'job' last_seen: '2026-09-05' --- # AI Red Team Engineer at White Circle - **Company:** White Circle - **Location:** •, United States - **Compensation:** $60k–$90k - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-07-06 - **Last confirmed live:** 2026-09-05 - **Apply:** https://jobs.ashbyhq.com/whitecircle/d2c5f461-55d2-4c20-97a6-41a0fa0b51e9 ## Job description TLDR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly, automate the repetitive attacks, and turn their findings into clear evidence that powers customer demos, security reviews, and sales conversations. You'll own hands-on adversarial testing end to end: find the failure, prove it, script it, and write it up. ## About us White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale. - We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others - We process over one hundred million API calls every month - We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model We’re a small, highly focused team. If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built – you’re the one we need. You will: - Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products. - Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse. - Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports. - Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases. - Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it. - Convert successful attacks into regression tests and product requirements. - Track new red-team and safety techniques and fold the useful ones into our tests. - Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations. You'll fit right in if you: - Genuinely love breaking things and reasoning adversarially. - Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty. - Have strong Python scripting skills. - Have experience testing APIs, web apps, backends, or SaaS products. - Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling. - Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion). - Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact. - Can separate real customer risk from low-impact prompt tricks. - Write clear, reproducible bug reports in clear English. - Can move fast without perfect requirements. - Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material. ## A BIG PLUS: - Experience with Burp Suite, Postman, Playwright, pytest. - Experience with modern LLM red-teaming automated agents and pipelines. - Familiarity with LangChain, LangGraph, LlamaIndex, RAG pipelines, AI agents, tool/function calling, and LLM-as-judge evaluation. - Familiarity with OWASP LLM Top 10, OWASP Web Top 10, MITRE ATLAS, or other AI security taxonomies. - Experience testing RAG systems, AI agents, tool-calling workflows, browser agents, or internal copilots. - Experience writing customer-facing security reports. - Experience with trust & safety, abuse prevention, fraud, moderation, or platform security. - Experience building eval pipelines, regression suites, dashboards, or CI-friendly security tests. - A track record in CTFs, red-team competitions, or responsible-disclosure / bounty programs. ## Why White Circle - Paid time off in line with your local regulations, no matter where you work from - Meaningful equity package - All the hardware, tools, and services you need - Covered subscriptions for AI agents - Team off-sites twice a year: we've recently been to the Alps and to Saint-Tropez ## How we hire 1. Intro call with HR (25 min) 2. Take-home test task 3. Technical interview (60 min) 4. Final call with CEO (45 min) Please submit your application in English ## About White Circle ## Company Overview - **One-liner**: White Circle is a control layer for AI in production, providing safety, security, evals, and performance optimization through a unified API. - **Entity Type**: Private (Seed stage) - **Headquarters**: Dover, Delaware, United States (operational hub in Paris, France) - **Founded**: 2025 - **Founders**: Denis Shilov (Founder & CEO) ## Core Business - **Primary industry**: AI Governance, AI Security, AI Observability - **Target customers**: B2B, Enterprise (companies deploying AI agents, chatbots, or LLM-powered applications) - **Mission or purpose**: To give organisations visibility into how their AI behaves, help them respond when things go wrong, and provide a unified system for improving reliability, safety and compliance. ## Products & Services - **White Circle Platform**: A single API that integrates three layers: - **Protect**: Custom low-latency guardrails for blocking unsafe inputs, preventing jailbreaks, detecting prompt injections, PII leakage, and sensitive data leaks. - **Observe**: Real-time analytics including topic classification, user behavior clustering, custom metrics, error rate analysis, sentiment analysis, and risk scoring. - **Improve**: Model routing, dynamic prompt engineering, context enrichment, and automated optimization based on labelled user feedback. - **CircleGuardBench**: A proprietary benchmark for evaluating AI moderation models across harm detection, jailbreak resistance, false positives, and latency. - **KillBench**: A benchmark of hidden LLM biases in critical decisions (life-and-death scenarios). ## Market Standing - **Valuation/Market Cap**: Not disclosed - **Key Metric**: $11M raised in seed funding (announced May 2026) - **Notable Investors/Partners**: Backed by prominent AI and technology leaders including Romain Huet (ex-Head of Product @DeepMind), Dirk Kingma (Anthropic), Guillaume Lample (Mistral), Thomas Wolf (Hugging Face), Olivier Pomel (Datadog), François Chollet (Keras), Mehdi Ghissassi (ex-DeepMind), Paige Bailey (DeepMind), and David Cramer (Sentry). - **Growth Signals**: - 21 employees (+283.3% YoY, +9.5% monthly growth) - Monthly website traffic: 32,090 visits (+682.7% yearly) - Active job postings: 6 (monthly job posting growth +20%) - SOC 2 Type II and HIPAA compliant - Operates in 7 countries (France, Russia, Netherlands, Serbia, Switzerland, Georgia, Spain) ## Competitive Advantages - **Proprietary self-adjusting models** optimized for low-latency and high performance - **Unified platform** that combines safety, security, evaluation, and performance optimization – no need for multiple point solutions - **Enterprise-grade compliance** (SOC 2 Type II & HIPAA) - **Strong network effects** from backing by leaders at OpenAI, Anthropic, DeepMind, Hugging Face, and Datadog - **Automated red-teaming** and continuous improvement through user feedback ## Strategic Focus - Accelerate product development and expand the team across the US, UK, and Europe - Grow global customer base across healthcare, finance, hiring, and security verticals - Continue building benchmarks (CircleGuardBench, KillBench) to shape AI safety standards - Maintain deep integration with the AI open-source ecosystem ## Why Work Here - **Culture**: “Smart, friendly, and fun” team with a strong engineering and research focus (43% technical staff) - **Location**: Based in Paris, France with a hybrid work model (office + remote) - **Benefits**: - Strong salary - 20 paid days off (plus ability to take more for recovery) - Free education (company covers learning costs) - Team offsites several times a year - **Open roles** (June 2026): QA Engineer, Product Engineer (Backend), DevOps Engineer, Product Designer, AI Engineer (Audio), Data Engineer - **Impact**: Build the future of AI safety at a well-funded, high-growth startup with world-class investors ## Sources 1. [whitecircle.ai](https://whitecircle.ai) 2. [whitecircle.ai/careers](https://whitecircle.ai/careers) 3. [LinkedIn – White Circle](https://www.linkedin.com/company/whitecircle) 4. [Tech.eu – White Circle lands $11M](https://tech.eu/2026/05/12/white-circle-lands-11m-to-help-companies-secure-ai-systems/) ## Other roles at White Circle - [Finance Lead](https://feeny.ai/job/finance-lead-white-circle-us-san-francisco-or-s3jtsbzkwt7f) — US San Francisco OR, NY - [Product Brand Designer](https://feeny.ai/job/product-brand-designer-white-circle-us-san-francisco-or-3v7jkx8gskhf) — US San Francisco OR, NY - [Founding Account Executive](https://feeny.ai/job/founding-account-executive-white-circle-us-san-francisco-or-rns9xb4d5gj6) — US San Francisco OR, NY - [Recruiter (Research)](https://feeny.ai/job/recruiter-research-white-circle-remote-1vekq2rdfrvz) - [DevOps Engineer](https://feeny.ai/job/devops-engineer-white-circle-paris-y84p4fve5nn1) — Paris, France - [Founding Events Lead](https://feeny.ai/job/founding-events-lead-white-circle-new-york-wczngn43ffxc) — New York, NY - [Partnerships Manager](https://feeny.ai/job/partnerships-manager-white-circle-london-c3ep31m67xse) — • London, United Kingdom - [Senior Data Labeler](https://feeny.ai/job/senior-data-labeler-white-circle-paris-vtp78vea9yp4) — Paris, France - [ML Infrastructure Engineer](https://feeny.ai/job/ml-infrastructure-engineer-white-circle-paris-cjvfwvec3k00) — Paris, France - [Multimodal ML Engineer](https://feeny.ai/job/multimodal-ml-engineer-white-circle-paris-kd1xveq65a91) — Paris, France