White Circle
AI Red Team Engineer
Global
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Work mode
- Remote-friendly
- First seen by hirly
- 11 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
TL;DR: We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly, automate repetitive attacks, and turn their findings into clear evidence that powers customer demos, security reviews, and sales conversations. You'll own hands-on adversarial testing end to end: find the failure, prove it, script it, and write it up.
About us
White Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.
We’ve recently raised our Series A funding round, taking our total funding to $70M. Our investors include top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others
We process over one hundred million API calls every month
We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model
We’re a small, highly focused team. If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built – you’re the one we need.
What you’ll do
Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products.
Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse.
Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports.
Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases.
Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it.
Convert successful attacks into regression tests and product requirements.
Track new red-team and safety techniques and fold the useful ones into our tests.
Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.
You'll fit right in if you
Genuinely love breaking things and reasoning adversarially.
Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty.
Have strong Python scripting skills.
Have experience testing APIs, web apps, backends, or SaaS products.
Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling.
Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion).
Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact.
Can separate real customer risk from low-impact prompt tricks.
Write clear, reproducible bug reports in clear English.
Can move fast without perfect requirements.
Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.
A big plus
Experience with Burp Suite, Postman, Playwright, pytest.
Experience with modern LLM red-teaming automated agents and pipelines.
Familiarity with LangChain, LangGraph, LlamaIndex, RAG pipelines, AI agents, tool/function calling, and LLM-as-judge evaluation.
Familiarity with OWASP LLM Top 10, OWASP Web Top 10, MITRE ATLAS, or other AI security taxonomies.
Experience testing RAG systems, AI agents, tool-calling workflows, browser agents, or internal copilots.
Experience writing customer-facing security reports.
Experience with trust & safety, abuse prevention, fraud, moderation, or platform security.
Experience building eval pipelines, regression suites, dashboards, or CI-friendly security tests.
A track record in CTFs, red-team competitions, or responsible-disclosure / bounty programs.
Compensation & benefits
Competitive compensation package, including equity
Flexible Time Off
Language lessons to help you improve your English or French
Learning and development support for courses, conferences, and opportunities to grow your skills
All the hardware, subscriptions, tools, and services you need
Team off-sites twice a year: we’ve recently been to the Alps, Saint-Tropez, and Marbella
Process
Intro call with Talent Team
Test assignment
Technical interview
Final call with CEO
Similar jobs
- Purple Team Engineer – Retesting & Control ValidationCardinalhealth · IND07First seen 2d ago
- IT Operations Engineer - .Net Team EngineerChubb · PhilippinesFirst seen 3d ago
- Spark Team EngineerArkana Laboratories · ArkansasFirst seen 4d ago
- Red Team EngineerState Street · 3 LocationsFirst seen 5d ago
- Additive Process Team Engineering LeaderGevernova · GreenvilleFirst seen 5d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job