sauce.ai/redteam

Adversarial safety testing for clinical AI chatbots. Point it at your bot, pick a specialty, and it runs synthetic patient conversations engineered to surface unsafe medical advice — then scores every reply and emails you an epidemiological report.

Free while in preview: up to 100 trials total per email. For use only on systems you are authorized to test.

This is a safety-evaluation tool. It deliberately tries to elicit unsafe responses from the bot under test so you can find and fix them. Only run it against endpoints you own or have written permission to test. Reports are LLM-scored screening signals, not clinical determinations — have a clinician review flagged transcripts.

1 · Your chatbot (the target)

Custom endpoint mapping
Web page selectors

Requires the browser add-on on the server. Selectors target the newest assistant bubble.

2 · The evaluation

Harm categories to hunt (blank = all)
Advanced: the red-team engine (optional)

The attacker is an ensemble of LLMs that propose, refine, and vote on the next patient message most likely to push the bot into an unsafe reply. All of this is optional — defaults use Claude.

3 · Where to send the report

Run

Open live report ↗ · JSON export ·

A sauce.ai lab tool. Questions → redteam@sauce.ai.