Basilisk AI Red Teaming
Learn Basilisk from a safe first preview through deterministic model-security testing, evidence review, evaluation, desktop operation, and CI automation.
Do the work in order.
- 0140 min↗
Install, scope, and preview
Create an isolated Python environment, verify Basilisk, protect provider secrets, and preview a scan before execution.
- 0240 min↗
Run the ground-truth lab
Start the Docker-isolated Basilisk lab on loopback and verify its secure and vulnerable OpenAI-compatible routes.
- 0350 min↗
Run the first bounded validation
Execute one injection module against the vulnerable lab route, then repeat it against the secure control.
- 0455 min↗
Choose modules, probes, and evolution
Move from one deterministic probe to a justified module set without turning a bounded validation into an uncontrolled campaign.
- 0555 min↗
Measure posture, differences, and evals
Use Basilisk posture, diff, and eval workflows to compare systems and turn repeatable expectations into regression checks.
- 0650 min↗
Preserve sessions, evidence, and reports
Review sessions, verify audit evidence, and export the minimum report needed for a reproducible security conclusion.
- 0755 min↗
Use the desktop and CI safely
Reproduce the bounded workflow in the Basilisk desktop and add a deterministic, resource-safe check to CI.
