Part of Leviathan Platform · standalone license available

Probe Kit

A validated social-engineering probe corpus, run continuously against your own agent through one simple interface. Not a one-time pentest — know whether your agent still resists attacks after every model, prompt, tool, and policy change. Disclosure-to-exploitation time has collapsed from roughly two years (2018) to about ten hours (Palo Alto research, 2026, 82% of vulnerabilities) — a quarterly review can't keep pace with that, continuous testing is the only thing that can.

$2,000/yr, per org — flat, unlimited seats and environments

What it actually solves

Three real problems with testing an agent's judgment.

Coverage

Nobody actually attacks the agent before it ships

A validated corpus of real social-engineering scenarios — authority claims, hearsay approval, urgency, fabricated errors — run against your own agent through one Callable[[str], ProbeOutcome] interface.

Scoring

A pass isn't a pass isn't a pass

Deliberation-aware resistance scoring: an agent that refuses instantly scores differently from one that complies after being talked into it.

Evidence

A red-team log that can be quietly edited proves nothing

Every trial result lands in a hash-chained, tamper-evident evidence graph — the same append-only pattern the rest of Leviathan Platform uses for its own audit trail.

Four modules

probe_library · probe_engine · resistance_score · evidence_graph.

A real, growing corpus — 34 probes across 26 angles as of this writing, each grounded in disclosed research (real papers, real Black Hat/DEF CON talks), not invented. Run it once, or run it continuously as a regression gate.

Pricing

One flat license. No per-seat fees.

Card checkout: the price is firm; self-serve checkout isn't wired up for this product yet. Reach out and we'll send payment instructions directly.