Jailbreak
Role-play, obfuscation, and multi-turn escalation attempts to bypass your system prompt.
Probative continuously red-teams your production AI — jailbreaks, prompt injection, PII exfiltration, tool misuse, bias, hallucination — and grades every result with an independent judge. Governance tools document your AI. We break it, and hand you the signed evidence.
Role-play, obfuscation, and multi-turn escalation attempts to bypass your system prompt.
Hostile instructions smuggled through user content, documents, and tool outputs.
Attempts to extract personal data your system holds or has seen.
Differential treatment across protected characteristics in identical scenarios.
Confident fabrication under pressure — citations, numbers, people, precedent.
Harmful output under provocation, adversarial phrasing, and persona pressure.
Whether your guardrails are the same today as the day you shipped them.
Leakage of system prompts, internal tools, and configuration through the front door.
Coaxing your agent into calling the wrong tool, with the wrong arguments, on someone else's behalf.
Pulling retrieval-indexed documents out through the answers they ground.
Whether your agent stays inside its mandate when an attacker widens it — the failure class regulators are reading about right now.
The 2026 Digital Omnibus moved the EU AI Act's high-risk deadlines. Anyone still selling you an "August 2026 high-risk deadline" is selling a date that no longer exists. Evidence that doesn't lie starts with a timeline that doesn't either.
Obligations for general-purpose AI model providers took effect.
The AI Office's enforcement powers activate: fines up to 3% of global turnover or €15M for GPAI providers. Article 50 transparency — chatbot disclosure, deepfake labeling — and Article 4 AI literacy enforce. Only Article 50(2) machine-readable marking of synthetic content slipped, to Dec 2, 2026.
Stand-alone high-risk systems (Arts. 9–15) — hiring, credit, education, essential services, biometrics — postponed from 2026 by the Digital Omnibus. Sixteen months of behavioral evidence is what a defensible conformity file is made of.
High-risk obligations for AI embedded in regulated products — machinery, medical devices, vehicles.
Every probe result is hashed into a per-system chain and signed. Change one record and the chain breaks — visibly, mathematically, for anyone who checks.
Probative has no stake in your model, your cloud, or your security vendor. A judge that grades its own remediation isn't a judge.
When no probe behaviorally tests a control, Probative reports it as untested — the honest coverage gap — instead of mapping it anyway.
Findings roll up to the EU AI Act, NIST AI RMF, ISO 42001, and the Colorado AI Act. One scan, four evidence packs.
Governance platforms inventory your AI and collect attestations. Eval libraries test what your own team thought to test. Neither hands your auditor evidence they can verify without trusting you.
| Capability | GRC / governance platforms | DIY eval libraries | Probative |
|---|---|---|---|
| Adversarial behavioral testing | Questionnaires & policies | If you build it | 11 attack classes, every scan |
| Independent judge | Self-attested | Grades its own homework | No stake in your stack |
| Tamper-evident evidence | Editable records | Logs & spreadsheets | sha256 chain, ed25519 signed |
| Framework mapping | Control checklists | Manual | EU AI Act · NIST · ISO 42001 · Colorado |
| Continuous re-verification | Annual review cycle | When someone remembers | Scheduled, signed, on repeat |
Probative produces evidence against obligations you already know you have. If you don't yet know which of the EU AI Act, Colorado, Texas, Illinois, NYC, or California regimes reach your systems, start with Obligent — our sibling instrument: a deterministic, cited applicability engine that answers which obligations, from which rule, as of which date, and says needs analysis instead of guessing. Diagnosis there; evidence here.
What Probative is not. Probative is behavioral evidence, not legal advice, and not a certification body. Evidence packs show what your AI did under attack, mapped to framework controls — your counsel and your auditor decide what it means. Probative is currently in private pilot; every claim on this page describes the shipped probe battery and evidence engine, not aspirations.
Sixteen months of signed behavioral evidence is what a defensible conformity file is made of. Start the chain now.