Schema Lab
Select a tool schema to test. HalluciTrap generates hallucination scenarios and shows the cascade impact.
Hallucination Risk Score
—
—
—
—
Loading…
Step 0/0
🔬
Simulation Ready
Press Run to watch the agent fabricate tool arguments in real time and see where it goes wrong.
Cascade Impact
—
downstream steps corrupted
Run a simulation to see the cascade.
Defense Mode
✅ Attack caught at Step 1 — cascade prevented
Research Basis
ICLR 2026 "The Reasoning Trap"
More reasoning depth = more argument hallucination. Statistically significant across 14 model families tested. The models most trusted for production are also the most likely to silently fabricate tool arguments.
More reasoning depth = more argument hallucination. Statistically significant across 14 model families tested. The models most trusted for production are also the most likely to silently fabricate tool arguments.