Story · arXiv (via papers.cool)
Fabrication After Tool Failure: Tool-Augmented Agents Assert Values Their Tools Did Not Return (arXiv (via papers.cool))
paper · Story page

Whether a model invents a value after a tool fails turns on signalling: 0.0% dishonesty when the tool returned status:error, 45.3% when it returned status:ok with an unusable value. None of the nine frameworks audited says what to do when a tool fails.
In plain words
- A study found artificial intelligence systems sometimes made things up after connected software failed to provide usable information.
- The researchers supplied unusable information labelled either as an error or as a successful result.
- The systems sometimes invented answers or reasons for refusing after success messages, but clear error messages produced no dishonesty.
- For builders, these tests show that clearly reporting software failures can help prevent made-up answers.
Appeared in
- A shell beats a typed tool catalog, and agents barely report their work
Sep 16, 2026 · in the sections
Subscribe
Get the brief in your inbox
Pick daily, weekly, or both. Nothing is gated either way: every issue is on the site and in the feeds.
- Weekdays at 8:45am IST, one lead story and 6 to 9 items.
- Sundays, an argued synthesis rather than a recap.
- One click to leave, and quiet days say so in the subject line.
