Archive · page 2 of 2
Older issues
- Daily
Mind viruses spread between LLM agents, and a one-line warning nearly stops them
Plus: Zed's Delta, Grok 4.6, and a memory harness ladder. 5 min.
Aug 13, 2026 · 5 min

- Daily
In LangChain's benchmark, only 7% of agent turns needed a frontier model
Plus: BDH-CQ's 29.5% on ARC-AGI 1, Meta's first Apache 2.0 model, and a TDD experiment at Thoughtworks. 5 min.
Aug 12, 2026 · 5 min

- Daily
Claude Code turns auto mode on for everyone; reward-hack monitors catch 28%
Plus: malicious skill files on arXiv, Stagehand v4, and Meta's local 30B. 5 min.
Aug 11, 2026 · 5 min

- Daily
A public harness reproduces DeepSeek's 82.7% on Terminal-Bench, 445 trials deep
Plus: 4 talks from AI Engineer, a GitHub retirement, and sycophancy for smart people. 5 min.
Aug 10, 2026 · 5 min




