The review bottleneck changed.
Whiteboard is an open-source desktop app built on Code OSS. Its README describes a shared canvas where coding agents can draw architecture, link diagrams and trace quotes to code, show an abstract syntax tree aware diff, and record autonomous decisions.[1]
That is a sharper response to agent work than another chat panel. The hard part is no longer obtaining a patch. The hard part is checking whether the patch still matches the requirement and whether the reviewer understands the choices hidden inside it.
Review should compress navigation, not compress away the reasons for a change.
Keep the links bidirectional.
A diagram without a code link becomes presentation. A code diff without the requirement becomes archaeology. A decision log without a challenge path becomes a sales pitch written by the system under review.
Whiteboard says its visualizations can jump to underlying code and its decision log can link agent traces to requirements and implementation. The project also lists current limits. It cannot edit files, multi-repository review is not well supported, and a shared review does not update after later edits.[1] Treat those as design boundaries, not footnotes.
Session plumbing still matters.
VS Code 1.139 runs agent harnesses in a dedicated agent-host process. The release notes say the same session can connect to multiple VS Code windows. They also describe Dev Container sessions on SSH, Tunnel, and WSL hosts, plus a central catalog that keeps lightweight session metadata separate from each full conversation database.[3]
That architecture helps agents run in the right environment and helps people find old work. It does not explain why a change was made. Session transport and review meaning are separate jobs. A useful cockpit needs both.
Human judgment needs a handle.
Fabrizio Ferri Benedetti argues that tech workers need to read proposals closely, construct arguments, separate evidence from inference, and recognize recycled ideas.[4] That is an opinion, not a software benchmark. It names the human work that review tools should make easier.
Give the reviewer four linked surfaces:
- Quote the requirement and keep its source.
- List each autonomous decision and the rejected alternative.
- Link the decision to the exact changed symbols and lines.
- Attach a test, result, and open challenge to that same route.