Generation speed moves the bottleneck.
Anthropic says Claude Sonnet 5.5 produces output more than 30 percent faster than Sonnet 5 and costs up to 30 percent less per task in its tests. The list price remains $2 per million input tokens and $10 per million output tokens. Anthropic attributes the lower task cost to using fewer tokens for the same work.[1]
Those are vendor measurements. They still matter. A faster model can shorten the draft loop. It can also increase the number of changes that reach review in one afternoon. If review capacity stays fixed, generation speed turns into a longer queue.
Spend the speed gain on evidence, not output volume.
One disclaimer cannot check four kinds of claims.
Glyph's essay on serious AI products makes a useful product argument. A warning that AI can make mistakes gives the user a job without giving them a workbench. The essay proposes visible claim checks, readable citations, human notes, and clear data provenance.[2]
The check depends on the claim. A quotation needs the original source. A number needs a repeatable calculation. A behavior claim needs an executable test. A design or policy decision needs a named owner and stated tradeoff. A generic "reviewed" badge hides these differences.
Automated review is another producer.
GitHub's current Copilot code review documentation separates model credits from the runner time used for repository context and tools. It also says reviews still run in a more limited form when the supporting Actions path fails. The same page prices Lite and Balanced review at different estimated ranges.[3]
That is a useful reminder. Review output has a mode, a cost, an execution path, and a failure state. Save those facts. Do not treat an automated comment as independent proof merely because it arrived in the review column.
Make the check visible before merge.
- Copy the exact claim from the generated patch, summary, or review.
- Classify it as source, number, behavior, or judgment.
- Attach the matching evidence. Do not substitute a test for a product decision or a citation for runtime behavior.
- Record who checked it and what remains open.
- Keep the claim open until the evidence can be inspected by the next reviewer.
The release should get faster only when the check lane gets faster too.