Researchers warn latent reasoning could hollow out CoT oversight
Six AI safety researchers argue that adopting latent reasoning architectures like COCONUT would sharply reduce CoT's oversight value; a viewpoint worth tracking.
Original event 2026-09-23
Six named researchers published an essay on LessWrong on September 23 arguing that a shift to latent reasoning architectures would sharply undermine chain-of-thought (CoT) as an oversight tool.
The essay names three architecture families: COCONUT, which would replace CoT entirely; full-bandwidth transformers, which would add a latent channel alongside CoT; and looped transformers, which deepen computation between text bottlenecks. The authors say such changes 'could very plausibly become adopted in the near future,' as vendors may trade monitorability for competitive advantage.
The authors distinguish two oversight mechanisms: models needing CoT to complete hard tasks, and models tending to verbalize plans even when unnecessary. They judge the necessity argument robust for several years absent architectural change, while the propensity argument is fragile under adversarial pressure; latent architectures would weaken both. Note this is an analysis piece, not a new experimental result — the authors themselves call the necessity case 'far from guaranteed,' and no vendor has announced adoption of such architectures.