norgitov/ trends
Technology · people · ideas
Back to discovery/LessWrong56 minutes ago

Controllable-CoT Leads to Covert Reasoning Capabilities

The study evaluates GPT-6 Astra's performance on multi-hop tasks with a secondary CoT-control instruction, demonstrating covert reasoning capabilities that outperform non-reasoning or filler token methods. It also examines open-weights models like Kimi K3, finding they lack CoT-controllability and covert reasoning. The research uses the inspect framework and Codex for dataset generation and evaluation.

Open original
SIGNAL FROM THE SOURCE
8
source points
Tracking sinceSeptember 22, 20268 source points
MomentumMore observations needed
PublishedSeptember 22, 2026Edward Cant
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

8 source points

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This research provides insights into AI models' ability to perform hidden reasoning, which could have implications for security and monitoring systems.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic