A model's written reasoning step corresponds to a distinct pattern inside it
A new study shows the reasoning steps models write correspond to distinct internal patterns. The signal is strongest in the middle layers.
Research
A new study shows the reasoning steps models write correspond to distinct internal patterns. The signal is strongest in the middle layers.
Timothy Gowers and Peter Sarnak credit large language models with serious mathematical skill but see hard limits on genuinely new ideas. The gap is not knowledge but the intuition to pick the right route.