Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Register
Metaverse Media Group
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
Metaverse Media Group

AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds

AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds

The Decoderby The Decoder
12 September 2026
Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model’s internal states, especially in the middle layers. That matters for AI safety, because models process more than their visible chain of thought reveals. The article AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds appeared first on The Decoder….


Jonathan Kemper


Sep 12, 2026

Image description

Nano Banana Pro prompted by THE DECODER

Can the distinct reasoning steps a language model shows in its text output also be found in its internal states? A new study put it to the test.

When a reasoning model solves a task step by step, it does different things along the way: reading data, breaking down the problem, retrieving a formula, running a calculation. Researchers at South Korea’s KAIST and Naver AI Lab wanted to know whether those reasoning steps can also be separated from one another inside the model’s numerical representations. They can, and the signal is strongest in the middle layers.

The team defined eight recurring reasoning operations, including extraction, decomposition, formula recall, deduction, and computation. They had three models (Qwen2.5-7B, Qwen3-8B, and Gemma4-31B) solve math problems, split the solution paths into segments, and then used GPT-5 to label each segment with one of those operations.

Two-part graphic from the paper. Left side shows the same Qwen3-8B response three times with color-coded highlights for Extraction, Recall, and Decomposition operations. Right side shows a scatter plot of segments along Decomposition score and Recall score axes.
The same response produces a different activation pattern depending on which reasoning operation is being probed. Segments cluster along the corresponding direction in the scatter plot. | Image: Jeong et al.

Reasoning steps are clearly separable inside the model

The different reasoning operations can be reliably told apart in the models’ internal representations, and this holds across all three models tested. The separation peaks in the middle layers.

The researchers checked whether simple word choice could account for the effect. A classifier that only looked at the tokens used performed worse than one analyzing internal representations. Position within the solution path didn’t explain it either. That means the internal states carry information about the type of reasoning step that goes beyond surface-level wording.

Left side shows AUROC values for eight reasoning operations in Qwen3-8B, Qwen2.5-7B, and Gemma4-31B with confidence intervals. Right side shows the average AUROC across all operations plotted over model depth stages Embed, Early, Middle, and Late.
Across all three models, reasoning operations can be reliably separated, with the clearest signal in the middle layers. | Image: Jeong et al.

Same words, different representations depending on the reasoning step

Common function words like “a,” “is,” or “the” show up across very different reasoning steps. In the early layers, their representations are still jumbled together, but by the middle and later layers they separate according to the surrounding operation. The same word gets a different internal representation depending on which reasoning step it belongs to.

The researchers also tested whether a reasoning step forms in isolation. When they blocked attention to the preceding 30 tokens through a targeted intervention, the signal for that operation weakened. Reasoning steps don’t emerge on their own but build on the preceding context.

Grid of 18 scatter plots for three operation pairs across the embedding layer and layers 1, 11, 21, 27, and 36 of Qwen3-8B. Each point represents one occurrence of a shared token.
The same words overlap in early layers and separate by surrounding operation in the middle and late layers. | Image: Jeong et al.

Even on incorrectly solved problems, the type of step the model was performing stayed identifiable, whether it was computing, retrieving a formula, or deducing. A flawed computation step still looked like a computation step internally, even when the result was wrong.

The separability held up in additional tests too. It replicated with Llama-3-8B, and for Qwen3-8B the trained classifiers transferred successfully to GPQA-Diamond and MATH-500. That said, the experiments are limited to math tasks and a handful of models. Whether these findings can be used to catch errors or steer a model mid-generation remains an open question for future work.

The relationship between text output and internal computation matters for AI safety. Reading the chain of thought is one of the few oversight tools available, according to OpenAI, but Anthropic showed that models only disclose the hints they used in 25 to 39 percent of cases. A method that translates a model’s internal vectors into readable text revealed that Claude Opus 4.6 processes more than what shows up in its output reasoning. And with OpenAI’s Astra model, the Recurrent Depth technique shifts part of the reasoning into internal numerical representations, which is the space the KAIST study investigates.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive “AI Radar” frontier report six times a year, full archive access, and access to our comment section.


Subscribe now

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: “AI Radar” — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI


Subscribe to The Decoder

Read the full article on The-Decoder.com
in AI
Reading Time: 5 mins read
0
0
22
VIEWS
Share on TwitterShare on Facebook

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now
ADVERTISEMENT

Related Posts

ChatGPT Images 2.5 vs Nano Banana 2: Which One is Better?
AI

ChatGPT Images 2.5 vs Nano Banana 2: Which One is Better?

2 hours ago
23
From Hacks to Bioweapons, Claude Misuse Is Now Everywhere
AI

From Hacks to Bioweapons, Claude Misuse Is Now Everywhere

5 hours ago
23
OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google
AI

OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google

5 hours ago
24

Comments

Please login to join discussion
ADVERTISEMENT

Latest News

  • All
  • Crypto
  • NFTs
  • Technology
  • Business
Bitcoin Price Takes a $3,215 Ride Just to Land Back at $77K
Crypto

Bitcoin Price Takes a $3,215 Ride Just to Land Back at $77K

Bitcoin.com News
by Bitcoin.com News
1 hour ago
22
AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds
AI

AI models’ written reasoning steps correspond to distinct internal patterns, a new study finds

The Decoder
by The Decoder
2 hours ago
22
Bitcoin’s Price Just Drew an $85M Bet From Last Year’s ETH Top-Caller
Crypto

Bitcoin’s Price Just Drew an $85M Bet From Last Year’s ETH Top-Caller

Bitcoin.com News
by Bitcoin.com News
2 hours ago
23
ChatGPT Images 2.5 vs Nano Banana 2: Which One is Better?
AI

ChatGPT Images 2.5 vs Nano Banana 2: Which One is Better?

Decrypt
by Decrypt
2 hours ago
23
Hunter Biden’s LAPTOP Burns $3.5M After Eric Trump and Beeple Mentions
Crypto

Hunter Biden’s LAPTOP Burns $3.5M After Eric Trump and Beeple Mentions

Bitcoin.com News
by Bitcoin.com News
3 hours ago
24
Betmgm Bets on Injury Refunds and a $50K Free-to-Play NFL Jackpot
Crypto

Betmgm Bets on Injury Refunds and a $50K Free-to-Play NFL Jackpot

Bitcoin.com News
by Bitcoin.com News
4 hours ago
22
Load More
Next Post
Bitcoin Price Takes a $3,215 Ride Just to Land Back at $77K

Bitcoin Price Takes a $3,215 Ride Just to Land Back at $77K

ADVERTISEMENT

Follow Us

Categories

  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
Subscribe to our Newsletter

© 2022 Metaverse Media Group – The Metaverse Mecca

Privacy and Cookie Policy | Sitemap

Welcome Back!

Sign In with Google
OR

Login to your account below

Forgotten Password? Sign Up

Create New Account!

Sign Up with Google
OR

Fill the forms below to register

*By registering into our website, you agree to the Terms & Conditions and Privacy Policy.
All fields are required. Log In

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Sign Up
This website uses cookies. By continuing to use this website you are giving consent to cookies being used. Visit our Privacy and Cookie Policy.

Subscribe to our newsletter

Get the latest news & win monthly prizes

Subscribe to our newsletter

For the Latest News and Monthly Prize Giveaways

Join Now
Join Now