Technologist Mag
  • Home
  • Tech News
  • AI
  • Apps
  • Gadgets
  • Gaming
  • Guides
  • Laptops
  • Mobiles
  • Wearables
  • More
    • Web Stories
    • Trending
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
George Santos Just Got Hit With Kalshi’s First-Ever Lifetime Ban

George Santos Just Got Hit With Kalshi’s First-Ever Lifetime Ban

31 August 2026
EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

31 August 2026
JBL thinks the future of gaming audio is smarter, not just wireless

JBL thinks the future of gaming audio is smarter, not just wireless

31 August 2026
Time Doesn’t Heal All Wounds

Time Doesn’t Heal All Wounds

31 August 2026
Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

31 August 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Technologist Mag
SUBSCRIBE
  • Home
  • Tech News
  • AI
  • Apps
  • Gadgets
  • Gaming
  • Guides
  • Laptops
  • Mobiles
  • Wearables
  • More
    • Web Stories
    • Trending
    • Press Release
Technologist Mag
Home » A New Trick Reveals AI Models’ Inner Thoughts
Tech News

A New Trick Reveals AI Models’ Inner Thoughts

By technologistmag.com11 August 20263 Mins Read
A New Trick Reveals AI Models’ Inner Thoughts
Share
Facebook Twitter Reddit Telegram Pinterest Email

Computer scientists recently discovered a way to extract the hidden “thinking” that frontier AI models perform as they work through complex problems.

The findings provide some evidence—although not conclusive proof—that certain Chinese models may have been trained by “distilling” reasoning information from US models that was supposedly hidden because of how closely some of their thinking or reasoning patterns seem to match. The researchers have also demonstrated that the method could be used to recover personal information, like passwords and API keys, from a model’s inner reasoning, although this vulnerability has been fixed.

“All major frontier model providers we tested share this vulnerability,” says Alexander Panfilov,⁩ a computer scientist at University of Tübingen in Germany who was involved with the work. “It can lead to personal information leakage, and it enables large-scale reasoning distillation attacks.”

Panfilov and colleagues from the University of Tubingen, the Max Planck Institute, the AI safety institute MATS Research, and the security company Snyk identified the same issue with frontier models from OpenAI, Anthropic, and Google that are accessed via an application programming interface or API.

In a paper laying out the work, the researchers show that the open-weight or downloadable Chinese model Kimi K3 from Moonshot AI produces a strikingly similar output to the hidden reasoning traces—the written-out reasoning steps involved in solving a problem—of Claude Opus 4.8 and GPT 5.6 Sol for certain prompts. Despite the similarities, they note that the work “cannot causally establish distillation.” They found that two other open-weight models, China’s DeepSeek and Inkling from the US company Thinking Machines, did not exhibit this kind of reasoning similarity with Claude Opus.

Moonshot AI and Z.ai did not respond to a request for comment by time of publication.

Distillation is a well-established, widely used technique for efficiently copying the capabilities of existing models over to new ones, and is especially common in the development of open-weight or fully downloadable models.

Lately, however, distillation has become a controversial topic, because of claims that Chinese AI companies use it to essentially copy the best US models. In February, OpenAI told US lawmakers that DeekSeek seemed to have copied one of its models to build a reasoning model called R1. In June, Anthropic told lawmakers that Alibaba had systematically distilled its models in order to build its own, called Qwen.

There’s no indication that Chinese AI companies used this specific technique to distill US-based AI models. But Panfilov and collaborators say that using their method would make it possible to distill more information from closed models than previously realized.

Mini-Me Models

Advanced AI models solve difficult problems by breaking them into constituent parts that are analyzed in turn in a kind of artificial reasoning or “chain of thought.” Companies tend to keep a proprietary model’s reasoning secret to prevent others from using them to train new ones. However, they typically also send an encrypted version of that reasoning to a user’s computer in a way that offloads some computation.

The researchers’ attack relies on the fact that most AI companies also provide related models of different sizes. Larger models are more capable but also more computationally expensive to run and more expensive to access. Users may choose smaller, weaker models for certain tasks to lower costs.

Panfilov and his colleagues found that feeding encrypted reasoning traces to a smaller version of the same model can reveal the hidden reasoning inside. The smaller models have received less alignment training, meaning that, unlike the bigger ones, they are less likely to refuse to reveal their inner thoughts.

Share. Facebook Twitter Pinterest LinkedIn Telegram Reddit Email
Previous ArticleAI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one
Next Article Meta’s new AI model runs entirely offline, but your GPU needs to keep up

Related Articles

George Santos Just Got Hit With Kalshi’s First-Ever Lifetime Ban

George Santos Just Got Hit With Kalshi’s First-Ever Lifetime Ban

31 August 2026
EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

31 August 2026
JBL thinks the future of gaming audio is smarter, not just wireless

JBL thinks the future of gaming audio is smarter, not just wireless

31 August 2026
Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

31 August 2026
The Best MagSafe Phone Grips

The Best MagSafe Phone Grips

31 August 2026
NASA’s new telescope can see 100 times more sky than Hubble in a single shot

NASA’s new telescope can see 100 times more sky than Hubble in a single shot

31 August 2026
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Don't Miss
EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

EXCLUSIVE: Just Play Dead director Martin Campbell and writer Dan Gordon on Samuel L. Jackson and Eva Green’s twisted thriller

By technologistmag.com31 August 2026

Samuel L. Jackson and Eva Green try to kill each other in their new action-thriller,…

JBL thinks the future of gaming audio is smarter, not just wireless

JBL thinks the future of gaming audio is smarter, not just wireless

31 August 2026
Time Doesn’t Heal All Wounds

Time Doesn’t Heal All Wounds

31 August 2026
Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

Samsung Galaxy S27 Ultra could trade 5x zoom for a much bigger telephoto sensor

31 August 2026
PlayStation Schedules Back-To-Back State of Plays For Thursday

PlayStation Schedules Back-To-Back State of Plays For Thursday

31 August 2026
Technologist Mag
Facebook X (Twitter) Instagram Pinterest
  • Privacy
  • Terms
  • Advertise
  • Contact
© 2026 Technologist Mag. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.