Topics
Follow your own topics →
DIFF.BLOG
New Following Discover Jobs
More
Top Writers Suggest a blog Upvotes plugin
Report bug Contact About
Sign up
Menu
New Following Discover Jobs Top Writers
More
Suggest a blog Upvotes plugin Report bug Contact About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS

TIL: How "Thinking" Models Actually Work

14 · Anup · April 14, 2026, 5:52 a.m.
artificial-intelligence Machine Learning transformers Reasoning Models
Summary
This post summarizes Lilian Weng's insights from her work, 'Why We Think', focusing on test-time compute and chain-of-thought reasoning in transformer models, specifically how the computation cost per token is linked to the model's parameters.
Read full post on www.anup.io →
MORE POSTS LIKE THIS
Autonomous Discovery of Wireless Communications Algorithms
Research Nvidia · Jul 24, 2026
Wireless Communications artificial-intelligence
recall vs reflect: Search Your Agent's Memory, or Ask It
Hindsight Blog · Jul 24, 2026
hindsight Agent Memory
Personalizing Airbnb search by learning from the guest journey
Airbnb · Jul 21, 2026
recommendation-system Travel
Controlling Reasoning Effort in LLMs
Sebastian Raschka · Jul 18, 2026
large language models Reasoning Modes
Old Painless Meets New Clueless: Anthropic LLM Fails the Palantir Test
Flying Penguin Blog · Jul 13, 2026
history Security
Smarter data generation for faster Speculator training
Red Hat · Jul 6, 2026
large language models speculative decoding
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub Continue with Google