#
DIFF.BLOG
New
Following
Discover
Jobs
More
Top Writers
Suggest a blog
Upvotes plugin
Report bug
Contact
About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join Diff.blog
TOPICS
TIL: How "Thinking" Models Actually Work
14
·
Anup
·
April 14, 2026, 5:52 a.m.
artificial-intelligence
Machine Learning
transformers
Reasoning Models
Summary
This post summarizes Lilian Weng's insights from her work, 'Why We Think', focusing on test-time compute and chain-of-thought reasoning in transformer models, specifically how the computation cost per token is linked to the model's parameters.
Read full post on www.anup.io →
MORE POSTS LIKE THIS
Autonomous Discovery of Wireless Communications Algorithms
Research Nvidia ·
Jul 24, 2026
Wireless Communications
artificial-intelligence
recall vs reflect: Search Your Agent's Memory, or Ask It
Hindsight Blog ·
Jul 24, 2026
hindsight
Agent Memory
Personalizing Airbnb search by learning from the guest journey
Airbnb ·
Jul 21, 2026
recommendation-system
Travel
Controlling Reasoning Effort in LLMs
Sebastian Raschka ·
Jul 18, 2026
large language models
Reasoning Modes
Old Painless Meets New Clueless: Anthropic LLM Fails the Palantir Test
Flying Penguin Blog ·
Jul 13, 2026
history
Security
Smarter data generation for faster Speculator training
Red Hat ·
Jul 6, 2026
large language models
speculative decoding
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub
Continue with Google