Thinking in Tokens
Home
About
Categories
Tags
Archives
Recent Posts
all posts
Test-Time Compute Explained: Why Sampling More Beats Training Bigger, and When It Doesn't
posted in
AI Scaling
Fri 25 September 2026
The Anatomy of an Agent Loop: Why One LLM Call Is Not Enough
posted in
AI Agents
Thu 24 September 2026
Agent Memory and Skills: What to Keep, What to Load, What to Forget
posted in
AI Agents
Thu 24 September 2026
Long-Context and State Management for AI Agents
posted in
AI Agents
Wed 23 September 2026
How AI Agents Turn Text Into Actions: Tool Calling, JSON, CodeAct, and MCP
posted in
AI Agents
Tue 22 September 2026
AI Agents Forget Instructions Because Context Compaction Deletes Them
posted in
AI Agents
Sat 19 September 2026
LLM Agent Costs Are Quadratic: The Token Trap and the Prompt Caching Fix
posted in
AI Agents
Sat 19 September 2026
What Actually Makes an LLM an Agent: The Formal Formulation, Policy Engine, and Harness
posted in
AI Agents
Thu 17 September 2026
What Is an AI Agent? A Practical Guide to Building Your First One
posted in
AI Agents
Sun 13 September 2026
Welcome to Thinking in Tokens
posted in
Introduction
Wed 09 September 2026