Series
AI Under the Hood
A guided path from one model interaction through tokens, calls, agent loops, memory, trust, and production systems.
Start with The Round Trip: What Happens When You Hit EnterFoundations
Follow one request from the interface to the model and back, then establish the mechanical truths the rest of the series will examine.
The Model
Look inside the prediction engine, from tokenization and next-token selection to training boundaries, hallucination, and sampling.
The Call
Trace one request in detail, including payload assembly, statelessness, context limits, streaming, and cost.
The Loop
See how external orchestration turns repeated model calls and ordinary code into tool-using agents.
Memory & Knowledge
Separate model state from application-managed memory, then examine embeddings, retrieval, and their failure modes.
Reasoning & Trust
Ask what visible reasoning can establish, why fluent explanations can mislead, and how applications engineer trust.
Context at Scale
Understand degradation, compaction, caching, rising cost, and the deliberate management of a finite context budget.
Multi-Agent & Production Systems
Compose agent loops into larger systems while accounting for isolation, reliability, security, and production constraints.