Posts

An AI Agent Flight Recorder Belongs in One Portable File, the Way HAR Did It for HTTP

Posts

LLM Response Caching Pays Off in CI, Evals and Agent Retries, Not in Chat

Posts

Most MCP Token Waste Is in Tool Results: Put a Deterministic Reducer Between Server and Model

Posts

Routing LLM Requests by Token Count, Privacy Tag and Cost Ceiling From One Config File

Posts

The New MCP Spec Caches Tool Lists but Not Tool Calls, and a Caching Proxy Fills the Gap

Posts

Why Private Domain Data Is the Real Key to AI That Actually Works