machine made worldsA journal of artificial intelligence

Ideas, collected

Local AI.

Essays, guides and observations. Find something worth sitting with.

13 articles

What is Prompt caching?

Prompt caching is a provider-billed discount for reusing a stable prompt prefix. What it caches, what it costs, and where the breakpoints go.

1 min read↗

What is Prefix caching?

Prefix caching reuses stored attention work across requests that share the same opening tokens. How it hits, and how builders keep it hot.

2 min read↗

What is KV Cache?

KV cache reuses past attention keys and values so long chats and agents run faster. What it costs, and how builders shrink it.

2 min read↗