Memory architectures that outperformed KV-cache in production
Field notes from teams using external memory, summarization, and retrieval to control long-running agents.
Agentic models are quietly rewriting the assumptions baked into a decade of product design. We talked to eleven teams shipping them, and the takeaway is less about intelligence and more about delegation.
Field notes from teams using external memory, summarization, and retrieval to control long-running agents.
Retrieval-augmented generation survived the hype cycle because the architecture became more disciplined.
How the current eval tooling landscape compares for teams trying to ship reliable AI workflows.
Most evaluations measure intermediate skill. Operators need finish-rate, recovery, escalation, and cost-to-complete.
The funding story is really a market-structure story: distribution, enterprise trust, and lower-cost specialization.
A new wave of sequence modeling research is pushing efficient attention back into production conversations.
The interesting question is no longer what can the model do — it's what should we let it do unsupervised, and on whose behalf.
Most evaluations measure intermediate skill. Operators need finish-rate, recovery, escalation, and cost-to-complete.
The tools worth trying, a prompt you can steal, and the news that affects your work. No hype, no jargon.