Tools
r/OpenSourceeAI
Viral '178x token reduction' claims debunked; memory is the hard problem
TLDR
A satirical debunk shows viral token-reduction math ignores output tokens, cache writes and multi-turn costs. The real point: retrieval is largely solved, persistent memory across sessions is the unsolved problem for.