I Can Generate Faster Than I Can Think: Tokenmaxxing, verification hell, and the creativity trap of infinite AI
- tokenmaxxing
- agentic coding
I used to think the bottleneck was execution.
02 / Notes
Essays, notes, and build logs on design systems, AI interfaces, and shipping polished web products.
I used to think the bottleneck was execution.
In 2023, almost everyone thought agent memory was mostly a context-window problem.
My mentor survived three part-time jobs and bone-deep poverty with a simple tool: the Pomodoro technique. It worked for him. And for years, it has worked…
Search has been with us since the Yahoo! days, and for most people it feels “solved.” You type a few words, hit enter, and out comes a list that’s “good…
I’ve talked a lot about the LongMemEval dataset on my blog. If you’re not deep in the agent memory space, you’ve probably never heard of it. It’s a poster…
If you’ve been keeping an eye on AI lately, you’ve probably noticed a big shift. Models have moved beyond just answering questions or generating clever…
Note: This article was updated on September 2, 2025, to incorporate clarifications from Sarah, the founder of Letta. The original version categorized…
The revolution in Large Language Models (LLMs) is undeniable, but their stateless nature remains a significant hurdle. Like Dory from Finding Nemo, they…
You’ve seen the hype: GPT-4o hits 128k tokens! Claude 3 digests novels! Sounds like AI memory’s cracked, time to pack up?
A couple of days ago, YouTube shoved this video in my face:
If you’ve been building AI applications, you’ve probably noticed something frustrating — LLMs have terrible memory. Sure, ChatGPT and Claude seem to…