Your AI gets dumber
the longer you talk.
The fix isn't a smarter model. It's a fresh chat that already knows where you left off.
Every answer in a chat gets read against everything that came before it. The longer one chat runs, the more old back-and-forth it's wading through, and the answers go vague, repetitive, or just wrong. People blame the model. It's the chat.
The fix is boring and it works: start a fresh chat, and hand it a short brief so it picks up exactly where the old one died. This is the block I paste, every working session, every day. Fill the brackets, paste it as the first message, and you're back in business without scrolling through an hour of history.
The two-minute version
Here's the whole thing, start to finish. Then grab the block below.
The brief I paste
Five fields. Fill the brackets, paste it as the first message of a new chat, and the new session starts already caught up. That's the entire move.
Five steps, any time a chat goes stale
This is measured, not a hunch
Researchers gave it a name: context rot. Test the top AI models on a long chat versus a short one and accuracy falls off a cliff. Here's how far the big models dropped once the chat hit about 32,000 tokens, a long but normal working session:
| Model | Short chat | Long chat (~32K) |
|---|---|---|
| GPT-4o | 99.3% | 69.7% |
| Gemini 1.5 Pro | 92.6% | 48.2% |
| Llama 3.3 70B | 97.3% | 42.7% |
| Claude 3.5 Sonnet | 87.6% | 29.8% |
Accuracy on a long-context recall test, short chat vs ~32K tokens. 10 of 12 models tested fell to half their short-chat score or worse by 32K. Source: NoLiMa, Adobe Research.
Most of them started slipping way earlier than that, some by just 2,000 to 8,000 tokens. And it isn't only the older models above. Anthropic's own developer tool reports that even the newest one, Claude Opus 4.8 (a million-token window), starts slipping around 40% full and flags you to start a fresh session by about 48%. That's the exact line I draw: reset before you're 40 to 50% in.
Newer models raise the ceiling. They don't remove the floor. The fix isn't waiting for a smarter model. It's the fresh chat above.
Sources: NoLiMa (Adobe Research) · Chroma, "Context Rot" · Claude Code degradation report
When to reach for it
The answers start slipping
You're switching tasks in the same project
You walked away frustrated
Rule of thumb: serious work gets a fresh chat early and often. Don't wait until it's obviously broken.
Three things that make it work
Write it at the end of the old chat, not from memory
Paste real material, not descriptions
One brief, one goal
Want it on autopilot? Make it a one-command skill
Skip this if you're just getting started. The copy-paste block above is all you need. But once you've pasted it a few times, you can turn the whole move into a single command.
If you use Claude Code (the terminal version of Claude), a skill is a saved instruction Claude runs on demand. I built one called /handoff. At the end of a session I type it, and it reads the whole chat, fills the five fields for me, saves the brief, and hands me the block to paste into the next one. No copy-paste, no writing it from memory.
Get the skill-creator (it builds the skill for you)
Give it this job (paste this)
Two rules that make it good
Now closing a session is one command, and the next one starts already caught up. (No terminal? The copy-paste block above is your version. It works just as well.)
Before you ask
This is how I actually run every working session with AI.
I put out one of these a week. The boring, useful part of AI for people who run real businesses. No hype, no jargon.
The brief is the beginner move. The real thing is wiring your business so AI runs parts of it. That starts with a free call, no pitch.