I use Claude on the $20/mo plan, mainly for research, "rubber duckying," and as an idea soundboard (I was using Fable for this but am now relegated to Opus 4.8 or Sonnet 5), and recently noticed a severe downturn in Opus 4.8 response quality in high-context situations. To attempt to remedy this, I disabled the memory system (now labeled "Legacy" even though to my knowledge there is no replacement).With this setting disabled, aside from the slight tedium of inputting the necessary context each time for task continuations, I have noticed a drastic improvement in response quality and accuracy—I'm one of those shmucks who whined a bit after Opus 4.7 came out, and my theory is some change was made to increase memory context that resulted in a decrease in actual response quality. In general, I wonder if the high-context abilities of SOTA thinking models are overstated.As a brief anecdote, I was working through a research task with upwards of 30 prompts. Switching to a new chat and condensing...
Want to discover more AI signals like this?
Explore Steek