It can as long as there's a shared prefix.
I've never heard of that. The cache is meant to support token generation, so unless it's trying to repeat itself, I think the previous cache would be little use.
An experience, by definition, must change your behavior.
Ok, yes, that is my argument, now how the fuck does a temporary cache, thrown out after every response, prove that the model is changing as a result of undergoing inference?