The fact that it can't tell the difference between a prompt and part of the data it is examining really kills your argument.
Also it's a word probability matrix, not actually reasoning or understanding. It looks at all the words it is fed, and comes up with other words that are most likely to be near those. That's why these tricks work. It injects noise that interferes with those probabilities