Oh man you're underselling it to the rest of the website haha.
But it's tough to understate just how fast this is without seeing it. 15,749 tokens/s is what I get, and most responses from the big models might be a bit over 1000 tokens, maybe 2000 if they're stretching it (including chain of thought). Longest I got deepseek to go recently was just a bit below 5000 tokens.
But at such speeds, all of these generation lengths - 1000, 2000, 5000 - are basically done in the blink of an eye. 5000 tokens will be written in a third of a second, or just slightly above the average reaction time.
Unfortunately Jimmy seems limited to ~1000 tokens generation so we won't be able to really push it to the limits lol.