Very serious. Your personal amount of usage means nothing at all in this conversation. It is entirely about tokens per watt. The amount of energy the memory operations involve scale incredibly well when people are accessing the same object in memory simultaneously. Last I looked it was around a 10x difference for the same models efficiency.
If you want me to be your personal search engine you’ll need to wait a bit, im making dinner right now and would rather look for the articles on my desktop.