For which you still need massive amounts of memory and compute to run reliably
2026's average gaming PC is massive amounts of memory and compute apparently
The gap will take decades to close, if it ever does.
lol there are plenty of open source models in the top 100 with multiple SOTA models released in the last few months alone
There's also smaller LLM's being made like https://eurollm.io/ which excel in their own ways
That, and the fact that chatbots and agents nowadays rely on all sorts of proprietary customizations
Funny that just came up: https://discourse.ubuntu.com/t/the-future-of-ai-in-ubuntu/81130?=0
Previously, to benefit from the full power of LLMs, you had to skew to higher parameter models. Recent developments in models like Gemma 4 and Qwen-3.6-35B-A3B demonstrate advanced capabilities such as tool-calling which enable LLMs to search the web, interact with external APIs and file systems, troubleshoot live systems and fundamentally reason about topics that lie outside of their initial training data.
The gap will take decades to close, if it ever does.
😁