Exactly, that's why I gave it the "No, but actually yes" preface.
I mean, not quite, but also yes.
Running your own model doesn't solve every problem with LLMs, but it sure as hell does solve a ton of them, for way less money! Local models are great, gwen3.8 on my 3090 at home is slower, but often better than the pay to play Claude from work.
Oh I'm not one of those "LLMs are good for nothing" people, I use them often at work and at home and they do have some real world use cases that are very compelling. However they are a far very from the "solve everything machine" they are being pitched as, and they do seem to be plateauing in their improvements, seems like there's not much further they are going to be able to go even with a billion people using them everyday, there just isn't any new training data that is actually useful and they don't make new data, only associations between data they have already ingested. My qwen 3.8 model running on my 3090 at home is often better than pay for play enterprise Claude at work, if not just slower. I expect local models will explode in the future due to the cost / benefit of running them.
While this is a legitimate concern for sometime in perhaps the next hundred years, our current path with LLMs is never going to produce super intelligent AGI.
They also shot a missile at a satellite and added a ton of high velocity junk into orbit just a few years ago, so not exactly the patron saints of protecting space over there either.
You are gone for six months, what are they going to do? Leave the dog to starve to death? The local pound or some other animal rescue organization picked them up and (hopefully) re-homed them (or otherwise put them down)
It was a Redhat project, which is now wholly owned by IBM.
Now I'M imagining a version of sub rosa in which the green mist turns into Kermit the frog and saranades her with a rendition of "It's not easy being green"
Check out Babylon 5, telepaths are a major part of the story for so many reasons.
He's just in it for the love of the grift.
I plan to once there is a version with turboquant and MTP as that huge context window is key.