Sure, but you're running a very small model compared to what we are talking about.
GLM-5.1 is over 200GB even when quantizied to 1-bit. Kimi K2.6 is even bigger. A framework desktop cannot run either of these. Qwen3.6 is significantly smaller and the model weights could fit, but consider the KV-cache you'd need for all of the company's users, and the throughput required to serve them all.
You're right that it is within reach for a company but framework desktop makes zero sense for this