I'm not sure if I understand you, it sounds like you think running a small llm locally on your computer will suddenly make it use like 10x more power. That's not how it works. It's the servers used to run the full sized models that use that much power, as each one has tens of thousands of processors running at once. And local llms do have usage, especially for accessability. I use a local llm for my home assistant instance so I can use voice commands, which is very helpful as a disabled person.
