Huh, what exactly are you using? I tried to run home assistant's, well, assistant, and qwen 2.5 3B just didn't properly call any tools. And bigger models become pretty slow on my ryzen 2400 "thinclient" home server. I've tried it via Ollama, even did some dirty hacks to use the iGPU, but that's about the size that I can run without waiting a minute for a simple prompt.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: