You can do better with a local model on a consumer video card, I can run ollama on my own busted laptop chip with 2gb vram.
Shout out to the Ollama devs, and llama.cpp devs. Absolute wizards.
You can do better with a local model on a consumer video card, I can run ollama on my own busted laptop chip with 2gb vram.
Shout out to the Ollama devs, and llama.cpp devs. Absolute wizards.