Sure you can run it on low end hardware, but how does the performance (response time for a given prompt) compare to the other models, either local or as a service?
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: