I locally host gemma 4 e2b on a very modest spec laptop to help summarize large texts, as a rudimentary translation tool, and just as a toy. It makes sense. I see small coding models on hugging face all the time, I think we've hit a point where it's just too efficient not to atleast entertain a little bit of slop in your workflow.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments