I've been playing with llama.cpp a bit for the last week and it's surprisingly workable on a recent laptop just using the CPU. It's not really hard to imagine Apple and others adding (more) AI accelerators on mobile.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: