There was just a subquadratic model (SubQ)released that performs about as good as the last round of frontier models, with a 12 million token window, but with 3 orders of magnitude less resource usage. Google also has their compressed models which greatly reduce processing. I think at some point the stuff we use frontier models for today will be running locally on our PC's and all this buildout will have been for naught or for people who have such good business cases that paying top dollar for a better model makes sense.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: