Technically, the capability is really simple, it could take them less than a day to implement. But, I think the reason he gave the "1 year" timeline is because that would greatly increase compute requirements, tools take more input tokens, and you have more requests because of them as well because it's basically prompt -> tool (start timer) -> response -> prompt -> tool (end timer) -> response whereas it's only two requests for it to hallucinate something. They're already struggling with keeping up with compute without adding tools into the mix
post
replies:
all 5 comments