20
you are viewing a single comment's thread
view the rest of the comments
[-] chesmotorcycle@lemmygrad.ml 10 points 1 week ago

Still underestimating China at their own peril:

Analysts were not expecting China to produce a model as powerful as Fable until early next year.

[-] yogthos@lemmygrad.ml 7 points 1 week ago

They just keep getting high on their own supply. While American companies are betting on the idea that if one model can pull away it's going to keep self improving and nobody will catch up, Chinese companies are betting there will be a plateau to this tech, and it's more important to focus on market dominance because they will catch up later.

And it's becoming clear that Chinese companies made the right bet because Chinese models are closing the gap now, which means we are getting to the point of diminishing returns already where it takes ever more effort into squeezing just a bit more capability on the frontier.

[-] chesmotorcycle@lemmygrad.ml 9 points 1 week ago

Agreed. And just to add to that, it's primarily the Chinese companies doing both the basic research and the clever optimization that will continue to make meaningful progress. Hyperscaling has real, physical limits after all. Lose-lose for the Yanks.

[-] yogthos@lemmygrad.ml 8 points 1 week ago

Yup, and also notable how China proceeded to invest in energy infrastructure before trying to scale out compute. The Americans put the cart before the horse here.

[-] chesmotorcycle@lemmygrad.ml 3 points 1 week ago

The Magic of the Market^TM^

[-] Orcinus@lemmygrad.ml 3 points 1 week ago

I don't know how true this article is but it claims Deepseek is gunning for AGI, which sounds like the opposite of a plateau. China's putting AI in charge of R&D for chrissakes, and with a lot of the post-silicon hardware seen in their labs as well as 6G we may yet see colossal jumps in capability.

[-] yogthos@lemmygrad.ml 7 points 1 week ago

I don't think there's anything that fundamentally prevents AGI from being developed, but I strongly suspect that it's going to take more than LLMs to do that. In fact, frontier is already moving towards world models where a physics simulation becomes the basis for reasoning. Deepseek gunning for AGI is not in any way at odds with the current methods hitting a plateau. We already know that simply making models bigger doesn't scale. It's also unlikely that feeding more data into models is going to help either. The problem is with how the architecture works, and making any substantial leaps from here will require rethinking that.

And putting AI in charge of R&D is tangential to improvement of models themselves. Current capabilities are already very useful, and when you put a model in a feedback loop, it can indeed work autonomously and produce useful results. So, current crop of models will absolutely lead to a ton of breakthroughs and new types of automation that weren't possible before. China is also very well positioned to take advantage of all that being the factory of the world.

The question is what it will take to make models themselves significantly more capable though. Look at the progress that was being made like two to three years ago. Every few months we'd see some incredible breakthrough in capability. That's basically stopped in the past year, and progress has become very incremental now.

this post was submitted on 17 Jul 2026
20 points (100.0% liked)

Technology

1456 readers
27 users here now

A tech news sub for communists

founded 4 years ago
MODERATORS