Despite the “official” coding score for GPT5 being higher, Claude sonnet still seems to blow it out of the water. That seems to suggest they are training to the test and the test must not be a very good test. Or they are lying.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
replies: