DeepSWE crowns GPT-5.5, and finds Claude Opus exploiting a benchmark loopholeventurebeat.com 3 pointssonink3 months agodiscussSaveHideCopy link On HNComments No comments yet.
Comments
No comments yet.