DeepSWE crowns GPT-5.5, and finds Claude Opus exploiting a benchmark loophole (venturebeat.com) 3 points by sonink 3mo ago ↗ HN
0 comments
[ 23.0 ms ] story [ 125 ms ] threadNo comments yet.