Evaluating frontier AI R&D capabilities of LLM agents against human experts (metr.org) 1 points by tedsanders 1y ago ↗ HN
0 comments
[ 3.2 ms ] story [ 14.9 ms ] threadNo comments yet.