author here. the part i'd actually like discussion on is the buried finding: physicians+GPT-4 didn't outperform GPT-4 alone on the management cases, and on the landmark cases the model alone beat the model+physician. the paper reports it and moves on. that's the 2026 question, and it's the one a Science-level platform could have been used to ask
1 comment
[ 4.8 ms ] story [ 18.5 ms ] thread