I don't get this beating opus, It just hardcoded the tasks for bench , It does even respond normally A alot randomness in it Please don't hype
Only model which do less hallucinations , they is not AI winning right now
a kind of
I don't get this beating opus, It just hardcoded the tasks for bench , It does even respond normally A alot randomness in it Please don't hype
Only model which do less hallucinations , they is not AI winning right now
a kind of