You picked a piss poor metric and you are getting piss poor results. You can absolutely train an LLM-based model to be superhuman at chess. We just don't care enough to do so. LLMs being even as good at implicitly…
If chess ELO is the best proxy for general intelligence, then surely, even Deep Blue was smarter than any human alive? You can also claim that the best proxy for general intelligence is being able to multiply large…
This. They've been saying "AI is an extremely powerful, extremely dangerous technology" back when actual AIs were image classifiers that would struggle to tell a cupcake apart from a dog. OpenAI was founded by people…
It's very easy to tune an LLM for a "default" like "be sycophantic" or "be contrarian". It's easy to instill a semi-rigid "response template" like "agree with most of whatever the user says, but find at least one thing…
The successes are in areas where you can get AI to execute an iterative loop autonomously. Try, fail, refine, retry - succeed eventually. Coding is, inherently, quite amenable to that. Is bioscience? They're certainly…
Xitter can't make a web UI that's actually usable to save their lives - so they're going after the one platform that does with lawfare. Despicable behavior.
The issue is, you kind of have to not get wrecked by the negative impacts to collect those positive impacts. Unfortunately, "AI that's competent at bioweapon R&D" might be an easier mark to hit than "AI that can put a…
Frankly, I would be more concerned if they didn't. Given what we know of how unaligned are they in practice? All the proto-Astra and Mythos incidents? Not hacking there would point towards pure benchmaxxing/eval gaming.
You picked a piss poor metric and you are getting piss poor results. You can absolutely train an LLM-based model to be superhuman at chess. We just don't care enough to do so. LLMs being even as good at implicitly…
If chess ELO is the best proxy for general intelligence, then surely, even Deep Blue was smarter than any human alive? You can also claim that the best proxy for general intelligence is being able to multiply large…
This. They've been saying "AI is an extremely powerful, extremely dangerous technology" back when actual AIs were image classifiers that would struggle to tell a cupcake apart from a dog. OpenAI was founded by people…
It's very easy to tune an LLM for a "default" like "be sycophantic" or "be contrarian". It's easy to instill a semi-rigid "response template" like "agree with most of whatever the user says, but find at least one thing…
The successes are in areas where you can get AI to execute an iterative loop autonomously. Try, fail, refine, retry - succeed eventually. Coding is, inherently, quite amenable to that. Is bioscience? They're certainly…
Xitter can't make a web UI that's actually usable to save their lives - so they're going after the one platform that does with lawfare. Despicable behavior.
The issue is, you kind of have to not get wrecked by the negative impacts to collect those positive impacts. Unfortunately, "AI that's competent at bioweapon R&D" might be an easier mark to hit than "AI that can put a…
Frankly, I would be more concerned if they didn't. Given what we know of how unaligned are they in practice? All the proto-Astra and Mythos incidents? Not hacking there would point towards pure benchmaxxing/eval gaming.