FWIW I believe we can hit AGI! but I think at this point it’s clear that benchmarks are ~meaningless. LLMs are spiky / alien intelligences which don’t map to our own expectations; the existence of a benchmark creates a…
but managing humans is simpler in many ways, too: there's accountability for one. and it's much easier to reason about the capability boundary of a human.
FWIW I believe we can hit AGI! but I think at this point it’s clear that benchmarks are ~meaningless. LLMs are spiky / alien intelligences which don’t map to our own expectations; the existence of a benchmark creates a…
but managing humans is simpler in many ways, too: there's accountability for one. and it's much easier to reason about the capability boundary of a human.