Live is short, required time to market is even shorter. Arbitrary constraints might work if you have nothing better to do, but if you have a goal, or specific ideas you want to train, you'll get more natural constraints.
Imagine agents thinking to themselves, "We are not alone"
This is good for satire, but the real solution is to use AI to become the CEO, or take over the scope of of product or decision-making roles. There are many business-savvy technical workers that get pigeonholed because…
This is consistent with his premise. See, the title says it: incentives are for losers. If you're already a winner you can escape from incentives. No contradiction.
They were famous for committing to buying compute with future money that many people thought they would go bankrupt. Anthropic was afraid of going bankrupt and didn't do the same.
Quantum mechanics is pretty fundamental but you wouldn't use it to describe a rock rolling down hill either.
A more concise explanation is that, computation is the transformation of information. Any transformation of information is computation, and is thus subject to some theory of computation. A physical process involves the…
That's true for Sony, but Nintendo doesn't sell the Switch at a loss.
"I know this" is different from "I know you know this", which is different from "You know I know this", which is still different from "you know I know you know this"
You're right that I jumped the gun and the analogy is not accurate. The point where they are similar is that you have the prefix as context as you try to type out the next word; it is a more deliberate form of reading,…
A bit tough to say this, but transformers are trained the same way.
>I'm comfortable calling MiniMax the more eager model in this set because that claim is backed by the artifacts, not by vibe. It repeatedly reached for locks, persistence, policy objects, fallback paths, decorators, and…
At this point letting an agent go like this is akin to not leashing your dog in public. It's not easy to draw an accurate line but probably there needs to be real punishment for doing these things.
I suspect this will be a significant problem blocking long-horizon tasks in practice, basically the more turns there are, the larger the chance the classifier produces a false positive. The disappointment of the user…
Or kids. Or work.
It's clear to me they are subsidizing inference in exchange for market share, and doing it at this scale makes the most sense if their target is getting more user data. Note that this sort of pricing isn't far off from…
The development and acquisition of valuable domain knowledge is a hard, risky, expensive and slow process. Because the valuable domain knowledge isn't yesterday's, it's today's and tomorrow's. In fields where domain…
This is a lot tamer than what Claude Code's team claims tbf.
Right, you could disagree on which things to prioritize over dollar profits. My main point is that these preferences are not irrational like was asserted. At the scale of a sovereign wealth fund or pensions, you need to…
An article with a title saying tokens per second throughput without any qualifier e.g. what size the model is should immediately be classified as spam.
Gemini 3.5 Flash is not good at coding in practice. Gemini 3.1 Pro too, in particular is known to be bad at tool calls. Many companies would love to have alternatives to Claude Code (as it's a significant risk to depend…
China will make sure they have a frontier lab, there's plenty of chance for Google to catch up once the compute crunch gets more serious.
Unsurprising he'd be cheered for saying what they wanted to hear. But perhaps whether or not his stance is correct, the students needed to hear this. They (we) have to believe human brains still have value and find a…
I'd love to go back to the 90s and live it again.
Right, indeed you need to first preserve the origin, but also that is trivially true for a linear map like JL.
Live is short, required time to market is even shorter. Arbitrary constraints might work if you have nothing better to do, but if you have a goal, or specific ideas you want to train, you'll get more natural constraints.
Imagine agents thinking to themselves, "We are not alone"
This is good for satire, but the real solution is to use AI to become the CEO, or take over the scope of of product or decision-making roles. There are many business-savvy technical workers that get pigeonholed because…
This is consistent with his premise. See, the title says it: incentives are for losers. If you're already a winner you can escape from incentives. No contradiction.
They were famous for committing to buying compute with future money that many people thought they would go bankrupt. Anthropic was afraid of going bankrupt and didn't do the same.
Quantum mechanics is pretty fundamental but you wouldn't use it to describe a rock rolling down hill either.
A more concise explanation is that, computation is the transformation of information. Any transformation of information is computation, and is thus subject to some theory of computation. A physical process involves the…
That's true for Sony, but Nintendo doesn't sell the Switch at a loss.
"I know this" is different from "I know you know this", which is different from "You know I know this", which is still different from "you know I know you know this"
You're right that I jumped the gun and the analogy is not accurate. The point where they are similar is that you have the prefix as context as you try to type out the next word; it is a more deliberate form of reading,…
A bit tough to say this, but transformers are trained the same way.
>I'm comfortable calling MiniMax the more eager model in this set because that claim is backed by the artifacts, not by vibe. It repeatedly reached for locks, persistence, policy objects, fallback paths, decorators, and…
At this point letting an agent go like this is akin to not leashing your dog in public. It's not easy to draw an accurate line but probably there needs to be real punishment for doing these things.
I suspect this will be a significant problem blocking long-horizon tasks in practice, basically the more turns there are, the larger the chance the classifier produces a false positive. The disappointment of the user…
Or kids. Or work.
It's clear to me they are subsidizing inference in exchange for market share, and doing it at this scale makes the most sense if their target is getting more user data. Note that this sort of pricing isn't far off from…
The development and acquisition of valuable domain knowledge is a hard, risky, expensive and slow process. Because the valuable domain knowledge isn't yesterday's, it's today's and tomorrow's. In fields where domain…
This is a lot tamer than what Claude Code's team claims tbf.
Right, you could disagree on which things to prioritize over dollar profits. My main point is that these preferences are not irrational like was asserted. At the scale of a sovereign wealth fund or pensions, you need to…
An article with a title saying tokens per second throughput without any qualifier e.g. what size the model is should immediately be classified as spam.
Gemini 3.5 Flash is not good at coding in practice. Gemini 3.1 Pro too, in particular is known to be bad at tool calls. Many companies would love to have alternatives to Claude Code (as it's a significant risk to depend…
China will make sure they have a frontier lab, there's plenty of chance for Google to catch up once the compute crunch gets more serious.
Unsurprising he'd be cheered for saying what they wanted to hear. But perhaps whether or not his stance is correct, the students needed to hear this. They (we) have to believe human brains still have value and find a…
I'd love to go back to the 90s and live it again.
Right, indeed you need to first preserve the origin, but also that is trivially true for a linear map like JL.