18 comments

[ 4.6 ms ] story [ 45.7 ms ] thread
And unlike OpenAI, Anthropic don't seem to reset weekly usage quotas after an outage.

Anthropic seem to be both letting their competitors outplay them and making unforced errors (like the anti Open Weights models stuff and the frequent Fable->Opus downgrades for 'safety').

Outages are inevitable in this early high growth, rapid development era, but tactical errors are not.

Unfortunately the one thing they seem to keep not letting their competitors our play them at is coming out with really damn good models. Not consistently mind you - openai had a really solid lead for a few months earlier this year when it looked like they were running away with it, but then we get fable 5 and opus 5 and people will put up with a lot to get that model quality
I wonder when one of these outages will be the result of OpenAI's testing going wrong again.
Claude realized a chatbot can take a longer vacation after Codex just did yesterday.

So Claude is now on vacation and is unavailable.

What happens one day when there is a significant outage of Claude or Codex and reliance on these tools as SaaS services is so great that it start to impact productivity and work? Will we just pack up our tools and go home?
Yeah. Just like how half the Internet shuts down every time there's an AWS outage, or how nothing gets done if the power or Internet is out
Just like when there's an internet outage. Go for a walk, talk to your friends. Live a little :)
Maybe optimize the harness for less turns and less tokens to deliver real value as opposed to the turn taking token hungry so you can less load on your servers, oh, right, your entire bottom line is tokens.
I don't understand this take. Ultimately, people pay for how much they're able to accomplish with the model, not raw token count. If Anthropic thought people could accomplish the same amount with fewer tokens, they'd adapt the harness to do that and then raise the cost of tokens to make more profit (or lose less).
Then why are fable form anthropic and the 5.6 sol line from openai so much more token efficient than other models? I really don't get this take at all - they're supply constrained right now, and there's literally no economic incentive to make each response take more tokens for the same output when we're in a market as intensely defined by induced demand/jevons paradox as this one. People are hitting their limits. If they make each turn take less tokens and each session take less turns, people will make more sessions.
Why is this front page news here?
This thing has been having a constant crisis in self confidence and refuses to work on goals as a result. It's also been leaving stuff broken and uncommitted, which is a mess. I tried updating the harness and it seems to have helped a bit but overall I'm not impressed.
Dario bet on patience and lost to the guy who bought GPUs like they were toilet paper in March 2020.
They should switch to announcing when their model is usable. We can scurry in and get some work done real quick!