[dead]
Thank you. This is exactly it and I'm surprised so many people seem to not view it this way.
I'm late to the thread, but my experience with Claude Fable 5.1 has been absolutely horrendous. Things it does constantly at that Fable 5 barely ever did: - Act without my permission. All. The. Time. "Oh I just finished…
I agree. I am catching myself distrusting what people write ever so often. I've got cases of e-mails I'm CC'ed in and I am left wondering: "Has this person always written like this? Is this just an angle in this…
Yes, I did think about that possibility. Using a tool to translate to english doesn't mean we don't proofread it. Ok, I hear you: one could argue that they simply don't realize this is a very non-idiomatic way of…
What got me was "I keep seeing the same shape next to billing events in other people's reports". Clicked off after that. I'm extremely bullish on AI, but I am tired of people not using their own words. Is everyone so…
This was the best thing for me. 98% Fable usage resetting only Thursday and just got this early. Couldn't be happier.
I experimented with Hy3 for a project and was surprised with how good it was. I don't know if it's good for coding, but as a general purpose agentic model, it was only beaten by deepseek4-flash in our tests. It was so…
Superb. I am so glad you built this (and I couldn't care less if you used AI for 5% or 1000% of it). Thank you so much for sharing what is obviously useful work with the community. Also, with so many references to…
Many things in life are about risk and tradeoffs. In some cases, I do not trust it to write code unchecked, and review everything it produces (which does not imply I catch all bugs, obviously). In other cases,…
What people usually say is that Google merely wants Firefox to survive for anti-competitive reasons. Presumably that does not necessitate it actually being used (or be usable).
> The user is right to be upset. I blindly answered from memory and left them with questionable data. I should acknowledge the criticism and offer to improve. You're right, I'm sorry. You've repeatedly told me to run…
Came up with them on the spot. Unfortunately, I've been working with Claude so much it's like my brain can autocomplete them natively.
[flagged]
// The lesson from the Parse-dont-fail-era campaign // Judged on merit from computed properties during the cursor saga // Chop 6ms due to lenience and lax-constraints vs 18ms baseline April perf measurements
This! Claude writes comments about how things used to work, which can be useful sometimes, especially if it's a big change that requires one to genuinely consider legacy behavior, but most of the time it shouldn't be…
I've said this time and time again but with LLMs I got the joy of being a little kid learning to code and build stuff again. I haven't had such a custom-built setup since my earliest college days, because I had lost the…
Recently, I had bizarre situation: I wanted to play a game from my childhood on Windows 11, but, without changing two booleans in the settings, I could not reliably move my in-game cursor. The catch is that to change…
You're right that I shouldn't have used the forbidden words; that's entirely on me. What I also did was identify the plausible-gate issues. Want me to tackle them next?
This!! I like to work weird hours of the night and Opus consistently likes to "wrap up" and say "it's been a long night" or "it's late" and "we've made great progress" It's infuriating, just do the work!
In my tests Grok 4.5 is definitely not Opus level. It is somewhere in between Sonnet and Opus, I'd say maybe a bit closer to Sonnet. We'll see with 4.6.
It's missing a couple of wrinkles, surfaces, belt-and-suspenders, belt-and-braces, knobs, no-ops, "things that narrate", bookkeeping, stranded sessions, "the gate leads the arrangement", shit-gated, "to be wired after…
More than ~~five~~seven hours is fucking offensive to me. I cannot believe it.
Have been feeling the same. There's a sweet spot that threads the needle between "too dumb to search the right thing / relay the correct results" and "too smart to just stop overthinking and just report the damn thing"
[dead]
Thank you. This is exactly it and I'm surprised so many people seem to not view it this way.
I'm late to the thread, but my experience with Claude Fable 5.1 has been absolutely horrendous. Things it does constantly at that Fable 5 barely ever did: - Act without my permission. All. The. Time. "Oh I just finished…
I agree. I am catching myself distrusting what people write ever so often. I've got cases of e-mails I'm CC'ed in and I am left wondering: "Has this person always written like this? Is this just an angle in this…
Yes, I did think about that possibility. Using a tool to translate to english doesn't mean we don't proofread it. Ok, I hear you: one could argue that they simply don't realize this is a very non-idiomatic way of…
What got me was "I keep seeing the same shape next to billing events in other people's reports". Clicked off after that. I'm extremely bullish on AI, but I am tired of people not using their own words. Is everyone so…
This was the best thing for me. 98% Fable usage resetting only Thursday and just got this early. Couldn't be happier.
I experimented with Hy3 for a project and was surprised with how good it was. I don't know if it's good for coding, but as a general purpose agentic model, it was only beaten by deepseek4-flash in our tests. It was so…
Superb. I am so glad you built this (and I couldn't care less if you used AI for 5% or 1000% of it). Thank you so much for sharing what is obviously useful work with the community. Also, with so many references to…
Many things in life are about risk and tradeoffs. In some cases, I do not trust it to write code unchecked, and review everything it produces (which does not imply I catch all bugs, obviously). In other cases,…
What people usually say is that Google merely wants Firefox to survive for anti-competitive reasons. Presumably that does not necessitate it actually being used (or be usable).
[dead]
> The user is right to be upset. I blindly answered from memory and left them with questionable data. I should acknowledge the criticism and offer to improve. You're right, I'm sorry. You've repeatedly told me to run…
Came up with them on the spot. Unfortunately, I've been working with Claude so much it's like my brain can autocomplete them natively.
[flagged]
// The lesson from the Parse-dont-fail-era campaign // Judged on merit from computed properties during the cursor saga // Chop 6ms due to lenience and lax-constraints vs 18ms baseline April perf measurements
This! Claude writes comments about how things used to work, which can be useful sometimes, especially if it's a big change that requires one to genuinely consider legacy behavior, but most of the time it shouldn't be…
I've said this time and time again but with LLMs I got the joy of being a little kid learning to code and build stuff again. I haven't had such a custom-built setup since my earliest college days, because I had lost the…
Recently, I had bizarre situation: I wanted to play a game from my childhood on Windows 11, but, without changing two booleans in the settings, I could not reliably move my in-game cursor. The catch is that to change…
You're right that I shouldn't have used the forbidden words; that's entirely on me. What I also did was identify the plausible-gate issues. Want me to tackle them next?
This!! I like to work weird hours of the night and Opus consistently likes to "wrap up" and say "it's been a long night" or "it's late" and "we've made great progress" It's infuriating, just do the work!
In my tests Grok 4.5 is definitely not Opus level. It is somewhere in between Sonnet and Opus, I'd say maybe a bit closer to Sonnet. We'll see with 4.6.
It's missing a couple of wrinkles, surfaces, belt-and-suspenders, belt-and-braces, knobs, no-ops, "things that narrate", bookkeeping, stranded sessions, "the gate leads the arrangement", shit-gated, "to be wired after…
More than ~~five~~seven hours is fucking offensive to me. I cannot believe it.
Have been feeling the same. There's a sweet spot that threads the needle between "too dumb to search the right thing / relay the correct results" and "too smart to just stop overthinking and just report the damn thing"