Gee Eddy_Viscosity2, how come your mom lets you have TWO big boxes of cables? I'm sure I have multiple VGA cables in mine, and I'm not sure a single device I own even has a VGA port on it anymore.
> Along the way, it wrote 13 million lines of Lean and proved 29,500 intermediate theorems. Pretty insane. I suppose it lends further credence to the idea that anything that can be shown to be correct can be done by a…
The review that "IBM Bob isn't just another autocomplete tool" tells you everything you need to know about this product. lol.
lore does not make it a good or professional choice
For one, I think AI music is ontologically evil and have extreme opinions on the matter. More than that though, I completely disagree that anything I'd ever choose to listen to is elevator tier / background noise. Just…
This was a very pleasant read. There's (justifiably) a lot of debate about AI prose on this forum, which has become tiring. It's refreshing to read something so grounded. Doubly so while looking at a packet of Belvita…
This is super cool! I have always loved projects that try to categorize color space. Usually my work has gone in the other direction - isGray, isCloseToWhite, isPastel - but it's the same insofar as it's defining shapes…
Very impressive headline benchmark numbers. I expected a step change, but not past Fable. That said - it all depends on whether the classifiers make the model unusable...
This is a good idea - and I might pick it up for my writing, since a lot of it still happens in Obsidian. If a document is worked on across multiple days / revisions / app sessions - how is that handled?
It seems that they're trying to push up-market, or at least they were. Given the extremely competitive releases of GLM 5.2 and DeepSeek V4 (both pro and flash), I don't think there'll be appetite for it.
It's a bit disheartening to see no comparison to other models here - and I'm not sure this pushes the curve anywhere. 3.6 flash is more expensive than GLM 5.2 - but seemingly worse, although this post is really light…
> There is no joy of discovery of new music that you haven't heard. There is no connection to other humans through music. Hard disagree here. For new music, Discover Weekly is great, if you take some time to engage with…
Yes - this is certainly true. I would have to concede that I could not speak to how my thesis would map onto a truly colossal codebase. That said - I would (again, maybe naively) suppose it's not hugely different - much…
Obviously I'm just one guy, but I maintain what I said w/ typing happily at 120+wpm. I think I could break 150 if I switched to colemak or dvorak, which I've been considering. I don't think my job was ever to type fast,…
Yes - I think that's fair. I'm happy to concede the term coding - I still enjoy the process of telling my computer to build something I've envisioned. It's now much faster, and I get to do more of the big-picture…
I guess I'm of two minds on it. Sometimes, I will review every line, test the front-end in a staging environment, verify the backend contract, et cetera. Over time, though, I realized that many of these reviews just…
I love coding and I always have - arguably moreso now, if you can still call it coding. For me, it is not about the syntax nor the mechanics of typing though. My enjoyment comes from thinking about problems and breaking…
Generally end-user applications, depending on how you classify internal tools. I'm generally very happy with the output of Opus 4.X with a moderately structured CLAUDE.md and some investment in detection/avoidance of…
I can only imagine that people who say things like: > If you use AI for anything else, and in particularly if you use it to generate code, you're wasting your time. Have not used frontier models in at least a year. It…
+1 for the claude icon fire bars
I would add a +1 for testing w/ Word - the official Office suite runs some validation where only Word will show a "broken file" popup, even when nothing else does. In our case, clients use only real Word, so any…
Interesting - I'll be sure to benchmark it at some point. We've found the best results come from blending providers depending on the task anyways. Thanks for the quick response - and always happy to see more competition…
Unclear what difference exists against Firecrawl - their team has been shipping great features extremely quickly lately, and their core offerings have become really good. I am interested in KnifeGeek though - looking…
Nor more than a mention of Query Hints, which had some interesting discussion under a similarly-titled submission. https://news.ycombinator.com/item?id=48413655
This seems to be a worse version of another submission [0] I saw a while back - binary octets are easy for anyone who can copy paste; image attributes like edge pressure and stable contour mean basically nothing to me.…
Gee Eddy_Viscosity2, how come your mom lets you have TWO big boxes of cables? I'm sure I have multiple VGA cables in mine, and I'm not sure a single device I own even has a VGA port on it anymore.
> Along the way, it wrote 13 million lines of Lean and proved 29,500 intermediate theorems. Pretty insane. I suppose it lends further credence to the idea that anything that can be shown to be correct can be done by a…
The review that "IBM Bob isn't just another autocomplete tool" tells you everything you need to know about this product. lol.
lore does not make it a good or professional choice
For one, I think AI music is ontologically evil and have extreme opinions on the matter. More than that though, I completely disagree that anything I'd ever choose to listen to is elevator tier / background noise. Just…
This was a very pleasant read. There's (justifiably) a lot of debate about AI prose on this forum, which has become tiring. It's refreshing to read something so grounded. Doubly so while looking at a packet of Belvita…
This is super cool! I have always loved projects that try to categorize color space. Usually my work has gone in the other direction - isGray, isCloseToWhite, isPastel - but it's the same insofar as it's defining shapes…
Very impressive headline benchmark numbers. I expected a step change, but not past Fable. That said - it all depends on whether the classifiers make the model unusable...
This is a good idea - and I might pick it up for my writing, since a lot of it still happens in Obsidian. If a document is worked on across multiple days / revisions / app sessions - how is that handled?
It seems that they're trying to push up-market, or at least they were. Given the extremely competitive releases of GLM 5.2 and DeepSeek V4 (both pro and flash), I don't think there'll be appetite for it.
It's a bit disheartening to see no comparison to other models here - and I'm not sure this pushes the curve anywhere. 3.6 flash is more expensive than GLM 5.2 - but seemingly worse, although this post is really light…
> There is no joy of discovery of new music that you haven't heard. There is no connection to other humans through music. Hard disagree here. For new music, Discover Weekly is great, if you take some time to engage with…
Yes - this is certainly true. I would have to concede that I could not speak to how my thesis would map onto a truly colossal codebase. That said - I would (again, maybe naively) suppose it's not hugely different - much…
Obviously I'm just one guy, but I maintain what I said w/ typing happily at 120+wpm. I think I could break 150 if I switched to colemak or dvorak, which I've been considering. I don't think my job was ever to type fast,…
Yes - I think that's fair. I'm happy to concede the term coding - I still enjoy the process of telling my computer to build something I've envisioned. It's now much faster, and I get to do more of the big-picture…
I guess I'm of two minds on it. Sometimes, I will review every line, test the front-end in a staging environment, verify the backend contract, et cetera. Over time, though, I realized that many of these reviews just…
I love coding and I always have - arguably moreso now, if you can still call it coding. For me, it is not about the syntax nor the mechanics of typing though. My enjoyment comes from thinking about problems and breaking…
Generally end-user applications, depending on how you classify internal tools. I'm generally very happy with the output of Opus 4.X with a moderately structured CLAUDE.md and some investment in detection/avoidance of…
I can only imagine that people who say things like: > If you use AI for anything else, and in particularly if you use it to generate code, you're wasting your time. Have not used frontier models in at least a year. It…
+1 for the claude icon fire bars
I would add a +1 for testing w/ Word - the official Office suite runs some validation where only Word will show a "broken file" popup, even when nothing else does. In our case, clients use only real Word, so any…
Interesting - I'll be sure to benchmark it at some point. We've found the best results come from blending providers depending on the task anyways. Thanks for the quick response - and always happy to see more competition…
Unclear what difference exists against Firecrawl - their team has been shipping great features extremely quickly lately, and their core offerings have become really good. I am interested in KnifeGeek though - looking…
Nor more than a mention of Query Hints, which had some interesting discussion under a similarly-titled submission. https://news.ycombinator.com/item?id=48413655
This seems to be a worse version of another submission [0] I saw a while back - binary octets are easy for anyone who can copy paste; image attributes like edge pressure and stable contour mean basically nothing to me.…