agreed. I'd also love to see more comparisons like this. I do think that a well speced out prompt can be finished faster by an agent then if prompted lazily. I've also found measuring this to be challenging, since…
also, this assumes humans are still the primary CLI consumers. With agents increasingly being the first-class users of command-line tools, building visual design tooling for terminal UIs feels like optimizing for a…
I get that "communicate your thought process" or "play along with the exercise" gets offered as the fix here. But that framing bothers me too. Why should simplicity require more justification than complexity? Google…
+ Also the fact that the Memory.md file was a hindrance to the quality of output
There are so many SOTA memory claims out there. There are also very few memory benchmarks (if any) that measure "memory" realistically. Not to mention the industry can't seem to decide on what "memory" even means. We…
agreed. I'd also love to see more comparisons like this. I do think that a well speced out prompt can be finished faster by an agent then if prompted lazily. I've also found measuring this to be challenging, since…
also, this assumes humans are still the primary CLI consumers. With agents increasingly being the first-class users of command-line tools, building visual design tooling for terminal UIs feels like optimizing for a…
I get that "communicate your thought process" or "play along with the exercise" gets offered as the fix here. But that framing bothers me too. Why should simplicity require more justification than complexity? Google…
+ Also the fact that the Memory.md file was a hindrance to the quality of output
There are so many SOTA memory claims out there. There are also very few memory benchmarks (if any) that measure "memory" realistically. Not to mention the industry can't seem to decide on what "memory" even means. We…