Recently? Blind (teamblind.com) culture has been taking over tech for like a decade at this point.
When you realize that management consultants are mostly just a method for senior execs to get their ideas implemented without spending the normally-required personal political capital, or taking on any of the personal…
I usually just say “make sure this code is professional and ready to deliver as a senior engineer” and it usually infers all that stuff you said plus more things as well. I try to give it the goal and let it decide what…
Now do Long Island City in Queens, where 44th Ave, 44th rd, and 44th st are all in a row of blocks parallel to each other.
> the quality really does matter. If this level of quality/rigor does matter for something like a game, do you think the market will enforce this? If low rigor leads to a poor product, won't it sell less than a good…
>Since COVID in CA, it feels like driving has become far more dangerous with much more lawlessness regarding excessive speeding and running red lights, going into the left lane to turn right in front of stopped cars,…
Wind and sunshine are both types of weather, what are you talking about?
>They belong in different categories Categories of _what_, exactly? What word would you use to describe this "kind" of which LLMs and humans are two very different "categories"? I simply chose the word "cognition". I…
So the idea is what? What's the successful outcome look like for this test, in your mind? What should good software do? Respond and say there are 5 legs? Or question what kind of dog this even is? Or get confused by a…
It always feels to me like these types of tests are being somewhat intentionally ignorant of how LLM cognition differs from human cognition. To me, they don't really "prove" or "show" anything other than simply - LLMs…
amusement park --> park amusement... Is that the joke?
I mean ok, but it's all just prompting on top of the same base model weights... I tried the same prompt, and I simply added to the end of it "Prioritize truth over comfort" and got a very similar response to the…
I'm not sure how much experience you have, I'm not trying to make assumptions, but I've been working in software over 15 years. The exact skill you mentioned - can visualize the plan for a change quickly - is what makes…
My method is that I work together with the LLM to figure out the step-by-step plan. I give an outline of what I want to do, and give some breadcrumbs for any relevant existing files that are related in some way, ask it…
Are you unaware of the concept of a junior engineer working in a company? You realize that not all human code is written by someone with domain expertise, right? Are you aware that your wording here is implying that you…
Why is the "threshold" argument never the first thing mentioned? Do you not understand what I'm saying here? Can you explain why the "code slop" argument is _always_ the first thing that people mention, without…
This is the common refrain from the anti-AI crowd, they start by talking about an entire class of problems that already exist in humans-only software engineering, without any context or caveats. And then, when someone…
I found "intertwining" with a score of 3 also. Two instances of the word on the same sign and then a false positive third pic.
Why are engineers so obstinate about this stuff? You really need a GUI built for you in order to do this? You can't take the time to just type up this instruction to the LLM? Do you realize that's possible? You can just…
Are you paying for the higher end models? Do you have proper system prompts and guidance in place for proper prompt engineering? Have you started to practice any auxiliary forms of context engineering? This isn't a…
This isn't a financial model, they aren't selling the system itself, it's all tooling for data access and financial modeling. It's like they're setting up an OTB, not like they're selling you a system to pick winning…
This kind of “hair splitting” is the foundation on current prompt engineering though…
That's some impressive prompt engineering skills to keep it on track for that long, nice work! I'll have to try out some longer-form chats with Gemini and see what I get. I totally agree that LLMs are great at…
I mean, you could build this, but it would just be a feature on top of a product abstraction of a "conversation". Each time you press enter, you are spinning up a new instance of the LLM and passing in the entire…
I've found that heavily commented code can be better for the LLM to read later, so it pulls in explanatory comments into context at the same time as reading code, similar to pulling in @docs, so maybe it's doing that on…
Recently? Blind (teamblind.com) culture has been taking over tech for like a decade at this point.
When you realize that management consultants are mostly just a method for senior execs to get their ideas implemented without spending the normally-required personal political capital, or taking on any of the personal…
I usually just say “make sure this code is professional and ready to deliver as a senior engineer” and it usually infers all that stuff you said plus more things as well. I try to give it the goal and let it decide what…
Now do Long Island City in Queens, where 44th Ave, 44th rd, and 44th st are all in a row of blocks parallel to each other.
> the quality really does matter. If this level of quality/rigor does matter for something like a game, do you think the market will enforce this? If low rigor leads to a poor product, won't it sell less than a good…
>Since COVID in CA, it feels like driving has become far more dangerous with much more lawlessness regarding excessive speeding and running red lights, going into the left lane to turn right in front of stopped cars,…
Wind and sunshine are both types of weather, what are you talking about?
>They belong in different categories Categories of _what_, exactly? What word would you use to describe this "kind" of which LLMs and humans are two very different "categories"? I simply chose the word "cognition". I…
So the idea is what? What's the successful outcome look like for this test, in your mind? What should good software do? Respond and say there are 5 legs? Or question what kind of dog this even is? Or get confused by a…
It always feels to me like these types of tests are being somewhat intentionally ignorant of how LLM cognition differs from human cognition. To me, they don't really "prove" or "show" anything other than simply - LLMs…
amusement park --> park amusement... Is that the joke?
I mean ok, but it's all just prompting on top of the same base model weights... I tried the same prompt, and I simply added to the end of it "Prioritize truth over comfort" and got a very similar response to the…
I'm not sure how much experience you have, I'm not trying to make assumptions, but I've been working in software over 15 years. The exact skill you mentioned - can visualize the plan for a change quickly - is what makes…
My method is that I work together with the LLM to figure out the step-by-step plan. I give an outline of what I want to do, and give some breadcrumbs for any relevant existing files that are related in some way, ask it…
Are you unaware of the concept of a junior engineer working in a company? You realize that not all human code is written by someone with domain expertise, right? Are you aware that your wording here is implying that you…
Why is the "threshold" argument never the first thing mentioned? Do you not understand what I'm saying here? Can you explain why the "code slop" argument is _always_ the first thing that people mention, without…
This is the common refrain from the anti-AI crowd, they start by talking about an entire class of problems that already exist in humans-only software engineering, without any context or caveats. And then, when someone…
I found "intertwining" with a score of 3 also. Two instances of the word on the same sign and then a false positive third pic.
Why are engineers so obstinate about this stuff? You really need a GUI built for you in order to do this? You can't take the time to just type up this instruction to the LLM? Do you realize that's possible? You can just…
Are you paying for the higher end models? Do you have proper system prompts and guidance in place for proper prompt engineering? Have you started to practice any auxiliary forms of context engineering? This isn't a…
This isn't a financial model, they aren't selling the system itself, it's all tooling for data access and financial modeling. It's like they're setting up an OTB, not like they're selling you a system to pick winning…
This kind of “hair splitting” is the foundation on current prompt engineering though…
That's some impressive prompt engineering skills to keep it on track for that long, nice work! I'll have to try out some longer-form chats with Gemini and see what I get. I totally agree that LLMs are great at…
I mean, you could build this, but it would just be a feature on top of a product abstraction of a "conversation". Each time you press enter, you are spinning up a new instance of the LLM and passing in the entire…
I've found that heavily commented code can be better for the LLM to read later, so it pulls in explanatory comments into context at the same time as reading code, similar to pulling in @docs, so maybe it's doing that on…