12 comments

[ 406 ms ] story [ 196 ms ] thread
Hi folks, author here and also author of the seminal OpenAI blog on this topic. Let me know how I can help you all let it rip.
The mother of all prompt injections.

Do others at OpenAI use this?

What percentage of this was written with AI?

> A command can feel ambient without making its credential model-visible.

I'd be wary of following guidelines on how to build human-computer interaction systems that weren't written with full human oversight. This kind of wording makes me wonder if these are actual recommendations or just the AI pattern-matching and hallucinating something that looks coherent.

(comment deleted)
I don't even know how to feel about this after glancing through the repo.
Serious question: how do we verify claims like these on the effectiveness of a harness?
Prompt engineering, context engineering, now harness engineering. I guess in a few months we'll have looped back to software engineering ?

All these buzzwords for basically what's glorified tweaking and configuration. All these years I should have mentioned I was doing Debugger Engineering and IDE Engineering