Not a lot that can. LLM proof fonts are spotty at best, I moved my hands up an to the right 3 keys and kept typing and i may as well of changed my font as obfuscated my message. I have an idea about negative space...…
First off, everyone here knows that a cybersecurity vulnerability audit is a hack request, and you don't hack without rolling up your sleeves and digging through the trash first, making a few phone calls, maybe take a…
these can work as long as the evidence is attached (to your point). And in healthcare absolutely must (your EMR example is terrifying). The patient may be allergic to nuts (confidence .84) >> cite: 1) lab value X, PRO…
I'm working on a universal one. My belief is that as agentic capabilities increase, behavioral attribution will be increasingly needed to maintain quality.
Agreed. However my research is showing that saturation of the rubric occurs in unique configuration that drive the transfer curve to near binary. Effectively turning a 1-10 rating system into yes/no.
[dead]
Has anyone actually used Grok to code? How does he do?
This is exactly right. Abstracted out of the process, or to a point of most optimal application.
I find some of the most interesting, and catastrophic failures in my agent fine-tuning come from the clamping down of non-determinism. It is totally the correct approach, but must be handled delicately. The…
I think stochastic parroting is really a very accurate description of what they do (if underserving of the overall usefulness of LLMs). As long as you consider they are parroting from the whole of human intelligence.…
[flagged]
Not a lot that can. LLM proof fonts are spotty at best, I moved my hands up an to the right 3 keys and kept typing and i may as well of changed my font as obfuscated my message. I have an idea about negative space...…
First off, everyone here knows that a cybersecurity vulnerability audit is a hack request, and you don't hack without rolling up your sleeves and digging through the trash first, making a few phone calls, maybe take a…
these can work as long as the evidence is attached (to your point). And in healthcare absolutely must (your EMR example is terrifying). The patient may be allergic to nuts (confidence .84) >> cite: 1) lab value X, PRO…
I'm working on a universal one. My belief is that as agentic capabilities increase, behavioral attribution will be increasingly needed to maintain quality.
Agreed. However my research is showing that saturation of the rubric occurs in unique configuration that drive the transfer curve to near binary. Effectively turning a 1-10 rating system into yes/no.
[dead]
Has anyone actually used Grok to code? How does he do?
This is exactly right. Abstracted out of the process, or to a point of most optimal application.
I find some of the most interesting, and catastrophic failures in my agent fine-tuning come from the clamping down of non-determinism. It is totally the correct approach, but must be handled delicately. The…
I think stochastic parroting is really a very accurate description of what they do (if underserving of the overall usefulness of LLMs). As long as you consider they are parroting from the whole of human intelligence.…
[flagged]