An AI slop to filter out AI slops -- the pinnacle of AI. Sadly this would not end just as a joke. We'll be doing this forever from now on.
> widen your reading, rather than trying to win a silly argument on a forum with 'logical fallacies' and other such humdrum wikiphilosophy. Yeah, sure, the paragraph there is pretty bad, I admit. I was having an…
> we are all way worse at reading compared to the average mid-20th centry reader? I would say English has changed. Modern English speakers, especially those in academia and technology, prefers explicit styles over…
No, my view is more like that he was too good at English to write bad English that he was criticizing. The examples he wrote are simply too sharp to fit well into his own criticisms. They are formulated so well that you…
Ah, I just read through it. That’s was a fun read. > This is a parody, but not a very gross one. … So, Orwell wrote the second sentence as a parody of the first sentence, using the modern English that he was criticizing…
I think Orwell did a great job there actually. Despite its bizarre look, the sentence is evocative and eloquent. It does make me get a clear mental image from the very first word. It leaves little room for roaming and…
Yeah, even Pangram is a bit problematic here. It has been notoriously fragile. Minor edits can flip scores from 100%-human to 100%-AI, because Pangram is crazy sensitive to local and surfacial features of text. Simple…
Obviously, NVIDIA is trying to own the AI development chain. Owning HF -- the discovery and distribution channel -- is one thing, but I think the biggest threat vector is the privileged access to HF platform data, that…
It doesn't answer at all. :/
One obvious loophole: upload your full outputs publicly. Anyone can feed those to their models. Wait, we're already doing it. /s
Yes, but the token expires upon logout, and, after reading your comment, I just double-checked the expiry (got 400 from all related endpoints).
You simply don't have enough testcases.
I've just let Claude analyze an HAR file containing XHR traffic on the Facebook timeline. It happily wrote a script for extracting contents, fetch images and videos, and rebuilt the timeline into a single HTML -- highly…
Lisp is a highly expressive language with little constraint, so there’s plenty of room for individuals to become strongly opinionated. That leads to a bit of elite mindset.
LLMs are only good at copying the surface, like frequently used words, sentence structures, use of metaphors, etc. They fail at copying the underlying thinking framework. LLMs cannot reproduce this, and they do notice…
I would call this one-row-per-contract-type, and this is the most general model for the problem (e.g. the model cannot be further broken down into finer level), thus, the most scalable model given storage is dirt cheap.
By saying coding is the hard part, you're insulting all engineers. Edit: Although the current job market is heavily distorted, there used to be distinction b/w developer and engineer in the past. As the mainstream…
This one(Zapscape) exploits the same module(shadow MMU) as Januscape, so it does workaround the issue. AFAIK the only code paths that activates shadow MMU are (1) lack of hardware EPT/NPT support (2) nested…
This one exploit "shadow MMU" in the nested virtualization path of KVM, so this one is more-limited than EPT/NPT vul'n (KVM defaults to EPT/NPT, nested virt'n is disabled by default). Nested virtualization is rather a…
> The model outputs are much more concise than when I try and talk to GPT-5.6 Sol about mathematics. By signalling expertise, Tao shunts the model into “talking-to-mathematicians” mode, not “explaining-to-amateurs” mode…
Most of the time, the domestic market in China is indirectly controlled through political and social influence, not through policy. It's not a fair market at all.
I think this kind of automated design process can un-hype humanoids. The point of humanoid robots is one-size-fits-all -- a generalist can readily fit into various scenarios. However, if custom designs are cheap and…
It's more like Claude models entirely suck at extracting key points. No matter how hard I emphasize that it needs to pick the "load-bearing" facts and claims, it cannot stop itself muttering around. It never nails the…
“State of THE art.”
> But I'm beginning to wonder whether that will even matter in a few months. At that point, even Bun itself doesn't matter. All intermediate tools don't matter if LLMs can reliably write something large. The problem is…
An AI slop to filter out AI slops -- the pinnacle of AI. Sadly this would not end just as a joke. We'll be doing this forever from now on.
> widen your reading, rather than trying to win a silly argument on a forum with 'logical fallacies' and other such humdrum wikiphilosophy. Yeah, sure, the paragraph there is pretty bad, I admit. I was having an…
> we are all way worse at reading compared to the average mid-20th centry reader? I would say English has changed. Modern English speakers, especially those in academia and technology, prefers explicit styles over…
No, my view is more like that he was too good at English to write bad English that he was criticizing. The examples he wrote are simply too sharp to fit well into his own criticisms. They are formulated so well that you…
Ah, I just read through it. That’s was a fun read. > This is a parody, but not a very gross one. … So, Orwell wrote the second sentence as a parody of the first sentence, using the modern English that he was criticizing…
I think Orwell did a great job there actually. Despite its bizarre look, the sentence is evocative and eloquent. It does make me get a clear mental image from the very first word. It leaves little room for roaming and…
Yeah, even Pangram is a bit problematic here. It has been notoriously fragile. Minor edits can flip scores from 100%-human to 100%-AI, because Pangram is crazy sensitive to local and surfacial features of text. Simple…
Obviously, NVIDIA is trying to own the AI development chain. Owning HF -- the discovery and distribution channel -- is one thing, but I think the biggest threat vector is the privileged access to HF platform data, that…
It doesn't answer at all. :/
One obvious loophole: upload your full outputs publicly. Anyone can feed those to their models. Wait, we're already doing it. /s
Yes, but the token expires upon logout, and, after reading your comment, I just double-checked the expiry (got 400 from all related endpoints).
You simply don't have enough testcases.
I've just let Claude analyze an HAR file containing XHR traffic on the Facebook timeline. It happily wrote a script for extracting contents, fetch images and videos, and rebuilt the timeline into a single HTML -- highly…
Lisp is a highly expressive language with little constraint, so there’s plenty of room for individuals to become strongly opinionated. That leads to a bit of elite mindset.
LLMs are only good at copying the surface, like frequently used words, sentence structures, use of metaphors, etc. They fail at copying the underlying thinking framework. LLMs cannot reproduce this, and they do notice…
I would call this one-row-per-contract-type, and this is the most general model for the problem (e.g. the model cannot be further broken down into finer level), thus, the most scalable model given storage is dirt cheap.
By saying coding is the hard part, you're insulting all engineers. Edit: Although the current job market is heavily distorted, there used to be distinction b/w developer and engineer in the past. As the mainstream…
This one(Zapscape) exploits the same module(shadow MMU) as Januscape, so it does workaround the issue. AFAIK the only code paths that activates shadow MMU are (1) lack of hardware EPT/NPT support (2) nested…
This one exploit "shadow MMU" in the nested virtualization path of KVM, so this one is more-limited than EPT/NPT vul'n (KVM defaults to EPT/NPT, nested virt'n is disabled by default). Nested virtualization is rather a…
> The model outputs are much more concise than when I try and talk to GPT-5.6 Sol about mathematics. By signalling expertise, Tao shunts the model into “talking-to-mathematicians” mode, not “explaining-to-amateurs” mode…
Most of the time, the domestic market in China is indirectly controlled through political and social influence, not through policy. It's not a fair market at all.
I think this kind of automated design process can un-hype humanoids. The point of humanoid robots is one-size-fits-all -- a generalist can readily fit into various scenarios. However, if custom designs are cheap and…
It's more like Claude models entirely suck at extracting key points. No matter how hard I emphasize that it needs to pick the "load-bearing" facts and claims, it cannot stop itself muttering around. It never nails the…
“State of THE art.”
> But I'm beginning to wonder whether that will even matter in a few months. At that point, even Bun itself doesn't matter. All intermediate tools don't matter if LLMs can reliably write something large. The problem is…