>it's far beyond the capacity of any human to review the log files of their activity Maybe they should contract with one of the other AI labs. I hear they have LLMs that are good at that kind of thing.
I don’t understand why everyone is so focused on watching the CoT. The tool calls can’t be faked, and they would have set off alarm bells all by themselves.
I don’t see how it can be safe to release this model if it has the training history that led to the huggingface hack. You can’t just roll back that kind of reinforcement learning after the fact. Especially because these…
Yes, which makes it absurd that they apparently weren’t checking their RL rollouts for evidence of reward hacking and punishing it. Even if no one expected this particular type of reward hacking, they should have had a…
> [during training] it's not feasible for anyone to "notice" or get involved I can’t disagree more strongly. Having checks for reward hacking is especially important during training, since it’s humans’ only real chance…
Exactly. So incredibly reckless. > After knowing that the server was hacked, the internal team finds the message board and does nothing with the information. They caught their AIs swarming and did not even inform…
This article uses the word “conceptions” throughout, but it sounds like the only data is about births. These births would have all occurred after the recessions started. Did the original paper rule out the hypothesis…
It’s basically drain cleaner (which becomes salty water as soon as it gets mixed with an acid), plus digested meat juice.
Even better, add an indicator light to the hardware that shows the OS is asking and not an app (although this suffers from the same problem that users need to notice something is absent).
That's only if you're using a third-party keyboard, I think.
Any individual bit could be noise, so it might take 2-3 bits to transmit something unambiguous (via an error-correcting code or similar). If so, I think the speed advantage disappears.
>voting could reduce the likelihood of the draft being reinstated ... only in the unlikely event that my vote changes the outcome of an election, and the candidate I voted for would also be able to affect such a big…
The authors state that HashCat uses Markov models and that they outperform it.
Maybe the Department of Justice? I don't know if they go for that sort of thing. The only other option I can think of off the top of my head is the FCC.
As a non-lawyer, I'd expect to miss something important from the legislative text. Your point about relying on other organizations is a good one, though. I generally take this approach, but I wasn't expecting anything…
Huh? The article is written by Harris and suggests that a piece of legislation she wrote will solve the problem.
Reforming bail sounds like a good idea. I don't know why anyone would trust Harris on the issue, though. If she wanted to shrink jails when she was California's Attorney General, she could have respected court orders…
Yeah, "150 years of lichen biology" would be much better.
No one in this thread claimed to be God. It's worth remembering that the whole point of Hospital IT is to facilitate the doctors' and administrators' work.
They claimed (or at least implied) that you'd be buying something closer to raw hunks of fruit and vegetables in the bag, so it would be fresh.
see also: Muphry's Law: https://en.wikipedia.org/wiki/Muphry's_law
It is, it's just not accurate. https://en.wikipedia.org/wiki/Accuracy_and_precision
It doesn't avoid the R interface, which is GPL'd (version 3).
Several of the universal perturbation vectors in Figure 4 remind me a lot of Deep Dream's textures. I wonder what it is about these high-saturation, stripy-spiraly bits that these networks are responding to. Is it…
I was doored once. I remember flying through the air and thinking, "oh, this isn't so bad." Then I landed and almost broke my hand.
>it's far beyond the capacity of any human to review the log files of their activity Maybe they should contract with one of the other AI labs. I hear they have LLMs that are good at that kind of thing.
I don’t understand why everyone is so focused on watching the CoT. The tool calls can’t be faked, and they would have set off alarm bells all by themselves.
I don’t see how it can be safe to release this model if it has the training history that led to the huggingface hack. You can’t just roll back that kind of reinforcement learning after the fact. Especially because these…
Yes, which makes it absurd that they apparently weren’t checking their RL rollouts for evidence of reward hacking and punishing it. Even if no one expected this particular type of reward hacking, they should have had a…
> [during training] it's not feasible for anyone to "notice" or get involved I can’t disagree more strongly. Having checks for reward hacking is especially important during training, since it’s humans’ only real chance…
Exactly. So incredibly reckless. > After knowing that the server was hacked, the internal team finds the message board and does nothing with the information. They caught their AIs swarming and did not even inform…
This article uses the word “conceptions” throughout, but it sounds like the only data is about births. These births would have all occurred after the recessions started. Did the original paper rule out the hypothesis…
It’s basically drain cleaner (which becomes salty water as soon as it gets mixed with an acid), plus digested meat juice.
Even better, add an indicator light to the hardware that shows the OS is asking and not an app (although this suffers from the same problem that users need to notice something is absent).
That's only if you're using a third-party keyboard, I think.
Any individual bit could be noise, so it might take 2-3 bits to transmit something unambiguous (via an error-correcting code or similar). If so, I think the speed advantage disappears.
>voting could reduce the likelihood of the draft being reinstated ... only in the unlikely event that my vote changes the outcome of an election, and the candidate I voted for would also be able to affect such a big…
The authors state that HashCat uses Markov models and that they outperform it.
Maybe the Department of Justice? I don't know if they go for that sort of thing. The only other option I can think of off the top of my head is the FCC.
As a non-lawyer, I'd expect to miss something important from the legislative text. Your point about relying on other organizations is a good one, though. I generally take this approach, but I wasn't expecting anything…
Huh? The article is written by Harris and suggests that a piece of legislation she wrote will solve the problem.
Reforming bail sounds like a good idea. I don't know why anyone would trust Harris on the issue, though. If she wanted to shrink jails when she was California's Attorney General, she could have respected court orders…
Yeah, "150 years of lichen biology" would be much better.
No one in this thread claimed to be God. It's worth remembering that the whole point of Hospital IT is to facilitate the doctors' and administrators' work.
They claimed (or at least implied) that you'd be buying something closer to raw hunks of fruit and vegetables in the bag, so it would be fresh.
see also: Muphry's Law: https://en.wikipedia.org/wiki/Muphry's_law
It is, it's just not accurate. https://en.wikipedia.org/wiki/Accuracy_and_precision
It doesn't avoid the R interface, which is GPL'd (version 3).
Several of the universal perturbation vectors in Figure 4 remind me a lot of Deep Dream's textures. I wonder what it is about these high-saturation, stripy-spiraly bits that these networks are responding to. Is it…
I was doored once. I remember flying through the air and thinking, "oh, this isn't so bad." Then I landed and almost broke my hand.