I had Muse Glimmer (from Meta / Facebook) quoting OpenAI's safety guidelines to me, and I had Poolside's Laguna (a smaller US company) with thinking traces about obeying Chinese law. Both of those are local models, and…
> A common issue is that it's rarely mentioned on which dataset KL-divergence is computed. It seems the most common dataset is wikitext Thank you for calling this out. Using Wikipedia snippets for these is a terrible…
I can't read Sam's mind, so I don't know the point. I was just trying to answer the other dude's (apparently insincere) question. > I think the release of kimi k3 is definitely arguably dangerous I understand the…
Maybe I'm reading too much between the lines, but I suspect the reason is to rub his nose in the duplicity or naivety depending on how generous you're feeling. Publishing the model would be a confession that he was…
There's a link you can click to see the personal background of the respondents. It's tough to know what "researcher in a field not listed" means, but it's possible that over 60% are not even physicists: 30.8% - A…
I would guess that 99% of Americans have a lower than average understanding of stats. Maybe you meant to use the median instead of the average? (Not every distribution is Normal)
This seems very cool, but I'm not sure I understand exactly what it's doing. Are they making a new speculative drafter for Qwen 3.8 27B? Maybe they're optimizing the MLX code for the decoder itself? Thank you in advance.
Your definition is not useful enough. There are people missing any or all of those, and any or all of those can be approximated by a machine.
I've thought about turning this upside down. To any person who is sure they know what has and doesn't have consciousness: If I say I don't have it, can you prove me wrong? As far as I'm concerned, it's a word without a…
I'm not claiming to have any expertise in this area, but I've got a list of things I try to apply when working with LLMs. Possibly relevant here is, "don't tell the model what NOT to do, show it what TO do". I think…
People over-quantize things, muck with the temperature and other settings based on superstitions or results from models they think are similar. There's lots of ways to make 3.8 27B dumber.
Not that my opinion matters much, but I like Deno. I never tried Bun.
Same here - my questions were sophomore level. I think it's notable that when I edited my question to say it was about Gemma 4, it answered without blocking. A cynic like myself would interpret that as evidence they…
What's the distinction between "protecting their turf" and "competitive reasons"? I see them as the same, but I could be missing something. That's an interesting thought on the current "safety blocking" being a trial…
I've gotten flagged for asking questions about tokens and tensors. That makes me believe it's not about safety, it's about protecting their turf. I cancelled my subscription - same fear about getting flagged too much…
All good, but that seems unrelated to what I said, and I'm not sure why you replied to me. Consider this though: Regulating AI models in the US benefits the data centers, not individuals. That's where regulated models…
I think you might misunderstand. The regulation isn't about what OpenAI and Anthropic can do. It's about what you, a citizen, can do.
Yeah, I think I'm seeing the same thing. I don't have all the answers, I just think it'd be a mistake to throw the baby out with the bath water on this model. It seems significantly better than the other dense models…
That's definitely not an argument I was making.
Any argument about regulation in the US which doesn't mention that China and other countries aren't bound by that regulation should be heavily questioned. Exactly who are you stopping from doing what you don't like?!?…
I don't really understand the argument you're making, but just to add a data point: DeepSeek V4 Flash 0731 is 167 gigabytes from the developer and as a GGUF with no additional quantization. It limps along on my 192GB M2…
It won't satisfy the people who just want to drop a model into their existing toolset and run, but I think there are a lot of ways to deal with this overthinking problem. For instance, it's a step backward, but I put…
[dead]
> [...] we evaluated the behavior of various Claude models in a setting with contradictory objectives. > We consistently saw a multiagent turf war... In fact, they sabotaged others with increasingly aggressive,…
One possibility is that you have Claude write something that would not get you in trouble and concatenate that with something you wrote by hand that would. Your signature from the safe material is now associated with…
I had Muse Glimmer (from Meta / Facebook) quoting OpenAI's safety guidelines to me, and I had Poolside's Laguna (a smaller US company) with thinking traces about obeying Chinese law. Both of those are local models, and…
> A common issue is that it's rarely mentioned on which dataset KL-divergence is computed. It seems the most common dataset is wikitext Thank you for calling this out. Using Wikipedia snippets for these is a terrible…
I can't read Sam's mind, so I don't know the point. I was just trying to answer the other dude's (apparently insincere) question. > I think the release of kimi k3 is definitely arguably dangerous I understand the…
Maybe I'm reading too much between the lines, but I suspect the reason is to rub his nose in the duplicity or naivety depending on how generous you're feeling. Publishing the model would be a confession that he was…
There's a link you can click to see the personal background of the respondents. It's tough to know what "researcher in a field not listed" means, but it's possible that over 60% are not even physicists: 30.8% - A…
I would guess that 99% of Americans have a lower than average understanding of stats. Maybe you meant to use the median instead of the average? (Not every distribution is Normal)
This seems very cool, but I'm not sure I understand exactly what it's doing. Are they making a new speculative drafter for Qwen 3.8 27B? Maybe they're optimizing the MLX code for the decoder itself? Thank you in advance.
Your definition is not useful enough. There are people missing any or all of those, and any or all of those can be approximated by a machine.
I've thought about turning this upside down. To any person who is sure they know what has and doesn't have consciousness: If I say I don't have it, can you prove me wrong? As far as I'm concerned, it's a word without a…
I'm not claiming to have any expertise in this area, but I've got a list of things I try to apply when working with LLMs. Possibly relevant here is, "don't tell the model what NOT to do, show it what TO do". I think…
People over-quantize things, muck with the temperature and other settings based on superstitions or results from models they think are similar. There's lots of ways to make 3.8 27B dumber.
Not that my opinion matters much, but I like Deno. I never tried Bun.
Same here - my questions were sophomore level. I think it's notable that when I edited my question to say it was about Gemma 4, it answered without blocking. A cynic like myself would interpret that as evidence they…
What's the distinction between "protecting their turf" and "competitive reasons"? I see them as the same, but I could be missing something. That's an interesting thought on the current "safety blocking" being a trial…
I've gotten flagged for asking questions about tokens and tensors. That makes me believe it's not about safety, it's about protecting their turf. I cancelled my subscription - same fear about getting flagged too much…
All good, but that seems unrelated to what I said, and I'm not sure why you replied to me. Consider this though: Regulating AI models in the US benefits the data centers, not individuals. That's where regulated models…
I think you might misunderstand. The regulation isn't about what OpenAI and Anthropic can do. It's about what you, a citizen, can do.
Yeah, I think I'm seeing the same thing. I don't have all the answers, I just think it'd be a mistake to throw the baby out with the bath water on this model. It seems significantly better than the other dense models…
That's definitely not an argument I was making.
Any argument about regulation in the US which doesn't mention that China and other countries aren't bound by that regulation should be heavily questioned. Exactly who are you stopping from doing what you don't like?!?…
I don't really understand the argument you're making, but just to add a data point: DeepSeek V4 Flash 0731 is 167 gigabytes from the developer and as a GGUF with no additional quantization. It limps along on my 192GB M2…
It won't satisfy the people who just want to drop a model into their existing toolset and run, but I think there are a lot of ways to deal with this overthinking problem. For instance, it's a step backward, but I put…
[dead]
> [...] we evaluated the behavior of various Claude models in a setting with contradictory objectives. > We consistently saw a multiagent turf war... In fact, they sabotaged others with increasingly aggressive,…
One possibility is that you have Claude write something that would not get you in trouble and concatenate that with something you wrote by hand that would. Your signature from the safe material is now associated with…