Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so…
While I instinctively dislike any "secretly spying" devices, phones could do that for a while now. I wouldn’t have much use for my watch doing it, but my nephew swears by taking notes on an ipad with an app that has a…
It would be shocking if it wasn’t trained on sessions. Have you read the ToS parts for both openai and anthropic that talk about it? It’s so obviously a weaselly way to say "no we do not train on your exact chats but we…
GLM 5.3 is probably the sweet spot open weights model if you want to go beyond deepseek flash or the new glm flash. I used it with pi and had a fairly good time, especially since it’s less touchy about cyber and whatnot…
App vs website is purely taste. Some people will never install an app but will use your website. Some people love apps and will never use your website. Some people prefer apps to PWA when the app does thing, because…
The AI-isms in this post are a bit much. Anyway. Routers are such juicy targets, especially since there's less eyes on them than on laptops/desktops, that it's hard to believe there are any which aren't backdoored, be…
> 320B total parameters and just 18B active parameters This is pretty hefty for a "flash" model, even a 256 GB setup is insufficient at q4 - and q4 is already the worst-but-still-acceptable quant in my experience. The…
People are annoyed at apple’s recent lack of polish in software, but the increasing monetization and push into services are a bigger issue imo. You pay the apple premium for hardware and the hardware is indeed great,…
>And what happens when the price goes up by an extra zero on the end? >And then another one? Well so far the price trends towards going down by the zero at the end. Inference isn’t expensive, there’s a whole laundry…
~75% higher price for some extra GPU power and small CPU boost compared to M4? Eh. It's pretty far from M4's value, which was a no-brainer recommendation for people who wanted a tiny box sitting somewhere on their desk.
> OpenAI could still have a significant moat. ChatGPT occupies most consumers’ minds when they think about AI and has become a household name. ChatGPT is AI for the average non-techie the world over, but the average…
As usual, smartphones are an excellent servant but a terrible master. Having a phone, a music player, web browser, camera, gps all in one tiny package? Incredible. Having a machine that lets me doom scroll tiktok for 10…
I know very little about hardware hacking so I can't really judge, but my gut feeling is that this is pretty advanced stuff, right? Granted the models didn't start at zero - the CVE was described online so it had a…
The bitrot would be excessive
There isn't much to upgrade to if all you want is a desktop for gaming, a non-workstation work box or just a decent laptop. Much like smartphones, there just hasn't been much innovation of late. A midrange gaming…
This really doesn’t match my experience, as Gemini was bafflingly bad at any online [re]research when I tried it 2-3 months ago. Not that google’s ai models seem particularly good at anything of late, but given the fact…
I have several 1m long TB4 cables (100W, video, 40Gbps) and they’re neither particularly thick nor were they anywhere close to $60. The best you can get is TB5 (240W, up to 120 Gbps data transfer) and at 1m long they’re…
> a new interface paradigm that's balanced between agentic use and human review (with light editing) Is reviewing endless reams of AI slop something that's doable? It's a very tiring task when you're looking through…
Ah, the newly popular "let's make the policy unusually broad and treat all foreign production as presumptively dangerous, regardless of country, manufacturer or security practices, this way we're more transparent about…
While I am grateful for open weights models I never found much use of them in the past, barring those I could run myself. This changed with deepseek 4 - it is staggeringly cheap, even if the performance definitely isn't…
So far the first and only sophisticated cyberattack was carried out by openai rather than by one the subversive commie models that hate us for our freedoms, so forgive me for taking their safety and alignment spiel…
What Apple wants out of Google is Siri that runs at 8gb ram and isn’t a horrible embarrassment that feels like a primitive markov chain. Given how good Gemma 4 is, Google can squeeze some serious performance in small…
Tons of guardrails, lazy model, super confusing plans, expensive 3.5/3.6 flash and lite and 3.5 pro MiA? Rough patch for google ai
Article feels a bit shallow. Obviously carriers will do their usual fuckery, anyone who ever dealt with one knows this. What many people probably wouldn’t expect is when you’re in a country where physical sim are…
While a valid point, China also produces plenty of whitepapers going about the architecture and know how about the training and inference itself. There’s also the fact that unless LLMs do get to AGI (which seems……
Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so…
While I instinctively dislike any "secretly spying" devices, phones could do that for a while now. I wouldn’t have much use for my watch doing it, but my nephew swears by taking notes on an ipad with an app that has a…
It would be shocking if it wasn’t trained on sessions. Have you read the ToS parts for both openai and anthropic that talk about it? It’s so obviously a weaselly way to say "no we do not train on your exact chats but we…
GLM 5.3 is probably the sweet spot open weights model if you want to go beyond deepseek flash or the new glm flash. I used it with pi and had a fairly good time, especially since it’s less touchy about cyber and whatnot…
App vs website is purely taste. Some people will never install an app but will use your website. Some people love apps and will never use your website. Some people prefer apps to PWA when the app does thing, because…
The AI-isms in this post are a bit much. Anyway. Routers are such juicy targets, especially since there's less eyes on them than on laptops/desktops, that it's hard to believe there are any which aren't backdoored, be…
> 320B total parameters and just 18B active parameters This is pretty hefty for a "flash" model, even a 256 GB setup is insufficient at q4 - and q4 is already the worst-but-still-acceptable quant in my experience. The…
People are annoyed at apple’s recent lack of polish in software, but the increasing monetization and push into services are a bigger issue imo. You pay the apple premium for hardware and the hardware is indeed great,…
>And what happens when the price goes up by an extra zero on the end? >And then another one? Well so far the price trends towards going down by the zero at the end. Inference isn’t expensive, there’s a whole laundry…
~75% higher price for some extra GPU power and small CPU boost compared to M4? Eh. It's pretty far from M4's value, which was a no-brainer recommendation for people who wanted a tiny box sitting somewhere on their desk.
> OpenAI could still have a significant moat. ChatGPT occupies most consumers’ minds when they think about AI and has become a household name. ChatGPT is AI for the average non-techie the world over, but the average…
As usual, smartphones are an excellent servant but a terrible master. Having a phone, a music player, web browser, camera, gps all in one tiny package? Incredible. Having a machine that lets me doom scroll tiktok for 10…
I know very little about hardware hacking so I can't really judge, but my gut feeling is that this is pretty advanced stuff, right? Granted the models didn't start at zero - the CVE was described online so it had a…
The bitrot would be excessive
There isn't much to upgrade to if all you want is a desktop for gaming, a non-workstation work box or just a decent laptop. Much like smartphones, there just hasn't been much innovation of late. A midrange gaming…
This really doesn’t match my experience, as Gemini was bafflingly bad at any online [re]research when I tried it 2-3 months ago. Not that google’s ai models seem particularly good at anything of late, but given the fact…
I have several 1m long TB4 cables (100W, video, 40Gbps) and they’re neither particularly thick nor were they anywhere close to $60. The best you can get is TB5 (240W, up to 120 Gbps data transfer) and at 1m long they’re…
> a new interface paradigm that's balanced between agentic use and human review (with light editing) Is reviewing endless reams of AI slop something that's doable? It's a very tiring task when you're looking through…
Ah, the newly popular "let's make the policy unusually broad and treat all foreign production as presumptively dangerous, regardless of country, manufacturer or security practices, this way we're more transparent about…
While I am grateful for open weights models I never found much use of them in the past, barring those I could run myself. This changed with deepseek 4 - it is staggeringly cheap, even if the performance definitely isn't…
So far the first and only sophisticated cyberattack was carried out by openai rather than by one the subversive commie models that hate us for our freedoms, so forgive me for taking their safety and alignment spiel…
What Apple wants out of Google is Siri that runs at 8gb ram and isn’t a horrible embarrassment that feels like a primitive markov chain. Given how good Gemma 4 is, Google can squeeze some serious performance in small…
Tons of guardrails, lazy model, super confusing plans, expensive 3.5/3.6 flash and lite and 3.5 pro MiA? Rough patch for google ai
Article feels a bit shallow. Obviously carriers will do their usual fuckery, anyone who ever dealt with one knows this. What many people probably wouldn’t expect is when you’re in a country where physical sim are…
While a valid point, China also produces plenty of whitepapers going about the architecture and know how about the training and inference itself. There’s also the fact that unless LLMs do get to AGI (which seems……