This is a nit but his name is actually Erik Satie not Eric Satre.
It seems like the entire thought process you’re trying to sell hinges on the idea that OpenAI reported the attack first. Did you forget that it was actually HuggingFace that reported it first, and OpenAI only stepped…
It’s just a boring and unhelpful complaint that afaict largely serves to soothe the commenters ego rather than point at anything insightful that’s useful or predictive. Point me to your favorite “next token predictor”…
I just prefer HN comments to be better reflections of reality. There is an unspoken expectation here that people here know what they’re talking about especially when it comes to technical matters. The rise of LLMs has…
Fair so let me be clear. I’m whining because the “next token predictor” reductionist point of view has been wrong and is only growing more wrong with time. Clearly these things can do things that actually matter. Do you…
Are they useful or not? Will they continue changing the world or not? People who choose one way or the other for describing them typically fall on one side or the other in these questions imo. What do you think? Will…
Every day I wake up and open HN. “LLM has made legitimate mathematical discoveries” —> Wow the rate of progress is amazing. Highly upvoted. “LLM does something not good” -> Does everyone else not realize LLMs are just…
There are articles with far fewer upvotes and comments ranking higher on the front page right now, despite being the same age or older than this one. HNs opaque ranking system at it again.
Everyday? Which 10 problems were solved by mathematics grad students in the past 10 days? OK I’ll grant that it’s not your obligation to be my search function (despite you making the wild assertion in the first place),…
I think the fundamental difference between our assumptions is you believe prompts to be optimized for tasks rather than model-task pairs. The only elaboration I can give you is empirical observations and model providers…
Experienced similar between 5.4-mini vs 5.6-luna in our own pipelines but after spending some time on prompt optimization and testing out various reasoning effort levels 5.6-luna was well worth it. Did you just replace…
We can assume (outside of Ollama) that they meant the strongest model from each lab. If you limit yourself to just looking at the literal strings in the list, literally none of these are models. What model is "Deepseek"…
How is this any different than what we have already? We've had this ability for ages (6+ months, decades in the AI world), you can literally today easily prompt CC or Codex to use subagents to accomplish tasks and…
As usual HN posters are hyper aware of other's credentials while ignoring that their BS in CS (if that) doesn't magically qualify them to assess everything in every domain. "I'm a software engineer, I'm sure if I had…
https://xcancel.com/trq212/status/2014051501786931427#m
Hm first Shazeer and now Jumper, DeepMind getting hollowed out this week.
Don't forget that the denominator (total number of outstanding shares) will be increased by this as well. So even if the market cap reacted exactly one to one like you're proposing the per share price wouldn't stay…
No they didn't, they predict they'll get that much. Also worth noting the prediction assumes running at MXFP4/FP8 quantization.
Frontier as in "Frontier Model" is a legitimate vocabulary term you should probably be aware of in 2026. It's not something the author made up or chose randomly, it's common parlance in the space.
Can you really look yourself in the mirror and say with a straight face that fundamentally nothing has changed about the relationship between the US and its allies? Do you really think Europeans will be quick to forgive…
Probably a better way to phrase it would be keeping “pace“. Yes, they are still behind but by about the same amount as they always have, they aren’t drifting further behind. Like two marathon runners, one a a mile…
I mean yes? Considering that we're here reading the news that they've agreed to this.
What exactly is the intuition behind taking something inherently linear like a sequence of 100 days and presenting it as a graph with no information given about the rationale or reasoning behind the edges.
[dead]
The symbol is not the thing. The map is not the territory. Ceci n'est pas une pipe.
This is a nit but his name is actually Erik Satie not Eric Satre.
It seems like the entire thought process you’re trying to sell hinges on the idea that OpenAI reported the attack first. Did you forget that it was actually HuggingFace that reported it first, and OpenAI only stepped…
It’s just a boring and unhelpful complaint that afaict largely serves to soothe the commenters ego rather than point at anything insightful that’s useful or predictive. Point me to your favorite “next token predictor”…
I just prefer HN comments to be better reflections of reality. There is an unspoken expectation here that people here know what they’re talking about especially when it comes to technical matters. The rise of LLMs has…
Fair so let me be clear. I’m whining because the “next token predictor” reductionist point of view has been wrong and is only growing more wrong with time. Clearly these things can do things that actually matter. Do you…
Are they useful or not? Will they continue changing the world or not? People who choose one way or the other for describing them typically fall on one side or the other in these questions imo. What do you think? Will…
Every day I wake up and open HN. “LLM has made legitimate mathematical discoveries” —> Wow the rate of progress is amazing. Highly upvoted. “LLM does something not good” -> Does everyone else not realize LLMs are just…
There are articles with far fewer upvotes and comments ranking higher on the front page right now, despite being the same age or older than this one. HNs opaque ranking system at it again.
Everyday? Which 10 problems were solved by mathematics grad students in the past 10 days? OK I’ll grant that it’s not your obligation to be my search function (despite you making the wild assertion in the first place),…
I think the fundamental difference between our assumptions is you believe prompts to be optimized for tasks rather than model-task pairs. The only elaboration I can give you is empirical observations and model providers…
Experienced similar between 5.4-mini vs 5.6-luna in our own pipelines but after spending some time on prompt optimization and testing out various reasoning effort levels 5.6-luna was well worth it. Did you just replace…
We can assume (outside of Ollama) that they meant the strongest model from each lab. If you limit yourself to just looking at the literal strings in the list, literally none of these are models. What model is "Deepseek"…
How is this any different than what we have already? We've had this ability for ages (6+ months, decades in the AI world), you can literally today easily prompt CC or Codex to use subagents to accomplish tasks and…
As usual HN posters are hyper aware of other's credentials while ignoring that their BS in CS (if that) doesn't magically qualify them to assess everything in every domain. "I'm a software engineer, I'm sure if I had…
https://xcancel.com/trq212/status/2014051501786931427#m
Hm first Shazeer and now Jumper, DeepMind getting hollowed out this week.
Don't forget that the denominator (total number of outstanding shares) will be increased by this as well. So even if the market cap reacted exactly one to one like you're proposing the per share price wouldn't stay…
No they didn't, they predict they'll get that much. Also worth noting the prediction assumes running at MXFP4/FP8 quantization.
Frontier as in "Frontier Model" is a legitimate vocabulary term you should probably be aware of in 2026. It's not something the author made up or chose randomly, it's common parlance in the space.
Can you really look yourself in the mirror and say with a straight face that fundamentally nothing has changed about the relationship between the US and its allies? Do you really think Europeans will be quick to forgive…
Probably a better way to phrase it would be keeping “pace“. Yes, they are still behind but by about the same amount as they always have, they aren’t drifting further behind. Like two marathon runners, one a a mile…
I mean yes? Considering that we're here reading the news that they've agreed to this.
What exactly is the intuition behind taking something inherently linear like a sequence of 100 days and presenting it as a graph with no information given about the rationale or reasoning behind the edges.
[dead]
The symbol is not the thing. The map is not the territory. Ceci n'est pas une pipe.