I thought it was just a couple of random old problems that had been solved, I didn't realize it was so many, but still I wonder if AI is really getting "better" at math than humans or if there's just a particular set or…
I use Sonnet 80% of the time when I'm on claude.ai and on Claude Code my main agent is Opus and typically my sub-agents are Sonnet. I use this setup mainly because I genuinely don't know if anything I'm doing is complex…
Do you have a prompt that can stop it from saying, "I'd push back on this"... or maybe with this it will now say "The algorithm pushes back on this".. Seriously however, I think this may also help with the natural urge…
Thanks for saying the quiet part out loud. Everytime I do a demo it seems like a waste of time compared to just actually building. Unless I'm doing something super complicated even taking time to set things up like…
That's a very interesting perspective but now that you mention it, I can't even think of the last time I daydreamed, even when I'm not on a device. Curious if phone usage or whatever doom-scrolling we do exhausts the…
Anthropic opening up Fable for subscription use instead of gating it behind an API is a clear response to Kimi3 and other Chinese models catching up. The plan was to make temper the subscription model to "mid-tier" and…
Because the unsafe product most likely has 200+ "positive" reviews!
In WW2 thousands of troops could die in a single day and we wouldn't know their names until months or years later. Today a single troop dies and it is the first story on the national news. Millions of civilians also…
Alibaba made $59B last year in gross profit, and the amount will most likely be appealed and reduced. The biggest issue by far isn't the counterfeit or unsafe products, its the fake reviews, that's where regulators need…
The reality however is that unless you're spending tens of thousands on hardware or setting up cloud instances you can't run anything near close to state of the art. In almost all cases right now, a subscription with a…
Findings like these surprise me. I had a NDE a few years ago and my observation was such that everything slowed down tremendously and I was able to process every instant like it was slow motion. This wasn't like a…
These ridiculous market caps can only be justified if LLMs can ultimately produce outputs that are above human capabilities OR reduce costs and perform human tasks for cheaper than the human rate. Neither of these…
I graduated with a Bachelors in Math in 2018 and there's an entire new math now? For people who actually know the curriculum side: where does geometric algebra fit? Is it something that should come after Calc III /…
How is prompt caching different than caching responses in a database? If you use the same prompt wouldn't you want the same answer? Or can this be used for some type of intermediate process where different questions may…
I thought it was just a couple of random old problems that had been solved, I didn't realize it was so many, but still I wonder if AI is really getting "better" at math than humans or if there's just a particular set or…
I use Sonnet 80% of the time when I'm on claude.ai and on Claude Code my main agent is Opus and typically my sub-agents are Sonnet. I use this setup mainly because I genuinely don't know if anything I'm doing is complex…
Do you have a prompt that can stop it from saying, "I'd push back on this"... or maybe with this it will now say "The algorithm pushes back on this".. Seriously however, I think this may also help with the natural urge…
Thanks for saying the quiet part out loud. Everytime I do a demo it seems like a waste of time compared to just actually building. Unless I'm doing something super complicated even taking time to set things up like…
That's a very interesting perspective but now that you mention it, I can't even think of the last time I daydreamed, even when I'm not on a device. Curious if phone usage or whatever doom-scrolling we do exhausts the…
Anthropic opening up Fable for subscription use instead of gating it behind an API is a clear response to Kimi3 and other Chinese models catching up. The plan was to make temper the subscription model to "mid-tier" and…
Because the unsafe product most likely has 200+ "positive" reviews!
In WW2 thousands of troops could die in a single day and we wouldn't know their names until months or years later. Today a single troop dies and it is the first story on the national news. Millions of civilians also…
Alibaba made $59B last year in gross profit, and the amount will most likely be appealed and reduced. The biggest issue by far isn't the counterfeit or unsafe products, its the fake reviews, that's where regulators need…
The reality however is that unless you're spending tens of thousands on hardware or setting up cloud instances you can't run anything near close to state of the art. In almost all cases right now, a subscription with a…
Findings like these surprise me. I had a NDE a few years ago and my observation was such that everything slowed down tremendously and I was able to process every instant like it was slow motion. This wasn't like a…
These ridiculous market caps can only be justified if LLMs can ultimately produce outputs that are above human capabilities OR reduce costs and perform human tasks for cheaper than the human rate. Neither of these…
I graduated with a Bachelors in Math in 2018 and there's an entire new math now? For people who actually know the curriculum side: where does geometric algebra fit? Is it something that should come after Calc III /…
How is prompt caching different than caching responses in a database? If you use the same prompt wouldn't you want the same answer? Or can this be used for some type of intermediate process where different questions may…