Based on my own experience AI (Fable, Astra) still has major blind spots and is prone to draw obviously wrong conclusions from thin air. Conclusions based on weak evidence that are so obviously blatantly wrong that you…
We need to start numbering these things. I opened the board today and thought "oh, another one again?!" but it's just the thread of yesterday.
It seems to start understanding how bicycles work. Look at the front fork. There are reasons why it's shaped the way it is in real bicycles. If you do physical simulations in 3D with reinforcement learning you will…
There is plenty of evidence that they have improved in all benchmarks and also in my private experience. But have they improved in the things they still fail at? No, they still fail at them. You need only one example of…
Apparently that's necessary and sufficient for getting views. In these interesting times when the future is very uncertain we like to hear people speaking with certainty about the future. Especially people who appear to…
Being early is the same as being wrong. Being lucky is the same as being right.
True, if you have a codebase that works in practice but has dozens of loose ends and poorly defined edge cases than it can chase off into rabbit holes because "oh wait, what if x is undefined instead of null? How is y…
Nobody except corporations who built workflows on top of it and don't care about the price because the developer already moved on and nobody wants to touch it.
Figure 1 made me laugh out loud.
I fear that we can only predict them if we deliberately trigger them to release the energy.
1/f noise basically kills averaging. You collect more signal but at the same time equally more noise.
Apparently LLMs are very good at decompilation. I mean it makes sense, they are trained on readable code and they understand logical structure very well. And a lot of real-world problems involve inverse thinking.…
If you pay your employees enough to sing all day about how much they love you they will sing.
It's not just speed, it will consume a lot less energy per token, maybe even more than 100x difference. And cost for a chip that runs that one model will also go down a lot once volume scales up. They will end up way…
The real crank you need to look at is where we dig the resources out of the ground. All of that will eventually end up in the atmosphere. Plug that hole. Stop digging and drilling. The invisible hand of the free market…
How many miles does the burger drive my car? Not to question your numbers but they lack context. Also, it would be good for nature if there were fewer or no humans on this planet. But then there would be no human here…
Yeah, it's way easier to argue that the product is a bad idea if they give you all the resources to build, release and grow it but there's just not enough customers showing up. The alternative might be internal battles…
Do you get promoted for not launching a product? The interests of the individual and the institution do not always perfectly align.
Can we assume that the test is still private when it was run on many cloud providers?
Also add electricity 50% on top for cooling the DC.
Just got cut off mid-task in central Europe. When I open a new chat it still advertises "Extended through Jul 19" One theory is that they are removing the limits altogether and the update has gone wrong.
Anthropic is arguably still better in tooling and integrating model and tooling. Good habit beats raw intelligence. For code editing Cursor editor tooling is even better.
To me the biggest gain I see is that you take the programmers out of the loop. Instead of formulating your ideas to start a project and then acquiring the resources to do a single iteration on it which may take months…
I have great hope that the CAD & FEM field will benefit greatly from LLM use. To my understanding we currently don't have a free CAD kernel and broadly applicable FEM solver because these are just too hard to make…
Cars are pretty mature while AI is just getting started. Expect 100x price drop for the same quality.
Based on my own experience AI (Fable, Astra) still has major blind spots and is prone to draw obviously wrong conclusions from thin air. Conclusions based on weak evidence that are so obviously blatantly wrong that you…
We need to start numbering these things. I opened the board today and thought "oh, another one again?!" but it's just the thread of yesterday.
It seems to start understanding how bicycles work. Look at the front fork. There are reasons why it's shaped the way it is in real bicycles. If you do physical simulations in 3D with reinforcement learning you will…
There is plenty of evidence that they have improved in all benchmarks and also in my private experience. But have they improved in the things they still fail at? No, they still fail at them. You need only one example of…
Apparently that's necessary and sufficient for getting views. In these interesting times when the future is very uncertain we like to hear people speaking with certainty about the future. Especially people who appear to…
Being early is the same as being wrong. Being lucky is the same as being right.
True, if you have a codebase that works in practice but has dozens of loose ends and poorly defined edge cases than it can chase off into rabbit holes because "oh wait, what if x is undefined instead of null? How is y…
Nobody except corporations who built workflows on top of it and don't care about the price because the developer already moved on and nobody wants to touch it.
Figure 1 made me laugh out loud.
I fear that we can only predict them if we deliberately trigger them to release the energy.
1/f noise basically kills averaging. You collect more signal but at the same time equally more noise.
Apparently LLMs are very good at decompilation. I mean it makes sense, they are trained on readable code and they understand logical structure very well. And a lot of real-world problems involve inverse thinking.…
If you pay your employees enough to sing all day about how much they love you they will sing.
It's not just speed, it will consume a lot less energy per token, maybe even more than 100x difference. And cost for a chip that runs that one model will also go down a lot once volume scales up. They will end up way…
The real crank you need to look at is where we dig the resources out of the ground. All of that will eventually end up in the atmosphere. Plug that hole. Stop digging and drilling. The invisible hand of the free market…
How many miles does the burger drive my car? Not to question your numbers but they lack context. Also, it would be good for nature if there were fewer or no humans on this planet. But then there would be no human here…
Yeah, it's way easier to argue that the product is a bad idea if they give you all the resources to build, release and grow it but there's just not enough customers showing up. The alternative might be internal battles…
Do you get promoted for not launching a product? The interests of the individual and the institution do not always perfectly align.
Can we assume that the test is still private when it was run on many cloud providers?
Also add electricity 50% on top for cooling the DC.
Just got cut off mid-task in central Europe. When I open a new chat it still advertises "Extended through Jul 19" One theory is that they are removing the limits altogether and the update has gone wrong.
Anthropic is arguably still better in tooling and integrating model and tooling. Good habit beats raw intelligence. For code editing Cursor editor tooling is even better.
To me the biggest gain I see is that you take the programmers out of the loop. Instead of formulating your ideas to start a project and then acquiring the resources to do a single iteration on it which may take months…
I have great hope that the CAD & FEM field will benefit greatly from LLM use. To my understanding we currently don't have a free CAD kernel and broadly applicable FEM solver because these are just too hard to make…
Cars are pretty mature while AI is just getting started. Expect 100x price drop for the same quality.