Biological minds in biological organisms are self referential in the way that you have a neural network that forms a model of the world. That model then "discovers" that it is "it self a part of the world" so it tries…
It's likely/It could be an effect of more reinforcement learning in training compared to earlier. You need loots of RL to learn to code well.
(1) LLMs collapse and start outputting garbage after a number of tokens if you do not sample and just pick the "best token" each time. This is a consequence of how they are trained.
Presumably it is just that no one has tried to do it all the way yet. But i know there is some experimental work out there. https://x.com/comma_ai/status/1956879764829995394
Interesting topic. That said I don't know how useful this is since LLMs are primarily trained using mode-covering training rather than Mode-seeking(RL) training, which means LLMs can not form (and does not have) the…
Yes there is. The use of Gödel numbers to represent a system that contain itself is infinite regression.
Gödels incompleteness is just an example of the fact that you cant determine the outcome of infinite regression (in the general case). The same as me asking you to give me the last digit of pi. I am a bit annoyed by pop…
I would think coaxial designs suffer as the lower rotor works in more turbulent non laminar air.
That would "using reason in chain of thought".
That would not be intuition that is just randomness. Intuition is not randomness. A jump in intuition comes from automatic processes reorganising the relational structure of conceptual models. There is no reorganisation…
Clearly it impossible to observe brain chemistry by looking at a person. Psychedelics does however have longterm behavioural effects that are statistically significant and to some extent observable. Using a definition…
Clearly LLMs cant do leaps of intuition since their "intuition" is locked after training ends. The only way a LLM can come up with new ideas if the "idea" appeared as a generalisation durring training or if it was…
It certainly looks like a Claude design to some extent; not all they way however.
Nice job, i have also been working on a few JEPA based models during the last few months. Trying to make more efficient LLMs. I feel like you hit the main issues in the use of jepa models (well except collapse but…
Wont this kill the kv cache? Also i am pretty sure neither open ai or anthropic leets you seed the agents own tokens.
Is the thinking even done in real tokens? I thought it was done using the pure residual stream. That is instead of collapsing the residual stream to a token you treat the final layers output as a vector of size d_model…
If you have a number that is 1000^90000000 that number is larger than the number of atoms in the observable universe.
The issue is apparently this commit (someone did a git bisect): https://github.com/RsyncProject/rsync/commit/859d44fa4f14207... Which is a fix to the security issue CVE-2026-29518:…
"A is not B instead A is blah blah" instead of just saying "A" is a very common pattern have seen in Claude. It is strange to read as the topic A has often not been introduced and introducing it by saying what it is not…
No you could rent virtualised servers way before AWS. AWS simply had good marketing. The virtualised server thing was not a AWS thing, the thing that was were their other services. For example instead of renting a…
Pragmatically enterprise tends to mean less refined, designed by committee and expensive. In this case i would guess it is mostly a justification for taking a part of the LLM pie.
Not sure that i understand your position exactly. But consciousness is also "just a story" (a complicated one) that the human body tells the human mind. We cant know from the outside if "the story" inside a LLM is…
I was not talking about the actual feeling in the moment. The point is the valence of the thing. Ie fear of a thing is a pointer to that thing having negative valence.
Most chatbots are not trained to have/emulate emotions so pain or fear of death is non existent. Therefore killing them and/or using them as slaves is not a moral issue. Thats how i reason. On another point, LLMs are…
Removes paradoxical stuff like claims that there are bigger and smaller infinities. Paradoxes comes from contradictions, a mathematical system that contains contradictions is a failed mathematical system.
Biological minds in biological organisms are self referential in the way that you have a neural network that forms a model of the world. That model then "discovers" that it is "it self a part of the world" so it tries…
It's likely/It could be an effect of more reinforcement learning in training compared to earlier. You need loots of RL to learn to code well.
(1) LLMs collapse and start outputting garbage after a number of tokens if you do not sample and just pick the "best token" each time. This is a consequence of how they are trained.
Presumably it is just that no one has tried to do it all the way yet. But i know there is some experimental work out there. https://x.com/comma_ai/status/1956879764829995394
Interesting topic. That said I don't know how useful this is since LLMs are primarily trained using mode-covering training rather than Mode-seeking(RL) training, which means LLMs can not form (and does not have) the…
Yes there is. The use of Gödel numbers to represent a system that contain itself is infinite regression.
Gödels incompleteness is just an example of the fact that you cant determine the outcome of infinite regression (in the general case). The same as me asking you to give me the last digit of pi. I am a bit annoyed by pop…
I would think coaxial designs suffer as the lower rotor works in more turbulent non laminar air.
That would "using reason in chain of thought".
That would not be intuition that is just randomness. Intuition is not randomness. A jump in intuition comes from automatic processes reorganising the relational structure of conceptual models. There is no reorganisation…
Clearly it impossible to observe brain chemistry by looking at a person. Psychedelics does however have longterm behavioural effects that are statistically significant and to some extent observable. Using a definition…
Clearly LLMs cant do leaps of intuition since their "intuition" is locked after training ends. The only way a LLM can come up with new ideas if the "idea" appeared as a generalisation durring training or if it was…
It certainly looks like a Claude design to some extent; not all they way however.
Nice job, i have also been working on a few JEPA based models during the last few months. Trying to make more efficient LLMs. I feel like you hit the main issues in the use of jepa models (well except collapse but…
Wont this kill the kv cache? Also i am pretty sure neither open ai or anthropic leets you seed the agents own tokens.
Is the thinking even done in real tokens? I thought it was done using the pure residual stream. That is instead of collapsing the residual stream to a token you treat the final layers output as a vector of size d_model…
If you have a number that is 1000^90000000 that number is larger than the number of atoms in the observable universe.
The issue is apparently this commit (someone did a git bisect): https://github.com/RsyncProject/rsync/commit/859d44fa4f14207... Which is a fix to the security issue CVE-2026-29518:…
"A is not B instead A is blah blah" instead of just saying "A" is a very common pattern have seen in Claude. It is strange to read as the topic A has often not been introduced and introducing it by saying what it is not…
No you could rent virtualised servers way before AWS. AWS simply had good marketing. The virtualised server thing was not a AWS thing, the thing that was were their other services. For example instead of renting a…
Pragmatically enterprise tends to mean less refined, designed by committee and expensive. In this case i would guess it is mostly a justification for taking a part of the LLM pie.
Not sure that i understand your position exactly. But consciousness is also "just a story" (a complicated one) that the human body tells the human mind. We cant know from the outside if "the story" inside a LLM is…
I was not talking about the actual feeling in the moment. The point is the valence of the thing. Ie fear of a thing is a pointer to that thing having negative valence.
Most chatbots are not trained to have/emulate emotions so pain or fear of death is non existent. Therefore killing them and/or using them as slaves is not a moral issue. Thats how i reason. On another point, LLMs are…
Removes paradoxical stuff like claims that there are bigger and smaller infinities. Paradoxes comes from contradictions, a mathematical system that contains contradictions is a failed mathematical system.