They are more than a accessory. You wouldn't say someone who funds a contract killing is merely an accessory to murder. VCs should likewise be considered as principal offenders, unless there is proof that the company…
Yes they are quite good, but are not able to run on a 16GB RX 9070. Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest…
Traffic jams can be a feature, if it's the game Prototype where you could barrel through cars like a freight train.
Qwen 3.8 27B is so far the closest I have run, though it suffers on speed compared to Gemma4 26B MoE model which I still use. Neither are going to match Opus or Sol though, but they can be as fast or faster depending on…
I haven't yet, though plan to when this model is released. The models that I've been daily driving (Gemma 4 26B, Qwen 3.8 27B) have fit nicely on my 3090. I think FreeToken only offers a perf increase for MoE models…
> It is still slow, a lot slower than what you are used to with claude and co. That really depends on the model, I run a few models locally. All at speeds comparable to or faster than Opus. In general we haven't reached…
True, but you can link them up over thunderbolt or Ethernet. If your goal is to run local LLMs, not all weights need to live on the same computer. You can segment the workload by layers and pass the activations along…
It will be interesting to see the intersection of this with inference engines like FreeToken which improve distribution of work for MoE models across CPU/RAM and GPU/VRAM. If all it takes for a competitive model to run…
It needs to be a little less than double the overall price for 256GB option, otherwise it's better to get 2 256GB and link them.
And you'd be okay with Anthropic being the ones that determine what "significantly large" means?
Yes I want so badly to work at Antropic. - 50% or more of my compensation locked up with what I believe has a less than 5% chance of keeping significant value - Work with engineers who believe there job now is to vibe…
No that is not inconceivable. What is inconceivable is believing that group of individuals is not a small minority, and the question will filter everyone else out.
In my experience hype driven companies are very resistant to that even when equity/cash ratio offered is more sane.
It would be a decent interview question if such a large portion of compensation did not come from equity. Otherwise it's basically the same as asking how someone would feel if all the sudden they lost 50% or more of…
It seems like at every turn the word private is avoided and replaced with "non-public". I understand that the system probably is sound, the data just isn't encrypted, which I wouldn't necessarily expect. Maybe it's…
Try it out with your own prompt, I suspect you'll be surprised.
I've done a few variations, I've been impressed with all of them. My favourite so far has been "Generate an SVG of a turtle flying a kite", result: https://imgur.com/a/bdKJPV4. Some will say conflating flying and flying…
Congrats on the release! Nice to see transitions API go away, though I am not entirely convinced of the async model. It's capability is impressive, yet I do not like hiding what values are async, and throwing to await…
Yes latency, and the usual preference of ownership over rentership. Similarly their are benefits to running local AI too, like data privacy and control.
The gains wouldn't be "free lunch", it's the result of time and effort researching optimal design and architecture. Even if the idea of "no free lunch" was taken liberally discounting the cost of research, it would only…
If the brain does rely on quantum effects, it's still possible the quantum effects in use are able to be simulated efficiently on a classical computer. For example if it's a matter of signal transfer rather than quantum…
You can find it today for gaming. Even despite the outlandish rise in hardware costs, there is very little demand for cloud gaming.
It may not make financial sense for someone retired, not into tech, and/or data privacy to host their own LLMs. However if usage of AI in day to day lives continues to increase, I think it will eventually make sense for…
We've barely even started on optimizations like advanced language aware grammars, and specialization routing (dynamically loading fine tunes or seperate weights for specific tasks or languages).
They should've done $0.01 off. Effectively the same discount, while being more outrageous.
They are more than a accessory. You wouldn't say someone who funds a contract killing is merely an accessory to murder. VCs should likewise be considered as principal offenders, unless there is proof that the company…
Yes they are quite good, but are not able to run on a 16GB RX 9070. Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest…
Traffic jams can be a feature, if it's the game Prototype where you could barrel through cars like a freight train.
Qwen 3.8 27B is so far the closest I have run, though it suffers on speed compared to Gemma4 26B MoE model which I still use. Neither are going to match Opus or Sol though, but they can be as fast or faster depending on…
I haven't yet, though plan to when this model is released. The models that I've been daily driving (Gemma 4 26B, Qwen 3.8 27B) have fit nicely on my 3090. I think FreeToken only offers a perf increase for MoE models…
> It is still slow, a lot slower than what you are used to with claude and co. That really depends on the model, I run a few models locally. All at speeds comparable to or faster than Opus. In general we haven't reached…
True, but you can link them up over thunderbolt or Ethernet. If your goal is to run local LLMs, not all weights need to live on the same computer. You can segment the workload by layers and pass the activations along…
It will be interesting to see the intersection of this with inference engines like FreeToken which improve distribution of work for MoE models across CPU/RAM and GPU/VRAM. If all it takes for a competitive model to run…
It needs to be a little less than double the overall price for 256GB option, otherwise it's better to get 2 256GB and link them.
And you'd be okay with Anthropic being the ones that determine what "significantly large" means?
Yes I want so badly to work at Antropic. - 50% or more of my compensation locked up with what I believe has a less than 5% chance of keeping significant value - Work with engineers who believe there job now is to vibe…
No that is not inconceivable. What is inconceivable is believing that group of individuals is not a small minority, and the question will filter everyone else out.
In my experience hype driven companies are very resistant to that even when equity/cash ratio offered is more sane.
It would be a decent interview question if such a large portion of compensation did not come from equity. Otherwise it's basically the same as asking how someone would feel if all the sudden they lost 50% or more of…
It seems like at every turn the word private is avoided and replaced with "non-public". I understand that the system probably is sound, the data just isn't encrypted, which I wouldn't necessarily expect. Maybe it's…
Try it out with your own prompt, I suspect you'll be surprised.
I've done a few variations, I've been impressed with all of them. My favourite so far has been "Generate an SVG of a turtle flying a kite", result: https://imgur.com/a/bdKJPV4. Some will say conflating flying and flying…
Congrats on the release! Nice to see transitions API go away, though I am not entirely convinced of the async model. It's capability is impressive, yet I do not like hiding what values are async, and throwing to await…
Yes latency, and the usual preference of ownership over rentership. Similarly their are benefits to running local AI too, like data privacy and control.
The gains wouldn't be "free lunch", it's the result of time and effort researching optimal design and architecture. Even if the idea of "no free lunch" was taken liberally discounting the cost of research, it would only…
If the brain does rely on quantum effects, it's still possible the quantum effects in use are able to be simulated efficiently on a classical computer. For example if it's a matter of signal transfer rather than quantum…
You can find it today for gaming. Even despite the outlandish rise in hardware costs, there is very little demand for cloud gaming.
It may not make financial sense for someone retired, not into tech, and/or data privacy to host their own LLMs. However if usage of AI in day to day lives continues to increase, I think it will eventually make sense for…
We've barely even started on optimizations like advanced language aware grammars, and specialization routing (dynamically loading fine tunes or seperate weights for specific tasks or languages).
They should've done $0.01 off. Effectively the same discount, while being more outrageous.