You know how there's a router mode to use the cheapest provider? That only takes into account uncached rates, last I checked. Make another one that takes into account effective rates (the ones that include cache).
Oh I tried in cloud. I'll give it another shot
So then they DIDN'T "learn the secret to cracking the problem". They simply knew that part of the problem was solved. Knowing a problem can be solved and knowing the solution are not the same thing.
I just gave it a try and it doesn't appear to be free, it used up some of my on demand usage. It does say 75% off though. Seems like for Pro subscribers SWE-1.7 is free, maybe SWE-2 is free for them?
It's the website's "grid-bg" effect. It's glitching out and also normally barely visible at all.
The writer is the CEO of... whatever this is: https://auraspark.com/ Completely slopped up website, zero human touch. I would even go so far as to call them a hyprocrite, for getting AI to do all that for them instead…
I think it's still a useful data point. For example, omp, which is pi with some default extensions, scores worse. I do agree that adding more configurations of Pi would help though.
The retain it for "the duration necessary to achieve the intended purposes", which could mean forever.
It's not that far off anymore. On my 7900 XTX 24GB, I can run Qwen3.8 27B with 131K context at Q4_K_M (55 tok/s with MTP). Excluding hardware cost, it's about $0.02 tok/M in and $0.40 tok/M out (cached in $0.0001). On…
To me, this seems like OPC UA / SiLA but instead of being software <-> machine semantics/control it's AI <-> machine semantics/control. Or in simpler terms, an AI-facing hardware abstraction / device-description…
Anything with motion. Sports, games, vlogs, and so on experience immense improvements. And it's not like it takes 2x the bandwidth, because inter-frame compression can be smarter about it. It's like 30-50% more…
I find the tokenizers most compelling. That's what the model is trained on, it's an immutable fact of the model and its architecture. You know for a fact that the model is at least related to other models that way. And…
Even at the hardware level, if it was a separate chip that the camera data passed through or something, that's not really good enough either, people have broken TPMs before. It'd have to be baked into the camera sensor.…
Yeah seems like it doesn't account for anti-fingerprinting at all.
Don't forget negative numbers. https://minusonelabs.com/ https://minus3labs.com/ https://minus9labs.com/ (borked but existed at one point) There are probably a few more
[dead]
Yes: https://huggingface.co/collections/ornith-ai/ornith-15
It's not just the chassis, it's also the screen, storage, RAM, ports, speakers, battery, and so on. And there is lots of demand for old mainboards, they sell pretty quick on Ebay. More sustainable to only buy and sell…
How exactly is this different from something like v86? It's definitely easier to embed but also not as customizable. Like if I don't need bioinformatics stuff, can I just exclude that?
TLDR: It keeps ~20k tokens of recent conversations, then hands the rest of the conversation to another model with a special system & user prompt. This then fills out a template with relevant information. See:…
The ENTIRE thing is AI generated. I'm not talking about the article. I'm talking about the entire website, the entire "product". https://0.mk/blog/zero-humans Also it's super easy to tell by looking at it, way too many…
I think it's just meant to make it more competitive, Gemini has kinda been behind in everything except maybe multimodal. It's only 3 weeks after Flash 3.6, so if they really wanted to, they could probably do a 3.8 Flash…
> afaik the PAYG subscribers are not subject to this lower limit ... In other words, if you want the old provisioning limits for free you just have to put your credit card info in. Always Free can mean different things:…
Most of the friction is just JS overhead for the computations, a compiled solver is like 1000x faster. If Anubis ever gets popular enough that scrapers care, it would be trivial to defeat. And last I checked you could…
It supports anything with ACP. So it can actually run Codex and Claude Code, not just the Zed Agent. Looks like omp supports ACP, so all you have to do is specify a Custom Agent in Zed and it should just work.…
You know how there's a router mode to use the cheapest provider? That only takes into account uncached rates, last I checked. Make another one that takes into account effective rates (the ones that include cache).
Oh I tried in cloud. I'll give it another shot
So then they DIDN'T "learn the secret to cracking the problem". They simply knew that part of the problem was solved. Knowing a problem can be solved and knowing the solution are not the same thing.
I just gave it a try and it doesn't appear to be free, it used up some of my on demand usage. It does say 75% off though. Seems like for Pro subscribers SWE-1.7 is free, maybe SWE-2 is free for them?
It's the website's "grid-bg" effect. It's glitching out and also normally barely visible at all.
The writer is the CEO of... whatever this is: https://auraspark.com/ Completely slopped up website, zero human touch. I would even go so far as to call them a hyprocrite, for getting AI to do all that for them instead…
I think it's still a useful data point. For example, omp, which is pi with some default extensions, scores worse. I do agree that adding more configurations of Pi would help though.
The retain it for "the duration necessary to achieve the intended purposes", which could mean forever.
It's not that far off anymore. On my 7900 XTX 24GB, I can run Qwen3.8 27B with 131K context at Q4_K_M (55 tok/s with MTP). Excluding hardware cost, it's about $0.02 tok/M in and $0.40 tok/M out (cached in $0.0001). On…
To me, this seems like OPC UA / SiLA but instead of being software <-> machine semantics/control it's AI <-> machine semantics/control. Or in simpler terms, an AI-facing hardware abstraction / device-description…
Anything with motion. Sports, games, vlogs, and so on experience immense improvements. And it's not like it takes 2x the bandwidth, because inter-frame compression can be smarter about it. It's like 30-50% more…
I find the tokenizers most compelling. That's what the model is trained on, it's an immutable fact of the model and its architecture. You know for a fact that the model is at least related to other models that way. And…
Even at the hardware level, if it was a separate chip that the camera data passed through or something, that's not really good enough either, people have broken TPMs before. It'd have to be baked into the camera sensor.…
Yeah seems like it doesn't account for anti-fingerprinting at all.
Don't forget negative numbers. https://minusonelabs.com/ https://minus3labs.com/ https://minus9labs.com/ (borked but existed at one point) There are probably a few more
[dead]
Yes: https://huggingface.co/collections/ornith-ai/ornith-15
It's not just the chassis, it's also the screen, storage, RAM, ports, speakers, battery, and so on. And there is lots of demand for old mainboards, they sell pretty quick on Ebay. More sustainable to only buy and sell…
How exactly is this different from something like v86? It's definitely easier to embed but also not as customizable. Like if I don't need bioinformatics stuff, can I just exclude that?
TLDR: It keeps ~20k tokens of recent conversations, then hands the rest of the conversation to another model with a special system & user prompt. This then fills out a template with relevant information. See:…
The ENTIRE thing is AI generated. I'm not talking about the article. I'm talking about the entire website, the entire "product". https://0.mk/blog/zero-humans Also it's super easy to tell by looking at it, way too many…
I think it's just meant to make it more competitive, Gemini has kinda been behind in everything except maybe multimodal. It's only 3 weeks after Flash 3.6, so if they really wanted to, they could probably do a 3.8 Flash…
> afaik the PAYG subscribers are not subject to this lower limit ... In other words, if you want the old provisioning limits for free you just have to put your credit card info in. Always Free can mean different things:…
Most of the friction is just JS overhead for the computations, a compiled solver is like 1000x faster. If Anubis ever gets popular enough that scrapers care, it would be trivial to defeat. And last I checked you could…
It supports anything with ACP. So it can actually run Codex and Claude Code, not just the Zed Agent. Looks like omp supports ACP, so all you have to do is specify a Custom Agent in Zed and it should just work.…