had my account since the gpt 3.5 days, disabled it once and its still disabled. though, now that i have advanced account security enabled, the setting is disabled entirely
the last model to use the gpt-4o base model was gpt 5.1, since then its been new pre-trains but this is a new one entirely to itself
make your only legit users mine fake crypto to access your site, only costs them 5% battery on an android device
steady lads, deploying more capital
if you're benchmaxxing then maybe bigger doesnt always mean better, but for general intelligence and big model smell, that couldn't be further from the truth the oss models are impressive but it's pretty clear how…
They mention it uses MXFP4 quant which is a blackwell capability but it looks like this is also supported by ascend 950 series according to marketing material
i wonder if they put an older cutoff date into the prompt intentionally so that when asked on more current events it leans towards tool calls / web searches for tuning
the model obviously knows things after the reported date but its just curious that it reports that date consistently could be they do it intentionally to encourage more tool calls/searches or for tuning reasons
with thinking off and tools disabled: Donald Trump won the 2024 U.S. presidential election.
you cant but its pretty reproducible across api and codex and other agents so i just thought it was odd. full text it gives: Knowledge cutoff: 2024-06 Current date: 2026-04-24 You are an AI assistant accessed via an…
API page lists the knowledge cutoff as Dec 01, 2025 but when prompting the model it says June 2024. Knowledge cutoff: 2024-06 Current date: 2026-04-24 You are an AI assistant accessed via an API.
Memory bandwidth is the biggest L on the dgx spark, it’s half my MacBook from 2023 and that’s the biggest tok/sec bottleneck
show us the benchmarks with "adaptive thinking" turned on
the MDM profile requirement is suspect though I get why they are doing it. but it doesn't inspire confidence to see that their profile is unsigned and still using the default micromdn scep challenge...
it should, lume is a thin wrapper around Apple's Virtualization.framework as i understand it
starting with M3+ you can use Hypervisor.framework/Virtualization.framework to spin up nested VMs. it would be amusing if that bypassed the limit.
I tried the periodic table in their examples using sonnet 4.6 on the $20/mo plan. After a few minutes Claude told me it reached the max message length and bailed. I pressed continue and eventually it generated the…
claude models with 'extended thinking' toggled answer very quickly and the quality of the answer is far ahead of what gpt 5.2 'instant' provides. i wont even bother using the non-thinking version of chatgpt because the…
“After creating a new account, I can confirm the quota drains 2.5x–3x slower. So basically Max (5x) on an older accounts is almost like Pro on a new one in terms of quota. Pretty blatant rug pull tbh.” lol
claude has half the context window size of codex and blows through a good percentage of it right off the bat by injecting a system prompt the size of don quixote
it sounds like the data can be involuntarily disclosed to an external third party (the attacker’s domain) purely because someone reviewed logs that auto-load remote images their log viewer renders the markdown and their…
i thought the article was going to go there, just redirecting the host to a self-hosted ip address serving the bin, but i was pleasantly surprised it didn’t! interesting to learn about the patching process and tooling…
like leveling to 99 in old school runescape
its possible to use gpt-5-high on the plus plan with codex-cli, its a whole different beast! i dont think theres any other way for plus users to leverage gpt-5 with high reasoning. codex -m gpt-5…
it only requires exponentially MORE money for linear returns!
had my account since the gpt 3.5 days, disabled it once and its still disabled. though, now that i have advanced account security enabled, the setting is disabled entirely
the last model to use the gpt-4o base model was gpt 5.1, since then its been new pre-trains but this is a new one entirely to itself
make your only legit users mine fake crypto to access your site, only costs them 5% battery on an android device
steady lads, deploying more capital
if you're benchmaxxing then maybe bigger doesnt always mean better, but for general intelligence and big model smell, that couldn't be further from the truth the oss models are impressive but it's pretty clear how…
They mention it uses MXFP4 quant which is a blackwell capability but it looks like this is also supported by ascend 950 series according to marketing material
i wonder if they put an older cutoff date into the prompt intentionally so that when asked on more current events it leans towards tool calls / web searches for tuning
the model obviously knows things after the reported date but its just curious that it reports that date consistently could be they do it intentionally to encourage more tool calls/searches or for tuning reasons
with thinking off and tools disabled: Donald Trump won the 2024 U.S. presidential election.
you cant but its pretty reproducible across api and codex and other agents so i just thought it was odd. full text it gives: Knowledge cutoff: 2024-06 Current date: 2026-04-24 You are an AI assistant accessed via an…
API page lists the knowledge cutoff as Dec 01, 2025 but when prompting the model it says June 2024. Knowledge cutoff: 2024-06 Current date: 2026-04-24 You are an AI assistant accessed via an API.
Memory bandwidth is the biggest L on the dgx spark, it’s half my MacBook from 2023 and that’s the biggest tok/sec bottleneck
show us the benchmarks with "adaptive thinking" turned on
the MDM profile requirement is suspect though I get why they are doing it. but it doesn't inspire confidence to see that their profile is unsigned and still using the default micromdn scep challenge...
it should, lume is a thin wrapper around Apple's Virtualization.framework as i understand it
starting with M3+ you can use Hypervisor.framework/Virtualization.framework to spin up nested VMs. it would be amusing if that bypassed the limit.
I tried the periodic table in their examples using sonnet 4.6 on the $20/mo plan. After a few minutes Claude told me it reached the max message length and bailed. I pressed continue and eventually it generated the…
claude models with 'extended thinking' toggled answer very quickly and the quality of the answer is far ahead of what gpt 5.2 'instant' provides. i wont even bother using the non-thinking version of chatgpt because the…
“After creating a new account, I can confirm the quota drains 2.5x–3x slower. So basically Max (5x) on an older accounts is almost like Pro on a new one in terms of quota. Pretty blatant rug pull tbh.” lol
claude has half the context window size of codex and blows through a good percentage of it right off the bat by injecting a system prompt the size of don quixote
it sounds like the data can be involuntarily disclosed to an external third party (the attacker’s domain) purely because someone reviewed logs that auto-load remote images their log viewer renders the markdown and their…
i thought the article was going to go there, just redirecting the host to a self-hosted ip address serving the bin, but i was pleasantly surprised it didn’t! interesting to learn about the patching process and tooling…
like leveling to 99 in old school runescape
its possible to use gpt-5-high on the plus plan with codex-cli, its a whole different beast! i dont think theres any other way for plus users to leverage gpt-5 with high reasoning. codex -m gpt-5…
it only requires exponentially MORE money for linear returns!