80 comments

[ 0.23 ms ] story [ 22.5 ms ] thread
> This is necessary as (a) the 5h limit allows us to smoothen the load on our compute, allowing to keep the plan generous in terms of weekly usage

Seems fair, but then maybe the 5h usage should also drain slower during off-peak hours? Perhaps it already does.

> and (b) users on the Plus plan are relatively casual and new users, but then also just accidentally eat through their whole weeks usage and then are confused, making it not a great experience.

Wouldn't this one be fixed by making the 5h limit a guardrail you can opt out of? If so, then this doesn't work as justification for a mandatory limit.

It feels like daily quests in MMOs, trying to force user engagement via cultivating addiction-style behavior patterns
This is such a bummer for those of us living in third world countries. I’m not a casual user, I just can’t afford to pay more.
Aren’t you better using other services than OpenAI if the price matters to you?
You might be thinking of API pricing? The subscriptions give you a lot.
Oh right, thanks for catching that
I’ve tried Claude Code and recently switched to Codex. I’ve found that $20 on Codex gets me 4–5 times more usage than on Claude Code.

I could burn through my Opus usage in 1 hour whereas with Sol, I can go 3–4 days.

Oolama? That plan is more than generous for $20.
From my experience:

* OpenAI was ~$170 per week in value

* Claude was around $500 per week in value.

* Ollama Cloud is giving me about ~$160 per week in value.

So right now (thing constantly change in the AI world), its not more generous then Claude.

You can't burn through a week's opus usage in 1 hour. You might be comparing the 5-hour limit of opus with the week limit of codex?
Personally I have been using Deepseek V4 flash and it is enough for my needs. You can use OpenCode Zen/Go and also get access to many other models, including OpenAI ones. Writing purpose specific agents can make cheaper models close the intelligence gap with the most expensive ones
There are many providers and harnesses. You could go almost 90% of the frontier models for the fraction of the price. You have 3 more free days of ox alpha.

Geofencing has been quite good at creating feeling of being ripped off in the developed countries - from which the bulk of the companies revenue come from - that companies seems to be dropping it lately.

Anyway - China's advances seems to have caught the big 2 unprepared, so price war seems to be going on.

And there are Chinese resellers for frontier models if you are desperate for high end frontier tokens.

Given that it's not possible to buy the USD 100,- or USD 200,- subscription on a team/corporate plan this effectively introduces the 5h limits again for all corporate users.

At this point the 40 USD/120 USD Cursor Team subscription suddenly becomes a lot more attractive. You'll get more generous limits and a single monthly limit.

[delayed]
No point in using /fast mode if you are gonna get stopped after 5 hours.
Fast mode doesn't use subscription usage (IIRC) so it's not going to be limited to 5hr.
I say its good because you can atleast have a break in between sessions
Haha. I've been pretty happy using similar limits on Claude precisely for this. It provides a really nice work:life balance without that nagging feeling that I should really get a few more hours in.
I can understand Plus to some extent but why would Work users be considered casual or new?
So we must move to 5 hour working days. Thanks OpenAI!
(comment deleted)
probably less - I sometimes manage to burn like most of my 5h credits on Claude 5x Max plan in 2h.

The annoying things is that you gave bigger tasks and is such session finished maybe 70% and after 5h pool resets cache will expire so that if you "please continue" it will wipe out again a lot of your 5h pool)

Lets normalize working 5 hours a day. :)
I’m not a user myself but find it not great for your customers to announce such a change for the next day. Cannot they announce a few days in advance?
I really don't mind blowing my weekly usage in one day, but getting my tasks stopped and waiting for 5h cap to finish - is just annoying.
Please don’t! Don’t limit my daily usage
"During this period, the company also reset users’ weekly usage early on several occasions to celebrate milestones"

Such a casino vibe.

also it sounds like users benefit but they mostly don't. resets mostly benefitted OpenAI and only gave users few additionals hours.
Hey I work in the casino industry and we have regulations!
>> During this period, the company also reset users’ weekly usage early on several occasions to celebrate milestones, including adding another million active users across the two (now unified) platforms.

> Such a casino vibe.

No no no! It means we're all in it together! Feel the parasocial vibes!

The vibe was people goofing around on Twitter. There's one guy making the call and the reset is kind of a meme.
I think it's great that oai resets users limits. What is this complaint?

Free drinks? Keep customers happy? Reward your heaviest users? Casino vibes

I think there can be many complaints about large scale AI providers but given that large companies are now pushing for local models it seems that we are luckily avoiding the situation where AI is a utility
I've seen a lot of people who seem to think it's inevitable that they'll get resets, making them more likely to over-spend early because they'll get a reset anyhow.

It could leave them out of availability later when they need it.

I used to think it was pretty great but now it just means you can't plan your usage across the week.

It's a slightly entitled complaint but when you 'waste' 80% of your remaining usage because you get a reset you weren't expecting, there's a ping of disappointment there.

If you don't like it just ignore it. You're just trying to game the system and extract the absolute most you can and you're upset that this thwarts your game and doesn't benefit you. People like you are why we can't have nice things
Yeah. I switched to codex 1.5 months ago to try it, and for some reason the "vibe" kept me here. It was really the unpredictable rewards from the constant resets though. I am not convinced it's a better product or that sol + any combination of luna is better than Claude.

I'll be switching back.

They now give out resets you can choose when to use, but with a "best before" date. With Anthropic you don't even know what model you'll have access to next week. A few months ago I wanted to try a subscription to Gemini but I couldn't figure out how to give them money. The other one thinks he's Mecha Hitler.

To an extent I get it, this is not like a normal SaaS where you just buy seats, but their commercial offerings, especially B2C ones, feel like complete amateur hour. If they were selling literally anything else I'd have ran for the hills a long time ago.

Imagine paying for this shit..
I left Claude because of this. I guess now I need to move on from OpenAI also
But where will you go?
I really don’t get why people use the native chat apps when you can use all models via their APIs with no restrictions.

It’s more expensive, yes, but I assume most people here use it for something that is worth spending money on?

From my rough understanding, tokens are about five to ten times more expensive than the prepaid plans
I get that it’s not nothing, but from personal experience it makes me just so much faster.

I’m basically replicating myself multiple times for a ~10% salary surcharge.

In my mind it’s impossible to not see it that way, unless you just want to capture the subsidy surplus for free.

In the last 30 days my LLM spend, if paid at API rates, would be $19,431 (and that doesn't track online chats or my Github Action reviews). Now I wish I was making $200K/mo but I'm not, no where close. I got all that for $220 on subscriptions.

Yes, LLMs are awesome and let me do things faster but I would be using them far less if my only option was the API. I used coding agents prior to subscriptions (Aider) and spent <$200 total before abandoning it due to cost. The results were good, but not worth the price for me.

It seems you’ve stumbled into the truth: a lot of things are only worth throwing LLMs at as long as the LLMs are cheap.
I see this as evidence that last weeks token burn rate wasn’t by accident or due to some “bug.”

I don’t mind the 5 hour window but this is a reminder to not hinge your workflows on any one tool.

Stupid question: Why are y'all not using the API? It’s more expensive, but you can do what you want.
It will get worse!

Tighter limits, forcing people to higher plans. Most will go as they are addicted and dependent on those tools. They need to use to finish the project they started as no human will touch that pile of code. If you didn't saw this coming, as it has been the standard business strategy from all kinds of SaaS in the last years, you haven't been paying attention.

There are solid, cheap options like xAI, z.AI, MiniMax, Kimi, DeepSeek, and Alibaba. Heck, I can run a coding agent on my 32GB MacBook Pro M1 right now and it's actually pretty good (just slow). No one is forced to keep paying OpenAI and Anthropic. That's their major strategic risk: distilled models are far cheaper to build and run and they're only a few months behind.
This sucks, but really the value of the $20 plan has been insane. I routinely get weeks of work out of the weekly limit, so yeah I'll just move up plans if needed.
Unlocking the weekly limit coming soon for $9.99.
Is there a world where we look back at how AI usage is charged today and we equate it with how we had minutes on AOL and how absurd it seems looking back?
AOL minutes seem antiquated, but not absurd.
I think without a doubt that will be the case, unless the trend of compute getting better over the last 60 years suddenly stops. It should become less tough to run a local model and our devices should become more powerful.

However I think there is still a significant runway for these models to scale, so there will always be some sort of offering from providers. I can't imagine that our current use of the context window will be how that looks in a handful of years.

Absolutely. LLM providers are likely to be the next telecom providers. Low and steady income, translating to safe dividend income for investors. Most of these companies have no moat, since most LLM implementation strategies more or less converge to the same result.
Absolutely. Tokens will flow like electricity, in the instances when you aren't just paying for electricity to run models locally. The token providers (well, at least the frontier labs) will fight it tooth-and-nail with the example of telecoms being "how to become a commodity" but as long as open models continue they will have no choice.

Any other outcome would be pretty terrible.