> This is necessary as (a) the 5h limit allows us to smoothen the load on our compute, allowing to keep the plan generous in terms of weekly usage
Seems fair, but then maybe the 5h usage should also drain slower during off-peak hours? Perhaps it already does.
> and (b) users on the Plus plan are relatively casual and new users, but then also just accidentally eat through their whole weeks usage and then are confused, making it not a great experience.
Wouldn't this one be fixed by making the 5h limit a guardrail you can opt out of? If so, then this doesn't work as justification for a mandatory limit.
Personally I have been using Deepseek V4 flash and it is enough for my needs. You can use OpenCode Zen/Go and also get access to many other models, including OpenAI ones. Writing purpose specific agents can make cheaper models close the intelligence gap with the most expensive ones
There are many providers and harnesses. You could go almost 90% of the frontier models for the fraction of the price. You have 3 more free days of ox alpha.
Geofencing has been quite good at creating feeling of being ripped off in the developed countries - from which the bulk of the companies revenue come from - that companies seems to be dropping it lately.
Anyway - China's advances seems to have caught the big 2 unprepared, so price war seems to be going on.
And there are Chinese resellers for frontier models if you are desperate for high end frontier tokens.
Given that it's not possible to buy the USD 100,- or USD 200,- subscription on a team/corporate plan this effectively introduces the 5h limits again for all corporate users.
At this point the 40 USD/120 USD Cursor Team subscription suddenly becomes a lot more attractive. You'll get more generous limits and a single monthly limit.
Haha. I've been pretty happy using similar limits on Claude precisely for this. It provides a really nice work:life balance without that nagging feeling that I should really get a few more hours in.
probably less - I sometimes manage to burn like most of my 5h credits on Claude 5x Max plan in 2h.
The annoying things is that you gave bigger tasks and is such session finished maybe 70% and after 5h pool resets cache will expire so that if you "please continue" it will wipe out again a lot of your 5h pool)
>> During this period, the company also reset users’ weekly usage early on several occasions to celebrate milestones, including adding another million active users across the two (now unified) platforms.
> Such a casino vibe.
No no no! It means we're all in it together! Feel the parasocial vibes!
I think there can be many complaints about large scale AI providers but given that large companies are now pushing for local models it seems that we are luckily avoiding the situation where AI is a utility
I've seen a lot of people who seem to think it's inevitable that they'll get resets, making them more likely to over-spend early because they'll get a reset anyhow.
It could leave them out of availability later when they need it.
I used to think it was pretty great but now it just means you can't plan your usage across the week.
It's a slightly entitled complaint but when you 'waste' 80% of your remaining usage because you get a reset you weren't expecting, there's a ping of disappointment there.
If you don't like it just ignore it. You're just trying to game the system and extract the absolute most you can and you're upset that this thwarts your game and doesn't benefit you. People like you are why we can't have nice things
Yeah. I switched to codex 1.5 months ago to try it, and for some reason the "vibe" kept me here. It was really the unpredictable rewards from the constant resets though. I am not convinced it's a better product or that sol + any combination of luna is better than Claude.
They now give out resets you can choose when to use, but with a "best before" date. With Anthropic you don't even know what model you'll have access to next week. A few months ago I wanted to try a subscription to Gemini but I couldn't figure out how to give them money. The other one thinks he's Mecha Hitler.
To an extent I get it, this is not like a normal SaaS where you just buy seats, but their commercial offerings, especially B2C ones, feel like complete amateur hour. If they were selling literally anything else I'd have ran for the hills a long time ago.
In the last 30 days my LLM spend, if paid at API rates, would be $19,431 (and that doesn't track online chats or my Github Action reviews). Now I wish I was making $200K/mo but I'm not, no where close. I got all that for $220 on subscriptions.
Yes, LLMs are awesome and let me do things faster but I would be using them far less if my only option was the API. I used coding agents prior to subscriptions (Aider) and spent <$200 total before abandoning it due to cost. The results were good, but not worth the price for me.
Tighter limits, forcing people to higher plans. Most will go as they are addicted and dependent on those tools. They need to use to finish the project they started as no human will touch that pile of code. If you didn't saw this coming, as it has been the standard business strategy from all kinds of SaaS in the last years, you haven't been paying attention.
There are solid, cheap options like xAI, z.AI, MiniMax, Kimi, DeepSeek, and Alibaba. Heck, I can run a coding agent on my 32GB MacBook Pro M1 right now and it's actually pretty good (just slow). No one is forced to keep paying OpenAI and Anthropic. That's their major strategic risk: distilled models are far cheaper to build and run and they're only a few months behind.
This sucks, but really the value of the $20 plan has been insane. I routinely get weeks of work out of the weekly limit, so yeah I'll just move up plans if needed.
Is there a world where we look back at how AI usage is charged today and we equate it with how we had minutes on AOL and how absurd it seems looking back?
I think without a doubt that will be the case, unless the trend of compute getting better over the last 60 years suddenly stops. It should become less tough to run a local model and our devices should become more powerful.
However I think there is still a significant runway for these models to scale, so there will always be some sort of offering from providers. I can't imagine that our current use of the context window will be how that looks in a handful of years.
Absolutely. LLM providers are likely to be the next telecom providers. Low and steady income, translating to safe dividend income for investors. Most of these companies have no moat, since most LLM implementation strategies more or less converge to the same result.
Absolutely. Tokens will flow like electricity, in the instances when you aren't just paying for electricity to run models locally. The token providers (well, at least the frontier labs) will fight it tooth-and-nail with the example of telecoms being "how to become a commodity" but as long as open models continue they will have no choice.
80 comments
[ 0.23 ms ] story [ 22.5 ms ] threadSeems fair, but then maybe the 5h usage should also drain slower during off-peak hours? Perhaps it already does.
> and (b) users on the Plus plan are relatively casual and new users, but then also just accidentally eat through their whole weeks usage and then are confused, making it not a great experience.
Wouldn't this one be fixed by making the 5h limit a guardrail you can opt out of? If so, then this doesn't work as justification for a mandatory limit.
I could burn through my Opus usage in 1 hour whereas with Sol, I can go 3–4 days.
* OpenAI was ~$170 per week in value
* Claude was around $500 per week in value.
* Ollama Cloud is giving me about ~$160 per week in value.
So right now (thing constantly change in the AI world), its not more generous then Claude.
Geofencing has been quite good at creating feeling of being ripped off in the developed countries - from which the bulk of the companies revenue come from - that companies seems to be dropping it lately.
Anyway - China's advances seems to have caught the big 2 unprepared, so price war seems to be going on.
And there are Chinese resellers for frontier models if you are desperate for high end frontier tokens.
At this point the 40 USD/120 USD Cursor Team subscription suddenly becomes a lot more attractive. You'll get more generous limits and a single monthly limit.
https://openai.com/index/premium-seats-chatgpt-business/
The annoying things is that you gave bigger tasks and is such session finished maybe 70% and after 5h pool resets cache will expire so that if you "please continue" it will wipe out again a lot of your 5h pool)
Such a casino vibe.
> Such a casino vibe.
No no no! It means we're all in it together! Feel the parasocial vibes!
Free drinks? Keep customers happy? Reward your heaviest users? Casino vibes
It could leave them out of availability later when they need it.
It's a slightly entitled complaint but when you 'waste' 80% of your remaining usage because you get a reset you weren't expecting, there's a ping of disappointment there.
I'll be switching back.
To an extent I get it, this is not like a normal SaaS where you just buy seats, but their commercial offerings, especially B2C ones, feel like complete amateur hour. If they were selling literally anything else I'd have ran for the hills a long time ago.
It’s more expensive, yes, but I assume most people here use it for something that is worth spending money on?
I’m basically replicating myself multiple times for a ~10% salary surcharge.
In my mind it’s impossible to not see it that way, unless you just want to capture the subsidy surplus for free.
Yes, LLMs are awesome and let me do things faster but I would be using them far less if my only option was the API. I used coding agents prior to subscriptions (Aider) and spent <$200 total before abandoning it due to cost. The results were good, but not worth the price for me.
I don’t mind the 5 hour window but this is a reminder to not hinge your workflows on any one tool.
Tighter limits, forcing people to higher plans. Most will go as they are addicted and dependent on those tools. They need to use to finish the project they started as no human will touch that pile of code. If you didn't saw this coming, as it has been the standard business strategy from all kinds of SaaS in the last years, you haven't been paying attention.
However I think there is still a significant runway for these models to scale, so there will always be some sort of offering from providers. I can't imagine that our current use of the context window will be how that looks in a handful of years.
Any other outcome would be pretty terrible.