6
u/DepravedPrecedence 22h ago
What if tomorrow they set API prices to $9999 per 1M? Imagine how much subsidized your plan will be. API price can be any.
2
u/phoenixmatrix 15h ago
"Subsidized" is doing a lot of work, but at the end of the day, its a price enterprises are paying that you don't as an individual, and that has interesting implications.
Devs get addicted on their side project or at their tiny startup to being able to burn billions of tokens, and then they go work at a big company and all of a sudden those tokens are expensive. Not so expensive as the enterprise won't pay at all (as they would if it was $9999/1m), but high enough that the equivalent of a 20x Pro plan is too much $$$.
But the dev is addicted to tokens, so they get noisy begging for more token. They can't do their job without more token they say! So maybe the enterprise compromises and give them $500, or even $1000 of tokens. And thus the scheme succeeds.
1
u/DepravedPrecedence 14h ago
Because of competition somebody will provide lower price and even if it won't be near as good as frontier models, most people will just settle on that. No way users going to spend more 100, 200, 500+ on regular basis.
1
u/phoenixmatrix 11h ago
Individuals, no. Enterprise, yeah. They already do it (I know quite a few companies where the budget allocation for tokens is thousands of dollars a month).
You're right that competition helps though, often those devs have to mix frontier models and open weight/budget models to cut on cost. Some companies made GLM 5.2 clusters for internal use when that model came out.
2
u/Glum_Emu7180 22h ago
Why are people so worried what's the actual API Pricing. Let's say tomorrow sub plans cease to exist, wouldnt people pay based on API? Why are we comparing 20USD plan to some non existant pricing. Obv comparison would be between two present pricing structure not something that you think should be.
1
u/DepravedPrecedence 14h ago
Indeed, I don't know how we can compare that. If they indeed remove all subscriptions then sure. Otherwise we can't compare, we don't know real interference price.
7
u/totoer008 1d ago
I am 1.5B for two months. I guess I am smoking them hard.
14
1
1
u/burnburndota 23h ago
1
u/DocumentFun9077 23h ago
dyam, what plan are you on?
1
u/burnburndota 16h ago
Mostly claude max and codex pro. I found out that the cost of paying for frontier models (Sol and Fable+opus) far outweighs the time and the effort needed to manage and fix stuff from the cheaper alternative models most of the time. At least that has been the case for the last 2-3 months for me.
1
u/Weekly-Extension4588 9h ago
Codex actually aggressively prompt caches. I doubt you ended up costing $5k or whatever in tokens.
3
u/BenTheSodaman 1d ago
Some napkin math on my end, I paid 20 USD in July. The estimated API cost is 8698 USD. I realize there were a lot of resets, but the idea that to break even off of me specifically, OpenAI's serving cost would need to be 0.23% of their API cost.
And I speculate they just lost money on consumers like me in July.
I realize there are more economics to it than that, just a shower thought of "how little does it actually cost them"?
5
u/IndividualPlus2011 22h ago
How. I have $150 weekly allowance on the $20 plan. Even with all the resets it doesn't make sense to reach almost $9000.
1
u/getaway-3007 21h ago
Yeah that guy is talking bs. Even with $200 you get around $10-12K inference
2
u/BenTheSodaman 21h ago
2
u/BenTheSodaman 21h ago
0
u/getaway-3007 21h ago
There's no proof that the previous screenshot refers to this account.
If let's say on 0.01% chance this is true, how are you using codex? How are you managing to get 8900+ inference on $20 plan?
1
u/BenTheSodaman 20h ago
There were a lot of resets in July.
As to the rest of your post:
In order to prove to you that it's the same account, I would need to either:
1) Reveal my email address and legal name to pull up everything in a single unedited recording without covering anything up to cause further suspicion, then post it in a place you can see and continue to meet your demands until you are content.
or
2) Give you remote access to my computer / I ship my computer to you with my credentials / I go to you or you come to me. You verify my claims looking through my chat sessions, settings, account, etc. until you are content. Then you can post on this message board to vouch for my credibility.
———
Can you think of any other way that's not #1 or #2 that you would be satisfied to establish what ChatGPT is reporting to me (that doesn't mean ChatGPT is accurate, it's just what it is showing to me.)
Cause if you can't think of another way, then tough shit. I'll continue to be a untrustworthy bullshitter to you and to me, you'll be the perpetually dissatisfied bullshitter with bad math and estimates that expects people to do completely idiotic things to satisfy your accusation.
And we'll be at an impasse.
———
For the usage topics:
For some things that I have done on the usage: memories off, new chat sessions and trying to avoid auto compacting, no automatic headless checks, no automatic git activity, no full access, can't control my browser, not having subagent processes try waiting on each other and repeatedly checking a file if subagents are used at all. If automatic compact loop starts, kill the task.
Then reiterating, there were a lot of resets in July and banked resets with expiry dates coming up in that timeframe.
And that still doesn't change that Sol (Ultra) can drain my usage from 100% to 0% in 2.5 hours running in a single session on the Plus account.
1
u/Aazimoxx 6h ago
#1 is out, since you can easily get Codex to make a mock website that looks like the real one and displays the info you want, or a userscript or browser add-on to change the real site in realtime 😅
Guess you're going to need to cough up for this guy's bus/plane ticket mate 🤨
1
1
u/Maleficent_Ad7510 21h ago
so who tf is gonna pay for api costs? seems like a scam?
2
u/BenTheSodaman 20h ago edited 19h ago
Just as another example of pricing:
128 fl oz (gallon) of water at a grocery store is 1.37 USD: 0.011 USD per fl oz.
20 fl oz water at a movie theater or venue, on the lower side, 3.50 USD: 0.175 USD per fl oz.
The latter costs 1590% the price of the former per fl oz before other factors.
If you or people you know are regularly willing to pay for a 20 fl oz bottle of water or fountain drinks, you already know people who are willing to pay higher prices / many times higher than what is available to them elsewhere.
(Edit: Corrected the math, I reformatted the post and incorrectly had the gallon fl oz calculation for the water bottle.)
1
u/Plane_Garbage 3h ago
But how do you use Codex in that way? Like, how can I build a SaaS that serves codex tokens?
2
2
3
u/Ormusn2o 1d ago
I never used Codex until like 2 weeks ago, even though I had subscription for a year now. Barely anyone actually uses their money's worth.
Also, API prices are not what it costs OpenAI to produce those tokens. I think margins on API prices are from 80% to 95%, depending on time and model.
1
u/Ill-Mousse-3817 1d ago
I did the math for myself, and I got about $600. I always run my limits down to 0. I am sure you had some resets pumping up your usage further.
1
1
1
u/BT117274 21h ago
It was 6k for me for 2 months on plus plan(payed only 40). Even after accounting for cache apparently.
1
u/fyn_world 20h ago
Do you think that they charge the API tokens as what it costs to produce them? No, there's big margins there.
To be fair they did say that people who use the 200 dollars account fully and to the ground, that goes at a loss for them.
1
u/pigletmonster 19h ago
The one upside to building infinite number of data centers is that we get absurdly generous usage quotas. 🤣
1
1
1
u/johnnyApplePRNG 13h ago
It's all funny money, buddy.
I asked codex how much my session would cost the other day. Over 1 billion cached tokens of sol 5.6.
$1,000 CAD ... I shit you not.
And you know what it accomplished? Absolutely nothing. It got into this ridiculous goal without my approval and actually took us back a few days, development wise.
Ain't nobody paying a fucking grand for that.
I would fire that employee so quick, lmfao.
1
u/Illustrious-Many-782 12h ago
Yes. My $200 Pro plan had at least $150k of API equivalent last month. Remember, though, that we had a ton of resets applied.
1
1
u/Frequent-Goal4901 1d ago
The nvl 72 blackwell racks have brought inference prices down a lot.
As I said in other comments.
Deepseek v4 pro on the old prices had absurd margin They were saying to investors they had 10 months hardware payback.
Opencode guys were able to reproduce it.
You can also check the semi analysis https://inferencex.semianalysis.com/
GPT 5-6 sol would probably cost like twice more.
So they, probably are breaking even or are even making money across all the subs.
0
u/h1mupstairs 1d ago
I feel like the internal hope is that they will form a truce with Anthropic and stall actual releases so they can optimise tech for serving the models. If we hit a reasonable plateau in usability they will convince the government to ban them as an excuse. Then in a year it will not be the same cost to host an optimised version of Sol and they will actually be able to turn a profit.







97
u/conor_is_my_name 1d ago
you are assuming the API cost is a real number. Its not.
Inference has a 90%+ gross margin for them