HN user

LUmBULtERA

1,358 karma
Posts1
Comments415
View on HN

Given how OpenAI got rid of their 5-hour limits and reset weekly limits so often, is Kimi really undercutting them on effective price?

ChatGPT Work 13 days ago

Literally no UI changes when I toggle between the two. It's 100% identical, one is not less technical than the other.

ChatGPT Work 13 days ago

Also, when you toggle btween ChatGPT Work and ChatGPT Codex, nothing changes. This is super confusing.

This is the weirdest thing. It's completely unclear even if I was using the app for "work" rather than "coding" why I wouldn't just use the "Codex" option. Why even bother toggling? Can Codex actually not do any of the things in "Work"?

GPT-5.6 13 days ago

They mean the same thing as GPT-5.6-Max, GPT-5.6-Plus, and GPT-5.6-Fast but require base knowledge for an analogy.

But do they though? When do you use GPT-5.6-Max-Low vs. GPT-5.6-Plus High? Or GPT-5.6-Fast-Xhigh? What's the Pareto optimal choice (outcome and price)? According to the benches it seems to bop around and the even if the benches are accurate the best choice isn't always consistent.

There could be a whole host of reasons. It may be during launch that compute is re-allocated from training to inference so that all users can try Fable. Soon that compute will re-allocate back to training until they can get more compute.

Claude Sonnet 5 22 days ago

Did Anthropic have Opus 4.8 and Sonnet 5 switched in the Agentic Search chart at first?

Claude Sonnet 5 22 days ago

That's yet to be determined. I think a lot of open-weight models are benchmaxxed and their usefulness for many tasks are not represented by those.

Fair to include A&G, but their marketing and training cost is about generating future growth. People say that the subscriptions are highly subsidized as if its fact -- it's certainly not a well-established fact. These people simply don't know the truth one way or another, but portray their not well supported opinion as fact.

My comment is about your statement "serving these tokens without paying for training is already expensive"...

One thing we do know from OpenAI's leaked financial document is that they are already profitable on inference, though that data is not broken down by cost and revenue of API vs. subscription. One important factor is that subscription inference can be optimized in ways to reduce cost (e.g., usage limits, batch optimization around API-prioritized inference, etc...). I think simply we do not know the actual cost of subscription interference for SOTA models.

Except there are plenty of inference providers worldwide (including the US) that serve open-weight models that are not subsidized, and are reasonable in cost. Or is your claim that those are all running at a loss?

3. We're massively overusing SOTA models. As long as you're on a subsidized subscription, you can use Claude Opus 4.8 high to write blog article meta descriptions. If you paid by token, you wouldn't do that.

This idea that the subscriptions are subsidized is repeated over and over, but I've never seen any proof of this. It seems to be entirely based on the inferred API cost the subscription usage could give you, but there are a lot of assumptions needed for that to follow.

GLM 5.2 vs. Opus 1 month ago

Subscription inference can also be cheaper than the cost of API inference if the provider wants it to -- providers can do flexible scheduling for subscription inference for example, around API inference, to lower its cost and get better utilization of the hardware.

GLM 5.2 vs. Opus 1 month ago

Fair enough, there is not strong specific evidence to the contrary except about overall inference being profitable for OpenAI (as well as the open weight model providers hosted throughout the world).

GLM 5.2 vs. Opus 1 month ago

My impression is that individual subscriptions are the loss leading hook

Except there is no evidence of this at all, just people comparing API and subscription pricing. The leaked financial info for OpenAI shows inference is profitable right now, though it does not show a distinction between subscription and API revenue... but if subscription revenue was so lossy, it would hard for total inference to still be profitable.