← All transcripts

Does the $200 Codex plan suck now? Transcript, AI Summary & Key Points

Theo - t3․gg · yesterday · Science & Technology · 33:33 · EN

Watch on YouTube

Answer

The $200 Codex plan doesn't uniformly suck. As 'the Astra plan' — buying it mainly to run GPT-6 Astra — yes, it does, and the Claude plan beats it there. But 6.1 Soul is so cheap that the halved cash value still yields roughly 5x more tasks than before, so for real-world usage (email/computer use, PR review via Claude) the plan still gives plenty of value; the Claude plan is still the better deal for shipping serious engineering work.

AI Summary

OpenAI's $200 Codex plan delivers far less than it used to: Tibo announced shortly before Dev Day that usage is calculated differently, roughly halving the cash value. The plan that used to yield up to $12,000/month of inference now yields about $2,000/month — the author's own 7-day usage came to $573.86, and he killed a whole account in 29 hours. The real story is model economics: 6.1 Soul is so cheap ($2/M in, $10/M out, 10¢/M cached reads) that the halved subsidization still buys far more tasks than before, making it arguably the best value a model has ever been — but GPT-6 Astra is expensive, and the new Ultrafast mode multiplies its cost so brutally that a $500 plan's weekly limit can be exhausted in about 2.1 hours (~$1,000/week of Ultrafast API value). The Claude plan remains the better value for shipping code (Opus 5.5 on the $200 plan is unbeatable, and Anthropic still gives a great deal), while the Codex plan shines for email/computer use and as a reviewer called from Claude via T3 Code. OpenAI compounded the damage by launching the $500 plan, Ultrafast, and the usage cut all at once — an optically disastrous rollout that burned the goodwill 6.1 Soul's efficiency should have earned.

Key Points

  • A $200 Codex plan used to get up to $12,000/month of inference; after the usage-calculation change announced by Tibo before Dev Day it effectively gives half as many dollars, and the author's actual 7-day usage was $573.86 (~$2,000/month), roughly a 10x drop.
  • A week of the author's API-equivalent spend across his machines runs about $20,000 — impossible to pay out of pocket, which is why subscriptions exist.
  • Both the subsidy number and the API price are 'fake': compute and energy cost for $100 of tokens is closer to $2–$5, with historical margins around 95%.
  • The Claude subscription only works inside Claude Code — you can't point it at Pi, Open Claw or Hermes, and Anthropic is 'that petty': mimicking a tool in Claude Code's system prompt makes the request fail. OpenAI's Codex subscription works anywhere as long as you're not reselling user traffic; running it on CI to review PRs is mostly allowed.
  • Anthropic's Fable limits usage to half your subscription allocation; OpenAI lets Astra consume your whole weekly limit, deliberately not adding a custom Soul limit.
  • Project Glasswing / Claude Mythos preview pricing was $25 per million tokens in and $125 per million tokens out, presumably keeping Anthropic's 95% margins; Fable 5/Mythos 5 launched at $10/M in and $50/M out instead, cutting margins from 95% to 87.5% and per-token profit from $118.75 to $43.75 — a deliberate move to make OpenAI's life harder.
  • OpenAI released Astra at the exact same price as Fable ($10/M in, $50/M out), showing its pricing is competitor-based rather than margin-based.
  • GPT-6.1 Soul launched at $2/M in, $10/M out with cached reads halved from 20¢ to 10¢ per million; with cached reads making up about half of OpenAI model costs, effective real cost is roughly 5x cheaper than planned, dropping margins to an estimated ~50% — a model the author calls the best value ever.

AI in practice

Used for

Agents

  • Stagehand demo agent — Fetch the latest releases from three GitHub repos by controlling a real browser. 2 held

Tools & resources

5 items

CNo. 0021
AIAINotes.us AI product

Claude Code

Open source · anthropics/claude-code

Claude Code is Anthropic's agentic coding tool for the terminal, IDEs, and GitHub. It uses natural-language commands to understand a codebase, create and read files, execute commands, run tests, explain code, manage Git workflows, and handle routine development tasks. It can also load persistent project context, run custom slash commands, use plugins with custom commands and agents, and operate with configurable autonomy while leaving actions such as final pull-request merging to a human. The official repository documents installation for macOS, Linux, and Windows, and identifies npm installation as deprecated.

TypeScript
Stars
★ 149,337
Forks
25,438
GNo. 1509
AIAINotes.us Tool

General Translation

Open source · generaltranslation/gt

General Translation is an open-source, full-stack internationalization and localization toolkit for React applications and other frameworks. It translates complete React components through a simple `<T>` wrapper rather than requiring content to be refactored into translation dictionaries, and supports text, variables, and currencies. The suite includes packages for Next.js, React, Vue, JavaScript, continuous localization, build integration, Sanity Studio, and an MCP server, as well as Locadex, an AI agent for automating internationalization in complex codebases. Projects can be initialized with `npx gt@latest` and connected to the service with an API key.

Mentioned in
5 videos
Kind
Other
ONo. 5304
AIAINotes.us AI product

OpenAI Codex

openai.com

OpenAI Codex is an AI coding agent offered through subscription plans. The videos describe using it for code review, email-related tasks, and computer-use tasks, and discuss subscription tiers and an Ultrafast mode.

Mentioned in
1 video
Kind
AI
SNo. 5095
AIAINotes.us AI product

Stagehand

browserbase.com

Stagehand is an open-source framework from Browserbase for building and evaluating agentic web-automation systems. It is used with browser sessions to let AI agents interact with web pages and complete tasks across the web.

Mentioned in
2 videos
Kind
AI
TNo. 1299
AIAINotes.us AI product

T3 Code

Open source · pingdotgg/t3code

T3 Code is an open-source agent-harness control surface for coding agents running on a developer's computer. It provides iOS and Android mobile apps, a web app, and an Electron-based desktop app for controlling locally configured Codex, Claude Code, Cursor, Grok Build, and OpenCode agents while using the user's existing provider subscriptions. The project can be launched with `npx t3@latest`, which starts its backend and local web interface on the machine. It also distributes desktop builds for Windows, macOS, and Arch Linux, and supports remote access from a phone or another machine. The repository describes the project as early-stage and warns that bugs are expected.

TypeScript
Stars
★ 25,759
Forks
6,631

Links mentioned

🔒 Full analysis locked

Unlock more videos and the full analysis

A credit unlocks one video's full analysis for good — the build steps, the tools and how each was used, the methods behind every use case. Pro opens the whole library instead, and raises how many videos you can analyse a day.

Unlock full analysis — free

Transcript

Searchable transcript of Does the $200 Codex plan suck now? — Theo - t3․gg (33:33). Search for a phrase, then click its timestamp to jump straight to that moment in the video.

Captions sourced from the original video on YouTube, published by Theo - t3․gg. The video, its captions and all related intellectual property remain the property of their respective owners; AINotes claims no ownership. Provided for research, accessibility and search — see the Transcript Notice and Copyright Policy.

00:00 Someday, far into the future, my children will be using their Sama coins to try and buy dinner. And I'll have to explain to them that when I was a kid, we used to use US dollars in order to get our Sama coins. And we used to pay $200 for $1,200 of Sama coins. I know I'm meing pretty hard here, but seriously, just a few weeks ago, a $200 plan on your codec sub would get you up to $12,000 a month in inference.

00:24 And on the Claw plan, it wasn't quite as much, but it was still 200 bucks in and $8,000 out. We got here in a way that was honestly kind of understandable. The $200 plans were introduced to feel unlimited back in the day where we would send maybe 10 or 15 prompts on our heaviest days and we would get a response back in under five minutes, usually just touching a few files here and there.

00:50 And when we were all using Sonnet to explore our code bases, $200 plans were practically unlimited. But as models got more and more capable and more and more expensive, neither OpenAI nor Anthropic wanted to take the optics hit of somebody being happy with how much usage they get for 200 bucks just to have it stop feeling unlimited when the new model comes out.

01:12 that has resulted in this crazy arms race where Anthropic and OpenAI kept subsidizing these plans more and more until we got to these crazy numbers. Right before Dev Day, Tibo made the biggest announcement that OpenAI has made arguably this year for us vibe coders. He announced that they would be changing how they calculate usage for the $200 plan. In effect, giving you half as many dollars as you would have had before.

01:34 That's pretty bad by itself, but I ran the numbers. I only got $570 out of a week of usage, which adds up to around two grand a month. So, is that the end? Did we really just go from 12K a month of usage to 2K and just smile and wave? Is there any reason to use the codec sub and not the quad sub? Or is this just some secret plot to force us all onto the new $500 tier instead?

02:00 And I can't wait to tell you those answers right after a quick break for today's sponsor. Agents need access to the internet to do tasks well. It's just a fact. It doesn't matter how smart the model is. If it doesn't know how to get to the internet and reference the things it needs, it's not going to do a very good job. That's why your agents need a great browser.

02:17 Ideally, today's sponsor, Browserbase. These guys built the perfect browser for your agents so that they have access to the entire web. And I don't mean just the browser. I mean all the different APIs, search capabilities, and more that your agents need to do all of this. None of that is my favorite part of browserbased, though. I've said for a while that Stageand is the coolest thing they're building.

02:34 If you're not familiar, Stage has an open source SDK that allows your agents to do things on web pages with plain English instead of having to define every single element that they're interacting with. I know showing code isn't the cool thing to do nowadays, but look at what I mean here. You can grab fields from the page by awaiting stage hand.observe description like find the email input or find the password input.

02:56 And once you've identified these fields, you can fill them with different variables and whatever else you want to do. This is so useful, especially when you combine it with their ZOD implementation that lets you get validated data out of a page. This is great for you and me writing the code, but let's be real, neither you nor me are writing the code.

03:10 That's why we built a demo with agents showing just how powerful this is. This demo is going to fetch the latest releases from these three different repos on GitHub. We click scan. We can see it opening and navigating in a real browser because this is actually controlling a browser. It's not just editing HTML for you. And now using those plain English instructions is able to find the right page and get the data we want.

03:30 What's even cooler with this is since it's not fixed on specific elements if the page gets updated in the future, you don't have to rewrite all your code to deal with one element moving slightly on the page. And now we have all the results. A nice little brief that shows all of the releases for the repos we put here. Make your agents more resilient and get them better data at soy.link/browserbase.

03:51 So what's going on here? If OpenAI wants so badly to be seen as a better value prop than Anthropic, why are they giving us so much less than they used to? It's just over, right? Like there's no reason to have the $200 Codex sub anymore, right? There's layers to this one. The biggest issue that we need to jump on top of now is the gap in these numbers.

04:11 The number on the right seems just insane compared to the number on the left. They're basically giving us a crazy cheap number here that's almost like made up. And when you compare that to what these models actually cost over API, it's just it's nonsensical. And when you look at my numbers for the last 7 days, it seems even crazier. I'm doing 20 grand of API spend in 7 days across my machines.

04:33 It's nuts. Obviously, I can't pay that. Like, I'm doing okay, but I I don't have it in me to pay 20 grand a month to use these things over API. The reality is that not only is this $200 number made up, so is the API price on the right. Both of these numbers are fake, and neither of them represent the actual costs to the companies when they're running these things.

04:55 The number on the right is the number that they have chosen to charge enterprises hitting these things over APIs to see what they can get away with. And the margins on these numbers have historically been insane. Upwards of like 95% against the hardware and the energy costs. Obviously, they put a bunch of money into training the model, hiring all the employees, doing everything else.

05:16 But the actual compute and energy costs for burning $100 in tokens is closer to like $2 to $5 than it is to that $100 number. Since this number is so outlandishly inflated by these crazy margins, this number feels insane. and the gap between them feels even more insane. But in an ideal world with enough competition, Anthropic and OpenAI would both have enough pressure to lower the number on the right that the gap between the subscription subsidization and just paying for the model for what it costs them.

05:47 That gap would get collapsed relatively quickly. And it seems like we are at the start of that. Now, we will get to what I mean there in a second. But I also want to talk about the weird things that made the Claude sub feel worse before. The first thing, the one that pissed me off the most by far was that you could only use it in Claude Code. The Claude sub isn't a sub to do $8,000 a month of Claude.

06:12 The Claude sub is a sub to do a large amount of inference in Claude code itself. You can't use your O here other places. You can use it in T3 code because in T3 code we are just calling claude code directly through the agent SDK. But if you wanted to throw it into something like PI or Open Claw or Hermes, you can't because those things aren't Claude code.

06:32 And if you think you can be clever and add the instructions to work like a tool like Pi in your Cloud Code system prompt, good luck. The request will fail. They are that petty at anthropic. badly. I am hoping that they're going to lift their restrictions here in the near future simply because a lot more people are maxing out their subs. So, it matters less if they're doing it inside or outside of cloud code.

06:53 Regardless, they've been awful about that. OpenAI has been the opposite. You can use your codec sub for basically whatever you want, basically wherever you want, as long as you're not using it to serve user traffic. So, if I put my codec sub in the cloud and then charge you $5 to access my subscription, that's not allowed. But if I use my codec sub to run on CI to review everyone's PRs when they come in, that is mostly allowed.

07:13 Just wanted to draw the distinction here because what you're allowed to do with your Codex plan goes a lot further than what you're allowed to do with your Claude plan. Claude also has a restriction on the usage of the biggest model. Fable usage cannot use your whole sub because Fable's price is not what it probably was meant to be. I have two sets of models on the screen here.

07:33 set one opus 4.5 to opus 5.5 and GBD 56 soul plus GBD6 soul the second group is Fable 5 and 5.1 and GBD6 Astra what is the difference between these two sets of models it seems obvious right the bottom ones are the big expensive ones and the top ones are the small cheap ones you're also noticing something weird I said GBT6 so soul not 6.1 which side do you think 6.1 goes on for the comparison I'm trying to make here GBD61 soul's on the other side of this line.

08:03 And you might be thinking, "Oh, is that cuz it's a big model and the others are smaller? Are you saying 61 Soul is bigger and more capable than Opus or some other dumb thing? I know how the comments end to go." I'm putting 61 Soul here for a different reason. Remember Project Glasswing, aka the announcement for Claude Mythos, when Anthropic confirmed they had this crazy model internally that they weren't going to sell?

08:27 They actually were going to sell. They were doing this big commitment early, but they planned to continue offering Claude Mythos to participating companies at $25 per million tokens in and $125 per million tokens out. My assumption here is that these numbers were calculated using the previous margins. If Anthropic was used to having a 95% margin on their compute, so they would spend $5 and have a hundred come in for their power and compute costs.

08:56 they planned to maintain those margins here. Let's just play with that 95 number. That would put the actual cost for a million tokens out not at the 125 they claimed, but at 625. So, let's just like write out these numbers quick. So, say Mythos was going to cost 125 per mill out and that would be $6.25 in cost. That keeps them at their 95% margin. But you might be confused about those prices because that's not the price Fable came out at.

09:30 And Fable and Mythos are the same model. I know some people seem confused about that. I'm not going to entertain that right now. They are the same model because when Fable and Mythos 5 came out, it was a different price. Fable 5 Mythos 5 are being offered at $10 per million in and 50 per mill out. Less than half the price of the Mythos preview. It's not like they made the model smaller and dumber before releasing it.

09:51 They were very clear about that. They actually just made it bigger and more capable. So again, we'll copy this and say Fable 5/MOS5. Price is now $50 per mill out, but it's the same model size, same hosting, which means that their costs probably didn't go down, which means their margins have dropped pretty massively from 95% to 87.5%. If you were to look at this as the profit that they're getting at a given time, it's even more brutal.

10:19 where in this instance they would have profited $118.75, here they're only profiting $43.75. The margin drop is big, but the return drop way bigger. They used to spend $625 in compute and make $118. Now they spend the same in compute and they make $43. The line I drew here is the start of the end. What is it the start of the end for? the subsidization.

10:51 Anthropic chose with Fable 5 and Mythos 5 to price it way cheaper than they originally planned in order to just really make OpenAI's life harder. And they succeeded with that. That was the start of the panic at OpenAI for sure. And even if the change in margins doesn't seem that much, the change in the revenue is world's difference. And if you bump that up to what was possibly a 98% margin before, these numbers get even crazier.

11:18 The point I'm trying to make here is that the amount of revenue Anthropic makes per amount of energy and compute used has gone down meaningfully with Fable, but they chose to do that because the margins are still insane. Like they're still just printing money with this. It's just that the amount they're printing has gone down significantly. And this also forced OpenAI to do the same.

11:36 I don't know what their planned price was for Astra, but they released it at the exact same price as Fable. So, we have no idea like what they would have charged before Fable because they made the model with a certain price target in mind probably. But, they did a few other things on top of that in particular with the subs that made it really brutal.

11:54 They took this hit on the cost. So, they might have intended to charge a 100 bucks per million out. Instead, they're only charging 50. So, their margins are worse. Meanwhile, Anthropic doesn't even give you your full usage of your sub with Fable. You only get half your allocation for Fable. the other half can can be used for whatever other models, but you're not using 100% of your usage with Fable.

12:16 It's limited massively. OpenAI chose to not do that and they let you use Astra for your whole weekly limit, not just a portion of it. And if you're curious where that 50% cut came from with Anthropic, it came from them not doing it for the three days before the ban reading all the data. And also here, it's not exact, but $43 is half to a third of $118.

12:38 They're profiting half as much. So they give you half as much. Makes all the sense in the world, right? Let's keep going down this rabbit hole a bit because there's one other model here that does not seem to fit that well. 6.1 Soul. Considering the way things are currently priced in the industry, my honest assumption would be that six soul was kind of meant to replace Terara because nobody used Terara.

13:04 And eventually they would do something more like Opus between soul and Astra. They didn't do that. In fact, they spit out 61 soul suspiciously close to six soul, which suggests that this was not necessarily what the model was meant to be. As such, I'm going to pretend that they were looking at the existing numbers for the competition. We've now established hopefully that OpenAI's pricing is based less on their margins and more on their competitors because they set Astra's price to be literally identical to Fable.

13:35 So, they were probably planning when they were training to make 61 soul a little closer to Opus, not Opus 5.5 cuz it didn't exist yet, but for Opus 5, which was $5 per mill in and $25 per mill out. We will assume these are the numbers that they were operating with. The new model they were planning was going to be $5 per mill in 25 per mill out. Nothing special with cash discounts.

13:59 We'll assume the same 95% margin. That means for $25 out it would cost them around $1.25 which gives them $23.75 profit. Not bad. You spend $1.25, you get $2375 out. Pretty good deal. But that is not the price that this model came out at. 6.1 soul came out differently. It came out, it's $2 per mill in, $10 per mill out and massive cash cost decrease.

14:30 The cash reads are half as expensive as they were before. They used to be 20 cents because it would be a 90% discount. So it would go from $2 per million to 20 cents per million. It's now 10 cents per million. Doesn't sound like a big deal. 20 cents to 10 cents, 90% to 95%. But when you look at the percentage of your cost and see that for OpenAI models, half your cost is cash reads, now it's 25% of your cost because you just took half your cost and you cut it in half.

14:55 So it would have been 50 cents per mill red from cash is now 10 cents per mill red from cash. That's a 5x decrease from what they had originally planned here. Crazy. So the result here is that your actual effective costs using 61 soul aren't just 2.5x cheaper than they expected. It's probably closer to 5x cheaper than they expected. What I'm trying to say is their margins for this model are garbage.

15:21 This 95% no. I would legitimately guess their margins for this model are like in the 50% range. Just seeing how much slower it is than six soul. It was definitely planned to be 525. So 210 with the 10 cent cash reads is just insane. So, they destroyed their profit margins. Like, they're still making money. I'm not going to say like a 50% profit margin is bad, but it's pretty obvious that they aren't just printing money on the inference the way they were before.

15:47 Remember earlier I was talking about that 50% limit for Fable, how they cut your limit in half when you were using Fable models. OpenAI had said they didn't want to put an arbitrary limit on Astra. That's what happened here, but it would have been the other way. Since the model they cut the subsidization for the most was 61 soul, they didn't want to put a custom soul limit.

16:06 They could have done that. OpenAI absolutely could have put in a custom limit that was your soul only counts for half your usage. But instead, they're taking the opportunity to try and collapse this gap in the subsidization a bit so that your account sub doesn't feel quite as insanely subsidized compared to paying API prices. I don't think they should have announced this the time they did.

16:28 I love the attempts at transparency from OpenAI. They were pretty transparent here. But this did not come off great, especially because so many people are still using Astra. But this is where we have to dig into something a little more painful, which is the effective costs. I ran a bunch of real numbers for my day-to-day usage. Previously, I was seeing similar to those 12K numbers before, especially when you combined the fact that we got a reset every 3 days on average many months.

16:54 you were getting 24 grand if you were truly maxing out every single window that you had. Now it's closer to $2500 a month instead of $24,000 a month. Roughly a 10x drop because I only saw $573.86 of usage. And I was able to kill a whole account in 29 hours in a loop. Not great. But let's scroll down a little because I have things I want to talk about in this.

17:21 My usage. My GBD6 Astra usage was for around 1,800 responses. And in these 800 responses, it did 766,000 output tokens and it cost $534. It was API prices. 61 soul did almost as many responses. It's a little less, but it was like 1,500 versus 1,800. It's not that big a gap. More input tokens, fewer cash input tokens. I think this might have been a bug in one of my setups or something.

17:47 I haven't figured it out. Don't care to. Again, very similar number of output tokens. $3827. Do you understand? Roughly the same amount of usage greater than a 10x decrease. The reason they made these changes, all of these changes, is they wanted to make 61 soul a great value. They wanted to make this $200 sub feel unlimited again the way it used to all the way back in the 54 and 55 days.

18:10 And they wanted to make sure they didn't go bankrupt because of the insane amount of compute that you're getting for that amount of money. The reason they had to do that is because of people like me and a lot of y'all who are token maxing right now. Since I use these subs and I push them to their absolute limit, if they didn't make these changes, I would be burning two to four times more compute from OpenAI every month.

18:32 And that's just insane cuz you can absolutely and I I do genuinely believe for any even vaguely real world use case, you're going to get practically unlimited use of a $200 plan as long as you only use 61 soul and you never touch max cuz max makes no sense at all. Do I think this is a good change, though? I'm going to be so real with you guys. There is more to talk about here, especially the $500 plan.

18:55 I'm going to be incredibly real and honest with y'all. If your goal here is to use your subscription to write code that you plan to ship, the insane value you get out of Opus 5.5 on the $200 Claude plan is impossible to beat. I thought I would cancel subs with Claude because it was so much more efficient, but I ended up just pushing my own limits to what I could do.

19:16 And I ended up with a six cla account and if I refresh now, it's probably closer to drained. Yeah, like I have killed three of my six. Two of them were dead this morning and just reset. I am pushing these things. I am very much getting more value than I would ever have expected out of my clawed plans. It's crazy. And the craziest part here is I have not touched any of my fable.

19:35 I have 600% of my fable. I can't use it on the debt accounts because that gets combined with your weekly. So, I can't use it in those accounts. But, it's hilarious that I'm just straight up not touching Fable anymore. So, I still think the Claude plan is the better value by far. But, I also think 61 Soul is the best value a model has ever been by far.

19:55 If I was forced to pay API prices, I'd be picking Soul without any question. But, I'm not. You get less subsidization from the Codeex plan, but the model's also so much cheaper that it balances out relatively well. And I'll show you this with a new project that I actually built with Astro Ultraast, funny enough, Slopolytics. I wanted it to be easier to visualize the score against the dollar spent for these benchmarks.

20:18 And you can see very clearly here 61 soul is an absurd value on the artificial analysis intelligence index, which is what this data is from. If I switch to the linear view, it's even funnier looking. Truly hilarious. I also ran this on terminal bench myself because I was not happy with how others ran Terminal Bench. And not only did 61 soul get the best score on it when you gave it access to its native harness, I ran it with codec and cloud code, so I got better numbers than a lot of others did.

20:44 Not only was the 61 soul max run the best score ever. The 61 soul x high run pretty much exactly tied opus 55, but opus cost $1311 per task. 61 soul cost 96 cents per task. A 13th the price. Do you understand? your subscription might not be as much dollar value, but the number of tasks you are completing with that money is way higher. So, if you're paying API prices and you're paying Opus, you better be damn certain that Opus is the only model that does what you need.

21:18 Because if you're comparing the cash value you get in the subs by the paid tokens, none of this ends up making sense. And this is the problem is we have two levels of like subsidization happening here. We have the amount of dollars you're getting for your $200 sub and we have the cost that those dollars are presented to the lab and we have how many tokens does it take to solve a problem.

21:42 Because if I switch this to show per tokens which I have in slopalytics do I have no I have tokens for task here you'll see that medium on opus 55 sorry high on opus 55 uses roughly as many tokens as max does on 61 soul and the tokens are cheaper. So cost per task is what we should be looking at. And realworld cost per task on something like terminal bench 4 is a 13x decrease.

22:08 So the dollar value of the sub being half as much ends up still being five times more tasks depending on the work you're doing. So again, I don't think 61 soul is the better model. I would take my cloud sub over my codec sub any day. But this isn't some crazy rug pull like everyone's making it out to be. unless you insist on continuing to use Astra with the much higher cost and much higher spiky nonsense rates.

22:33 But here's where I have to stop being nice to OpenAI cuz I'm about to crash out really, really hard. And everybody who's accused me of being paid by OpenAI to say nice things, this is going to prove that I'm not, which means I do need to be paid. So, we're going to do one more quick sponsor break first. If I told you that you could 5x your potential users with one simple trick, you'd probably say that I'm insane.

22:53 And I understand because you probably didn't know about General Translation. These guys are incredible. They made it so easy to translate realworld applications across various different surfaces. There are so many little edges that are hard to get right when you build translations into your apps. Localization is just not an easy problem to solve. I had to do it myself at Twitch and it was miserably difficult.

23:16 I'm very thankful for my team of much smarter engineers for helping us all through that. Going to do a thing I don't get too much anymore and show you some code because it really shows how magical what they built is. In React applications, you use their helpers in your JSX to actually indicate what section should be translated. They also have helpers for handling annoying things like number formatting and pluralization, which is not trivial in lots of different languages.

23:35 Where things get much more powerful is their dictionaries that allow you to have consistent translations across various surfaces. So if you have a feature with a specific name, you don't have to worry about that feature being named incorrectly in three different places. Now you can finally know for a fact that the same translation and the same verbiage will be used across all these different surfaces.

23:53 We try setting this up for the T3 Code marketing site. And when I say us, I really mean our agents because agents are really, really good at using general translation. And after two prompts, we have everything we need to showcase our application in all sorts of different languages. And the best part is we can share these definitions across all of our services, whether it's our mobile app, our desktop app, our website, and more.

24:12 Reach the 85% of the world that doesn't speak English at soy./gt. Time for the crash out. I'm going to do more of a detailed version of this in a future video, but I'm going to give you the cliffotes of it now. The $500 plan was a mistake. Not the existence of a $500 plan. Doing it right now, absolute optical shitow. To tell people that the 20x plan is now 10x, but don't worry, you can get the 25x plan if you pay 2.5 times more pisses everyone off.

24:42 Myself included. This was just a mistake. If they were going to do this, which obviously they had to eventually, they should not have done these things at the same time. It is such an unnecessary self-inflicted L that if I defend anything they do, you guys are going to call me a paid show. And I understand because this is such an absurd blunder optically that it makes them look insidious.

25:06 The reality is that these subscription plans are not how they make their money. They make their money off the consumer plans like the $8 and $20 plan and selling API access to the labs, not from these $200 subs. This is not real profit for them. These are all about marketing. So, if this $500 plan did not help them with marketing, it failed. And I can comfortably say it failed.

25:25 There was a bigger mistake though, and this is the one that's going to get a dedicated video in the near future. Ultraast. All of the things I have said up until this point made sense for one reason. 61 Soul is so cheap that cutting your cash value in half does not meaningfully decrease the amount of value you're getting from your subscription. As long as you move to Soul, because the model's so much cheaper, that 50% cut doesn't feel like a cut.

25:51 Ultra Fast is the opposite. Astra is priced at a dollar per mill in cash, $10 per mill in normally, $50 per mill out. God forbid you have a bigger context window. Now you're paying double. Then you turn on ultra fast. You're not hallucinating. That just went from $10 in to $60 in. The cash input just went to $6 in. Cash rights are now $75 and the output tokens are $300.

26:16 And that's before the long context. God forbid you don't have a context window limit on which thankfully we all do in codeex inputs now $120 in cash inputs $12 in cash rights are $150 bucks and the output tokens are $450 per mill out. I had Astra on ultraast review two PRs. It cost $600 dollars. $600. We're now in the category where it would have been roughly the same cost to pay an engineer.

26:53 It would have been slower, but I could have paid them. So, OpenAI the same day put out the Pareto Frontier cheapest model for the value you get ever by far. Like, let me go back to Slopolytics and show you the Pareto Frontier really quick. They own it. Luna and Soul are so deep in the Pareto line that you can't even see them when I highlight it because they're just the Pareto line.

27:17 Opus has a section that isn't that you see highlighted there when I hover because low and medium are not worth using compared to just using 61 soul. So they put out the cheapest model for that level intelligence by far. Didn't put it on cerebrus cuz that's how ultraast runs. It's on different infra that is built for crazy inference and it is incredibly expensive for them to run it.

27:38 They just don't have the compute available and it's like special chips from Cerebra. So they're trying their hardest to make enough to support this. It's crazy costs. I get it. But they put out the cheapest thing ever and the most expensive thing ever the same day and the day before announced a cut, a massive cut in how much cash value you get with your sub.

27:55 This is all just pathetic. And the reason it happened Oh, Astro Ultra Fast is Nvidia. I thought that they were used. So they straight up just haven't gotten Cerebrris working. Have they just given up? That's why they did this tweet because Cerebrris didn't work out for Astro ultrafast. Damn, I didn't even know that. that the cost makes even more sense now cuz that means that they are like massively overprovisioning in order to make it that fast.

28:16 God. Yeah. No, the margins on the ultra fast stuff are Like way way lower. Apparently Jane Street bought all the Cerebrus chips and that's why they aren't able to use them. Fun. Good to know. Regardless, this sucks and makes the value of that $500 plan feel awful because they're branding it as the plan with ultraast. Ultraast is an ultraast way to burn through your limits.

28:39 I did a bunch of math with my $500 plan using ultraast. You get about $1,000 of ultraast API value per week on that $500 plan, which means from our estimates 2.1 hours of generation. Yeah, my weekly limit can be killed in 2.1 hours. This was a mistake. They probably shouldn't have put it in the subscription at all. And combining all of these things at once, the release of 61 soul with way lower margins that people don't understand.

29:09 Combine that with the limits being cut in the accounts so drastically before people understand this change. Combine that with a new $500 plan that obviously people are going to feel pressured to upgrade to. Combine that with ultraast, which is the fastest way to burn through that new plan. And any goodwill that could have been won by 61 soul's efficiency has effectively been burned even faster than my limits.

29:32 That's the problem. They did not release these things in an order that makes sense so that we could learn these changes. And I don't blame anyone for not understanding that 61 soul has way lower margins. And I'm also not surprised that there are people who upgraded to this new plan excited for the ultraast and then were blown away when their account was locked after like an hour or two of prompting.

29:56 All of these things suck. All of these things hurt the trust that OpenAI has. And I'll be real, almost all of these things are OpenAI's fault. Doesn't mean I don't think they can come back. Just means that Astros put them in a tough place right now. The biggest benefit OpenAI used to have is that they were better at post- training in RL, so they could make smaller models that were really, really capable.

30:14 Anthropic figured this out with the recent Sonnet and Opus releases, so that gap isn't really there anymore. And here we are. I guess Sam was right. I'm an Anthropic fanboy in the end. But yeah, the value you get on your account is a lot lower. My ultra fast usage on the $500 plan gets me around $4,000 a month, whereas just a few weeks ago, my $200 plan used to get me $12,000 a month.

30:41 So, my price got two and a half times higher and my usage got three times lower. That just feels bad. As long as you're thinking of it that way, it feels bad. I went from months ago being able to use 54 or even 55 literally all day and if worst case I'd hit like 10 to 15% of my usage in that window to not even surviving half a day with my $200 or even $500 plans.

31:01 It does feel bad at OpenAI needs to fix it. They need to fix it by making things that are actually better than Astra. So there's no reason to reach for that model. They need to make things that are as capable as Opus 5.5, that are as pleasant to use, that are way faster than they have currently for reasonable prices, that can feel unlimited in our subscriptions, and they need to never ever ever brand the $500 plan as the ultra fast plan.

31:27 That was a mistake from the front. I warned them ahead of time. I still think that was a bad decision. So, to answer the original question, does the $200 sub suck? Now, if you are thinking about it as the Astro plan, the way I used to think of Claude as the Fable plan, yes. If you're buying the $200 plan just to use Astra or you would buy the $200 Claude plan to use Opus and Fable, yeah, go get the Claude plan.

31:47 But if you can do both, I personally get a ton of value from having Claude spin up sole sub agents to review its work because OpenAI models I've still found are a bit more thorough with their analysis and will find things deeper than enthropic models tend to. So, I do a wield and the vast majority of my codeex usage is either me using it to like do my email and computer use or it's coming from Claude calling Codeex via T3 code in order to review Claude's work.

32:18 And for those tasks, I have found it to be incredible and I get plenty of value out of my $200 subs. But the Claude plan is absolutely mandatory in my opinion if you want to ship serious engineering work regularly. It's just such a good value and you get so much. But I do suspect in the near future as OpenAI and Anthropic get more and more competitive with enterprises and start driving those API prices down more and more that the amount of subsidization we see here will go down and people will complain.

32:47 There will be lots of crazy graphs people are showing on Twitter how much worse a value it is. But I promise you, you will still get more work done with the new models on the same plan in a year than you're getting from it today or something went really, really, really wrong. Hopefully, you now better understand the change to that $200 plan. Do I think everyone should rush to go sign up for right now?

33:10 No. Do I think the Claude plan is a better deal? Honestly, kind of. Yeah. But I also think 61 Soul is one of the best values in models ever. and you shouldn't discount that just because your discount rate is a bit lower than it was before. Let me know how y'all feel and how you think I'm still an OpenAI hardcore fanboy after a video where I tear them apart like this. I I know how this one will go. But until next time, peace nerds.