Quick answer: for still images, yes. xAI's own plan comparison table shows image generation checked under the Free column as of September 10, 2026, and video generation dashed. The most widely cited paywall story is a January 2026 restriction on X itself, and the pages still carrying it have not looked at the pricing page in months.
My read on Grok Imagine: the stills are the free part, the video is the paywall, and nobody selling you an alternative wants you to know the difference.
Published September 10, 2026. Plan rows verified on xAI's pricing page the same day.
Quick decisionYour situation
What to doYou want still images and you pay nothing
Use it; xAI's table checks Imagine on FreeYou want video
That is the paid line, on every tier above FreeYou want 1080p video
That is the $100 tier, SuperGrok PlusYou want to know your limit
Nobody publishes one. Here is the meterYou are choosing a rung to pay for
Free vs SuperGrok, per triggerxAI's own table says images are on the free plan
This is the whole page, so I am putting the receipt first.xAI's pricing page, captured September 10, 2026. Image generation (Imagine): checked under Free. Video generation: dashed under Free, checked on everything above it.
Two rows, one conclusion. Stills are in. Motion is what you are actually buying.
That single dash is worth more than the ten blog posts above this one on Google, because it tells you exactly which half of the product is behind the wall.
The widely cited January 2026 story happened on X
So why does the entire first page of search results tell you Grok Imagine is not free?
In January 2026, X restricted Grok's image generation and editing for non-paying users on X itself, after a backlash over what people were making with it. Gulf News covered it at the time and noted something the aggregators dropped: people who were not paying could still reach Grok's image features through the separate app and website.
That nuance did not survive. "You cannot edit an image in an X reply anymore" became "Grok Imagine is paid," and the pages carrying that claim have been ranking ever since without anyone rechecking it.
Look at who is ranking for this question and you will see why nobody corrected it. Most of page one is either a post from before the change or a site selling you a different image generator, with the disappointment as the hook.
xAI's own pricing page contradicts itself, and you should know that
I am not going to pretend the answer is cleaner than it is.Same page, same day, scrolled up. The Free card lists four things. Images is not one of them. "Image and video generation" first appears on the $30 card.
So the comparison table says images are on Free and the plan card does not mention them at all. Both are on xAI's pricing page right now.
My read, and I want to be clear that it is a read: the table is the more specific artifact because it distinguishes stills from video row by row, while the card is marketing shorthand for what the $30 unlocks. The card is not wrong that images get better above Free. It just is not the place to learn whether the free plan can make one.
Here is what I checked and what I did not. I read xAI's published table on the day of writing, on my own paid account, and I screenshotted it. I did not create a free account, so I have not personally watched a $0 login press generate and get a picture back. If you are on free right now, you can settle it in ten seconds, and your ten seconds beats my inference.
What I would not do is trust a number. xAI publishes no free-tier image cap, so every "3 images a day" and "5 to 15 a day" figure you will find is someone's estimate rather than something xAI has stated. Paid plans have run on one shared weekly compute pool since June 2026.
And here is the third contradiction, which cuts against me rather than for me. xAI's FAQ says the free tier keeps its own limits on "Chat and Voice" that reset on their own schedule. It names Chat and Voice. It does not name Imagine. If you wanted to argue the free plan does not really get images, that sentence is your best evidence, and I would rather hand it to you than pretend the table is the last word.
What an hour on the paid version actually produced
I have a SuperGrok Plus account, so this section is the paid experience. I am not going to dress it up as a test of the free tier.
I typed one thing and got four.One prompt, four variations, with Generate More and Think Harder waiting underneath. The play button on each tile is the video path.
Four variations from one prompt is the part that changes how you work. You are not iterating on one image and hoping. You are picking.
There is also a Discover surface, which is where you go to find out what the thing is capable of before you have any idea what to ask it for.Discover, same session. "Generate a cartoon" with no other direction.
It will not make you copyrighted characters. Ask for Mario and it refuses, which is the correct answer and also not a real limitation for anyone making a header image.
I will say the thing I did not expect to say: I used to think this was trash. It is not trash.
One is hands-free, the other is a control surface
This is the distinction that actually decides which tab you open, and I have not seen anyone else write it down.ChatGPT images
Grok ImagineHow you drive it
Describe it, it decides
Toggle settings, dimensions, elementsClosest comparison
An assistant
MidjourneyGood when
You want it handled
You know what you want and want to steerThe cost of that
Less control
More decisionsWith OpenAI you say it and it generates. It is more included, more hands-free, and that is a genuinely good experience when you do not want to think about it.
Grok Imagine goes the other way. You can select different elements, change specific things, and get very specific about dimensions. It almost feels like a Photoshop.
Neither of those is a criticism. They are two different products for two different moods, and knowing which mood you are in is the entire decision.
Not everyone agrees with me, and here is who
I formed my impression in about an hour. Named people who spent longer landed in different places, and you should have both.
TechRadar tested the image editing features and published the verdict in the headline: they are fun but will not replace Photoshop any time soon. That is a direct contradiction of my "almost feels like a Photoshop" reaction, and I am leaving it in rather than picking the quote that agrees with me.
We are describing different things. I meant the interaction model feels like an editor rather than a slot machine. They mean it will not do your actual retouching work, and they are right about that.
Notebookcheck filed it under "where AI brilliance meets mild trauma," which is about as accurate a summary of using this product as three words can be.
invideo put it as comparable to mid-tier specialist models on stills and trailing the premium options on motion. That matches the pricing table, which is a decent sign both are describing reality.
What flipped me, specifically
It was not one output. It was that it was intuitive, and that it was fast.
The generation speed is the thing. It felt slightly faster than the OpenAI image model I normally use, and I did not time it, so take that as an impression rather than a benchmark.
What I can count is the output: 4 options where I expected 1. That is the part that changes behavior, because you experiment instead of committing.
On raw quality, I am not going to overclaim. Compared against the newest OpenAI image model, which gets genuinely precise on the fine detail like skin and wrinkles, Grok is not the precision winner. But it generates some flawless images, and the gap is much smaller than it was the last time I looked.
That is the honest shape of it: I went from dismissing it to keeping it open in a tab.
Does it change my stack? Partly, and I will say where it does not
Yes, more than I expected, but not everywhere.It adds a real option for images. I thought the field was Nano Banana and OpenAI's image model. Grok has earned a place in that set.
It will not make my thumbnails. No reason to switch. What I use works, and thumbnails are the one place where I am not experimenting.
It might make blog images. A quick explanatory graphic, the kind of thing that is not worth a long session. That is the use case I would test next.
I will not be running it through a CLI. For me this is a browser tool, not part of an automated pipeline.If I am honest, I have not found the killer use case for it in my own work yet. It is on the radar now, which it was not last week, and that is a real change even though it is not a migration.
Where this fits
If your question is really about which plan to pay for, that is free versus SuperGrok, one trigger per rung, and the $100 question is answered separately. If you got here from the app-building side of Grok, Build mode is the other thing that is quietly on the free plan. For how this sits against ChatGPT generally, Grok vs ChatGPT.
The rest of what I actually run is on my Claude at work hub.
How I checked this: I read and screenshotted xAI's pricing page on September 10, 2026, and used Grok Imagine on my own SuperGrok Plus account the same day. I did not test a free account, and I say so above rather than implying otherwise. Outside verdicts are linked to the publications that wrote them. Last reviewed September 2026.
Quick answer: Grok Build is a mode in Grok's chat window. You click the hammer icon under the message box, pick Build, describe what you want, and it builds a working website or app live in the conversation with a Publish button waiting at the end. There is nothing to install. There is also a completely separate developer product called Grok Build that runs in a terminal, and that is the one Google shows you first.
The Ship Lean rule on Grok Build: prototype in whatever you already pay for, and never let one vendor own all of it.
Published September 10, 2026. Plan availability verified on xAI's pricing page the same day.
Quick decisionWhat you are trying to do
Where to goGet a one-page site or a small app without touching code
Build mode, in the Grok chat windowDay-job software engineering
The Grok Build CLI, or Claude CodeFind out whether your plan includes it
It does. Every plan, including freeWork out what it costs you
It draws on the shared weekly poolDecide whether to pay xAI at all
My $100 SuperGrok Plus verdictYou searched Grok Build and Google handed you a terminal
There are two products named Grok Build, and they are not versions of each other.
The one most search results describe is the CLI, announced May 25, 2026 as "a powerful new coding agent for professional software engineering and complex coding work." You install it with a curl command. It runs in a terminal, plans before it edits, and works across a codebase. People who have put real hours into it rate it seriously, and it is genuinely not for you if you do not write software.
The one you found in your chat menu is Build mode, announced two months later on July 28, 2026. xAI's own sentence for it: "There's nothing to install and nothing to configure, and you never have to touch the code unless you want to."
Same name. One needs a terminal, one needs a sentence.
If you are reading this because you clicked something in Grok and a preview pane opened, you want the second one, and the rest of this page is about that.
Your chat menu has five modes, not four
Every mode guide ranking on Google right now lists four: Auto, Fast, Expert, Heavy. Count the menu.My own session on September 10, 2026. Five modes, Build checked, on a SuperGrok Plus account. The prompt is sitting right there: "build a website for a local landscaping website."
Here is what each one is actually for:Mode
What it does
Gated?Auto
Picks between Fast and Expert for you
NoFast
Answers quickly, thinks less
NoExpert
Thinks longer before answering
Yes. xAI's table shows a dash under FreeBuild
Makes websites, apps, games and dashboards
No. Checked on all seven plan columnsHeavy
Runs a larger group of agents on one problem
Yes, and the most restricted of the fiveHeavy is the confusing one, because Heavy is both a mode in this menu and a plan tier. I am on SuperGrok Plus at $100 and I do not have Heavy mode. I do have Build. That is the distinction the guides are collapsing when they tell you Build is an advanced feature you have to buy.
xAI does not publish a price for the Heavy tier, so I am not going to print one. My own account is the evidence I have: at $100 a month, Heavy mode is not in reach and Build is.
One more thing the menu quietly buried: xAI's own FAQ now says "Grok Studio is no longer supported. Use Grok Build instead." If you are searching for Studio, this is where it went.
Build stopped being a top-tier feature on August 19
At launch, Build mode was an Early Beta limited to SuperGrok Heavy subscribers. That is a real fact, it was true for three weeks, and it is why half the internet still says you need the top plan.
It stopped being true on August 19, 2026, when xAI published a follow-up post whose load-bearing sentence is: "Grok Build is now available on every plan, on the web and on mobile."
An announcement is not the product, so here is the pricing page on the same day.xAI's pricing page, captured September 10, 2026. Read the Grok Build row across all seven columns, then read the Grok Bot and Expert rows for contrast.
You can watch that gate close in real time in the reaction to the launch. @FFBuncho posted at Build Mode's announcement that you could publish an app to an xAI domain instantly with no domain contract, and added a note of disappointment that it looked to be SuperGrok Heavy only for now. That reservation was correct when it was written and stopped being true three weeks later, which is the whole problem with searching this question.
The same August 19 update also added remixing other people's published apps, custom domains and GitHub export. Small sites like Basenor still carry the Heavy-only headline in their URL, which is a decent illustration of how fast this fact went stale rather than a knock on them.
What one prompt actually produced in my first hour
I typed one sentence: build a website for a local landscaping website. That is the whole prompt, redundant wording and all.
It read the workspace instructions, connected to a machine, loaded design and scaffold skills, and started building while I watched a preview pane fill in. Then I closed the browser by accident, came back, and it had finished anyway.The finished page. Note the Publish button top-left. The business, the address and the phone number are all invented by the model.
It named the business Rowan Hill, dated it to 1994, placed it in Dutchess County, wrote the headline "Ground that feels as if it has always belonged," and built a working nav that scrolls you down the page.
It also built a real form, not a picture of one, though nothing is wired up to receive what someone types into it.Name, phone, email, a service dropdown and a free-text field, from one sentence of instruction.
My honest reaction in the moment was "holy shit," and I want to be precise about why: not because the output is remarkable in isolation, but because the input was one careless sentence and the thing finished on its own.
I would not hand that landscaping page to a real business yet
No. It needs tweaking, and the tweaking is my fault, not the model's.
A one-sentence prompt gets you a one-sentence site. Steal the prompt I should have written instead:The brand: the exact colors, the logo file, real project photos.
The place: the town and the service area, named.
The benchmark: one or two competitors ranking first in Google Maps for that keyword in that town, and what those sites actually have.
The brief: what the client wants that the competitors do not do.
Then let it iterate, because the second pass is where it gets good.It is also one page. You scroll, you click the menu bar, it scrolls for you. That is genuinely fine for a demo and it is not a business's website.
So the honest verdict: as a demo, yes. As something you invoice for, not off one prompt. What it changes is the cost of showing someone a mockup, and that is not nothing when you are pitching local businesses.
If that is the business you are in, the longer version of that pitch is selling AI builds to local businesses.
The other Grok Build is a terminal app for engineers
Briefly, because you may have landed here from a developer comparison.
The CLI is available to SuperGrok and X Premium Plus subscribers, runs interactively or headless, and runs multiple agents in parallel. Composio put 50-plus hours into comparing it against Claude Code, and Analytics Vidhya ran the same comparison under the headline "I Tested Both So You Don't Have To." I have not run that head-to-head myself, so I am pointing you at people who did rather than manufacturing a test.
I would ignore the SWE-bench numbers floating around this comparison. They are vendor-reported, they get copied between blogs without attribution, and they will not tell you anything about whether a landing page comes out right.
The model behind the CLI is grok-4.6. If you see a page claiming Grok 4.7, that page is guessing.
Grok Build usage limits: the same weekly pool as everything else
There is no separate Build allowance and no published Grok Build weekly limit. Per xAI's FAQ, Build sits on the one shared weekly usage pool alongside API, Chat, Imagine and Voice, and different products drain it at different rates. xAI publishes no per-tier number for any of them, so anyone quoting you a Build limit is inventing it.
The free tier works differently: it keeps its own Chat and Voice limits that reset on their own schedule, separate from the weekly pool.
I wrote the full mechanics up separately, because it is the single most misreported thing about this product: how Grok's usage limits actually work.
When to open Build instead of Claude Code
My actual rule, in the words I used out loud: choosing between these is like choosing between Macy's and JCPenney. They are more or less the same thing and it depends what is on sale.
What is on sale, for most people, is the ecosystem they already pay for.If you already pay for X Premium and a SuperGrok plan, Build is a no-brainer. You are not going to go buy a $200 Claude Code plan to do the thing you already have.
If you already pay for Claude Code or Codex, Build is your backup. When I run out of usage credit on Codex or Claude Code, I have Build sitting there for a dummy test or a throwaway prototype.
If you are picking one for serious work, I still think Claude Code and Codex are the two serious players. Lean on Claude Code when the copy has to be right; lean on Codex when the work is research-shaped.
If you are optimizing purely on budget, the interesting case is the top Grok plan. Instead of a $200 Claude Code plan plus a $200 Codex plan, one Grok subscription gets you Build, plus Grok Bot with its own separate usage pool, plus Cursor, which xAI's own Grok Bot launch post names alongside your Grok plan. That is a real proposition, not a gimmick.And here is the part I care about more than the feature comparison: you would be living entirely in their ecosystem. That is too risky for me. Do not put all of it on one vendor, however good the bundle looks this month.
The delta between these tools is small enough that people invent personalities for them. The delta that matters is which company owns all your work if they change the terms.
Where this fits
If you are working out the plan question rather than the tool question, start with what free Grok actually gets you, then whether the paid tier is worth it. If you came here from the image side of Grok, that is Grok Imagine, which shares this same weekly pool and has a very different answer on what free gets you.
Everything I run day to day lives on my Claude at work hub.
How I checked this: I used Build mode on my own SuperGrok Plus account on September 10, 2026 and screenshotted the session. Plan availability, mode gating and the announcement dates come from x.ai's own pages, read the same day. I have not tested the Grok Build CLI, and I say so where it comes up. Last reviewed September 2026.
Quick answer: Claude Code waits for you. Grok Bot works without you. That one line decides almost every routing question between them, and it is why I run both daily and have never had them compete for the same task.
Before anything else, the correction that half this SERP needs.
Grok Bot is not Grok Build
These are two different xAI products and several pages ranking for this query compare the wrong one to Claude Code without saying so.Grok Build is xAI's coding tool. That is the product that lines up against Claude Code as a coding tool.
Grok Bot is xAI's agent product. Named persistent bots, each with standing instructions, running on a cloud virtual machine.You do not have to take my word for the split, because xAI's own pricing page lists them as separate rows on different plans:xAI's pricing page, captured September 10, 2026. Grok Build ships on every tier including Free. Grok Bot only on the three individual SuperGrok tiers, and not on Business or Enterprise.
Different rows, different availability, different products. If a page tells you "Grok Bot is a terminal coding agent," it has not opened either one.
For the record: I have access to Grok Build and have not used it, so it gets no verdict on this page.
Claude Code waits for you. Grok Bot works without you
The obvious hypothesis is "does the job need my files." That is the line most comparisons draw, and in my setup it is wrong.
Grok Bot has access to my repo. It knows how to pull it, and if my computer is off, it knows how to SSH into my Mac mini.
It is even routing its browsing through that machine, so the bots do not look like bots and get slowed down by captchas. Nothing shady, just not being treated as a scraper while reading public pages.
So "whose files" does not separate them. Both can reach mine.
The real line is what kind of work it is:
Claude takes the content. All of it. Writing, building, the things where judgment about quality matters and I want to be in the loop while it happens.
Grok Bot is the sidekick. It runs analytics on everything. How is the long-form doing, how are the shorts doing, how is LinkedIn doing, how is the newsletter doing. It logs the wins and the losses. When something pops, it tells me. When someone comments on a thing, it logs it and tells me.
Consistent grunt work, on a schedule, that produces a record.Claude Code
Grok BotShape
Task-shaped, resets around each job
Employee-shaped, persistsWhere it runs
Your machine, your terminal
Its own cloud VM, one per accountWhose logins
Yours
Its own, authenticated onceWhen it works
While you are there
On a schedule, without youWhat I give it
Content and building
Analytics, inbox, monitoringWhat happens at 3am
Nothing
The job runsCost
Included on every paid Claude tier
From $30 on SuperGrokFailure mode
Stops
Can keep spendingWhy not just build the grunt work as workflows?
Fair question, and I did, for a long time.
The difference is not capability. It is that a conventional workflow can break silently.
You find out three weeks later that a thing stopped, and by then the ledger has a hole in it. It is also less interactive: you cannot really talk to it.
A hand-built workflow stack is more customizable, and that is a genuine advantage. It is also its problem. It becomes a machine assembled from scraps. Some good parts, some bad parts, some generic parts, all bolted together and none of it fluid.
Grok Bot is one system, same parent, same heart. Everything flows the way it should because it was built as one thing. More unified, less yours.
That tradeoff is the whole choice, and which side you want depends on whether you enjoy owning the plumbing.My Grok Bot roster as captured August 28, 2026. Named bots with standing instructions on one shared cloud machine, not a workflow canvas.What it grew into by September 10: a Chief of Staff bot that the other bots report into. I ask it one question a day.
Has Grok Bot taken a job away from Claude Code for good?
No.
That is the honest answer and I think it is the most useful sentence on this page, because it is not what a launch-week comparison would tell you.
Grok Bot did not take work off Claude Code. It picked up work that was not happening at all.
The analytics ledger did not exist before. The inbox got triaged by me, badly, or not at all.
Nothing migrated. Something new started.
The thing that makes me think that could change: Codex has genuinely surprised me lately, to the point of generating a tweet. Which makes me suspect the next Grok model might eventually take on work I currently keep on Claude. That model is not out and I am not going to pretend to know what it does.
For now: no job has moved, and I do not switch between them mid-task. I use both.
Where Claude Code actually lives
The mirror image of the roster above is that Claude Code is not a roster. It is a session in a terminal on my machine, plus Cowork as its non-terminal sibling in the desktop app.Claude Code and Cowork in the Claude desktop app. Same brain, two cockpits, both attached to my actual machine and my actual files.
And the honest caveat about the terminal: when I finally got into Claude Code it clicked immediately, but I am used to looking at a terminal window. I do not recommend that for everybody. If the terminal is the blocker rather than the capability, the fork you want is Cowork vs Claude Code, not this page.
Cost and risk, which are not symmetrical
Claude Code is included on every paid Claude tier. Pro at $20, Max 5x at $100, Max 20x at $200, per Anthropic's Max plan page and claude.com/pricing. You are not buying Claude Code, you are buying capacity for it. When you run out, it stops.
Grok Bot is included from $30 on SuperGrok, per xAI's pricing page as of September 10, 2026. Each eligible plan carries a weekly Grok Bot allowance, and overage is reported to bill from model and token cost with no Grok Bot specific spend cap yet.
That asymmetry deserves a sentence of its own. One of these fails by stopping. The other can fail by spending, unattended, at 3am. I have not been surprised by a bill, and I would still watch it closely for a first month.
There is a second risk that is structural rather than financial. All your bots share one cloud machine, along with its logins and its files.
That is convenient, and my read of it is that the bots are not meaningfully isolated from each other. Whatever one bot can log into, you have effectively handed to the whole roster. Scope what you connect accordingly.
Claude Code's risk profile is the opposite and more familiar: it is on your machine, in your files, and the blast radius is local.
Which one to start withYou want an agent doing something useful by Friday and you do not code. Grok Bot. Pick one job you do at the same time every day and hate. Not nine.
You want to build or write things and be in the loop. Claude Code, on any paid Claude plan you already have. And if the terminal is the problem, Cowork.
You are comparing coding tools specifically. Then you want Grok Build against Claude Code, not this pairing.
AI does real work in your week. Both. I do.The reason for "both" is not greed. There is no AI that does it all, and if you are running a business on this you do not want to rely on one. They get cancelled, they get rate limited, they go down, they get changed underneath you. Two vendors means one bad day is not your bad day.
If I were being properly serious about it, the most important jobs would run locally on hardware I own, with the paid frontier models sitting on top for the heavy thinking. Then a subscription change is an inconvenience rather than an outage. I am not there yet. I am running on paid models and I know that is a risk I am choosing.
The neighboring comparisons: Grok Bot vs Claude Cowork if you are choosing between the two agent products, what is Grok Bot for the week-one field test, Grok Bot use cases for the full roster and what each bot produces, and Grok vs Claude if you actually meant the models. More at Claude at Work.
Published and last reviewed September 10, 2026. The Grok Bot and Grok Build plan rows were read and captured that day from xAI's pricing page, screenshotted above. Grok Bot's VM, login and connector mechanics were verified for my week-one field test on August 28, 2026. Claude Code tier inclusion and Claude plan prices were verified for Claude Max 5x vs 20x against Anthropic's Max plan support article and claude.com/pricing. The routing rule, the repo and SSH setup, and the "no job has moved" answer are my own, from my own accounts. The overage-billing point is secondary reporting and labeled as such. A published head-to-head benchmark circulating for this pairing was left out because I could not trace it to the benchmark's own publisher.
Quick answer: these two do not compete, and every page comparing them is comparing the wrong layer. GPT-6 Astra is a model you hand hard work to. Grok Bot is a set of agents with their own computer that run standing jobs while you are asleep. I pay for both and they have never once contended for the same task.
Every result ranking for this pairing is an API aggregator: tokens per second, dollars per million, intelligence index. Useful if you are building software. Useless if you have two subscriptions and one Tuesday.
So here is the version for someone with two subscriptions and one Tuesday.
The category error, first, because it decides everythingGPT-6 Astra
Grok BotWhat it is
A model
Named persistent agentsWhere it runs
Your ChatGPT session, in Work or Codex
Its own cloud virtual machineWhose logins
Yours, in your browser
Its own, authenticated onceWhose files
Whatever you give the session
Its machine, under /workspaceWhen it works
While you are there
On a schedule, without youWhat happens at 3am
Nothing
The job runsHow it is billed
Inside your plan allowance
Weekly allowance, then reported overageRead the last two rows again. That is the entire decision.
Astra is the better thinker in this pairing. But it has never once told me something while I was doing something else.
What Astra actually ran for me
Astra shipped September 3, 2026. OpenAI describes it as state-of-the-art on computer use, browsing, software engineering, cybersecurity, science and professional work, with a context window over a million tokens.
In my week it does two things.
Research, and challenging Claude. It is genuinely good at pushing back on Fable when I want a second opinion on something I already believe. That is a real job and it is worth paying for. Smart enough that the disagreement is worth reading rather than annoying.
And this week, a game. The demos of people building games with Astra are real and I wanted in, so I spent a weekend on a Pokemon-style game. The graphics it generated genuinely stunned me. The game did not get built. Between that and my first week of running maximum effort on everything, I burned through two full allowances and produced something ugly.
That is not a knock on the model. I did not ask it the right way and the learning curve is real. But I want it next to the demo reels, because the demo reels are all successes.Astra running in ChatGPT Work on my Pro $100 account. Note the surface: on Plus, Astra appears only in Work and Codex, never in the Chat picker.
The honest cost note: at $100 I feel the ceiling for the first time. With GPT-5.6 Sol I felt it a little, but it was manageable. This is not.
If you want to build something serious with Astra, my read is you want the $200 plan. At $100 you miss out.
What Grok Bot actually ran for me
Grok Bot launched August 11, 2026. Named agents, each with standing instructions, sharing one cloud machine per account along with its logins and files.
I use it for orchestration. Analytics, publishing, intelligence, all happening in the background.
Every day, my inbox. Starting at 9am, about three times a day, a bot checks and manages it. I do not open it to triage. I open it to read what is left.
Every day, my numbers. A folder of bots scrapes my own analytics roughly three times a day and writes to a ledger. Shorts, newsletter, LinkedIn, comments. Wins and losses get logged rather than remembered. I am no longer posting blind, which sounds small and changed how I work more than any model upgrade this year.
Anything that drops in my space. A news scout watches X. Yesterday it told me OpenAI had shipped a new image generation model, and it was right that it mattered. We are now looking at integrating it into our YouTube thumbnails. That is a bot changing what I did with my afternoon.
And the sniper job. When I see a tweet I want a take on, I dump it into Grok Bot. It reads it instantly because it lives where the tweets are. Pasting that into a Claude chat is a headache. I do not want to open a terminal just for that. This is the fastest path from "interesting" to "what do you think."My account today. Inbox Hawk runs three passes a day on a routine and pings me with only what needs a human. That is the job Astra does not have a shape for.
The benchmark parity, compressed, because it should not decide this
Since somebody will ask: on Artificial Analysis, Astra and Grok 4.6 both sit at Intelligence 61. Astra's context is larger, Grok 4.6 is marginally faster, and API prices differ substantially in Grok's favor. Those are third-party figures and I am reporting them, not endorsing the methodology.
None of it should move you, for one reason: you are on a plan, not an invoice. Price per million tokens describes a bill you will never receive.
And you are not choosing between two models anyway. You are choosing between a model and an employee.
The billing asymmetry nobody prints
This is the most practical difference on this page and I have not seen it anywhere else.
ChatGPT stops you. When your allowance runs out, you wait, use a banked reset, or buy credits deliberately. The worst case is frustration.My ChatGPT meter, September 10, 2026. A ceiling, a reset date, and a $0 balance with auto-reload off. The failure mode here is that work stops.
Grok Bot bills you. Each eligible plan includes a weekly Grok Bot allowance, and extra usage is reported to be billed from model and token cost with no Grok Bot specific spend cap yet. A scheduled agent running unattended with no ceiling is a different category of risk from a subscription that simply stops.
One of these fails by stopping. The other can fail by spending. Watch the second one for a month before you trust it with a schedule.
On price itself, the record needs correcting. Grok Bot is included from SuperGrok at $30, per xAI's pricing page as of September 10, 2026.
Secondary coverage still ties it to SuperGrok Heavy bundling at around $300, and Reworked reported exactly that at launch. Whatever was true then, the page today says otherwise. It is also, notably, not listed as included on Business or Enterprise.xAI's pricing page, September 10, 2026. Grok Bot on the three individual SuperGrok tiers, and not on the two a company would buy.
If you can only pay for one
I asked myself the real version of this: if I had to cancel one tomorrow, which goes?
Astra goes. And I want to be clear that this is not a verdict on quality, because Astra is great and I would miss it and I would have to work out what to do about that.
Which sounds like it contradicts my own buying order above, and it does not. If you have neither, buy the model first, because a model is useful on day one and an agent is useful once you have taught it a job. I have already done that teaching. That is the whole difference.
Grok Bot is serving a more important purpose right now. Tracking the analytics, keeping things updated, doing the things I would otherwise not do. Losing Astra costs me a very good research partner and a sparring opponent. I would feel that. Losing Grok Bot means a category of work simply stops happening in my business.
The way I keep describing the difference: a conventional automation is like a teddy bear you have to move yourself. The moment you stop moving it, it dies. Occasionally it gets up on its own to do a boring chore, and sometimes it trips and falls and stays down. Grok Bot plays on its own, walks more smoothly, and when it falls it can get up or call one of the other bots for help.
That is a feel, not a benchmark. It is also why I would cancel the smarter one first.
So which do you buy?You mostly ask questions and want better answers. Neither of these. You want a good chat plan, and that is ChatGPT Plus or Claude Pro.
You have hard, long, multi-step work. Astra, on ChatGPT. And if you intend to build something serious with it, budget the $200 tier, because $100 has a ceiling you will meet.
You have recurring jobs you do at the same time every day and resent. Grok Bot, from $30. Start with one bot, not nine.
AI does real work in your week. Both, in that order. And honestly, for a reason beyond capability: running two vendors means one closed meter does not end your day.There is no AI that does it all. If you are running a business on this, do not rely on one. If I were making real money from it, I would be on the top tier of all of them and consider it a no-brainer, because it makes the money back.
The neighboring decisions: what GPT-6 Astra actually is, which ChatGPT plans get Astra, what is Grok Bot for the week-one field test, and Grok Bot use cases for the full roster. If your comparison is really against Claude's assistant, that is Grok Bot vs Claude Cowork. More at Claude at Work.
Published September 10, 2026. Grok Bot plan inclusion was read and captured that day from xAI's pricing page, screenshotted above. Grok Bot's VM and connector mechanics were verified for my week-one field test on August 28, 2026. Astra's launch description and context window come from OpenAI's September 3 announcement as captured on September 4; OpenAI's pages could not be fetched directly on September 10 and those facts are not re-dated. The benchmark parity figures are Artificial Analysis's published index, reported as third-party. The overage-billing point is secondary reporting and labeled as such. Screenshots are my own paid accounts, sidebars cropped. Grok 4.7 is unreleased and is deliberately not discussed anywhere on this page.
Quick answer: if you do not write code, Claude Pro at $20 is almost certainly enough, and the advice telling you otherwise was written for developers. The $20 plan does not break when you ask things. It breaks when you start making things.
That distinction is the whole page. Chat, writing, research and reading documents barely move the meter. "Help me build this website" and "edit this video" move it a lot.
So before you spend the extra $80, there is a two-week test worth running. It costs nothing and it usually answers the question for you.
What you are actually buying, so we can stop guessing
The thing that makes this decision feel harder than it is: people assume Max unlocks a whole better Claude.
It mostly does not. The feature list is the same, and between Max 5x and Max 20x it is identical, which I proved row by row on Claude Max 5x vs 20x.
But there is one real gate between Pro and Max, and it would be dishonest to flatten it.Claude Pro
Claude Max 5x
Claude Max 20xPrice
$20/mo, or $17/mo billed annually
$100/mo
$200/moAnnual billing
Yes
No
NoClaude Code included
Yes
Yes
YesFable 5 and 5.1
Not in plan limits. Usage credits only
Included, up to 50% of weekly limits
Included, up to 50% of weekly limitsEverything else
Same
Same
SameWhat extra money buys
Baseline
Capacity
More capacityAnthropic's own plan matrix. The Fable row is the one place Pro and Max genuinely differ.
Since July 19, 2026, the newest model is not part of what $20 buys. You can still use it, you just pay usage credits on top. That is the honest version, and it is covered properly on is Claude Pro worth it.
So the accurate framing is: Max buys capacity, plus in-plan access to the newest model. It does not buy different features, and it does not buy a smarter Claude on the models Pro already includes.
Worth noticing the annual row too. Pro has an annual option and neither Max step does.
How the limits actually work, in one paragraph
Two clocks. A five-hour rolling session window, and a separate weekly limit that applies across all models. You need room in both.
Anthropic does not publish a message count for any tier. Every "you get 45 messages" number you have read is somebody's estimate, usually derived from a coding workload, usually with no date on it.
The full mechanics live on Claude usage limits explained. For this decision, the only thing you need is that the meter measures work done, not messages sent.
The signature of the wall
Here is the part that decides it, and I want to be precise because this is the sentence I would hand someone on the $20 plan.
You are not going to hit the limit by asking chat questions.
Where people start running out is the moment the request changes shape. "Can you help me build a website." "Can you help me edit a video." "Can you create a video." That is when you start burning real credit.
So the useful question is not "how heavy a user am I." It is: does my week contain jobs where Claude is producing an artifact over many steps, or is it mostly answering me?
If it is mostly answering you, Pro is enough. Genuinely. I say that as someone paying $200 a month who would happily take your $80.
The clearest public illustration of the other side is the Claude Max class action, where lead plaintiff Karl Kahn is reported by Engadget to have upgraded to Max 20x for heavy coding and had a single five-hour session consume 15% of his weekly quota.
Read that as workload evidence, not as a plan warning, and remember it is an allegation in a filed complaint rather than a measurement. One session, allegedly 15% of a week, on the most expensive tier available.
If it is even roughly right, it tells you what making things at volume costs. Nothing you do in a chat window comes near it.My own meters, September 5, 2026. This is a Max 20x account on a normal working day, which is the point: even heavy use looks unremarkable until the work changes shape.
The correction I would make before you spend anything
There is a mistake underneath this question that costs more than the $80, and it is thinking about it as a coding decision at all.
If you zone in on code, you get stuck on "I need to build an app" and "I need a website." Those are real use cases. Most people do need a website, and plenty of people could save money building a small app they use regularly.
But that framing hides the actual opportunity. Even if you never write a line, you should understand what code can do for you.
Say you are a video editor and you do not code. Claude can still edit and polish your videos, because underneath it is running scripts and tools rather than typing into a timeline.
I have processes running on ffmpeg and Python that edit my videos. I do not need to understand any of it. I put a video in, a polished video comes out, and I still review it.
That is the thing worth $80, if it applies to you. Not a better chatbot. Hours back.
And the numbers are not close. A simple short takes about an hour to edit properly. Fourteen of those is fourteen hours, and that is just the cutting phase.
Then there is polishing, which is roughly another fifteen. Conservatively, 20 to 30 hours on short-form alone, before you touch long-form.
I spend about 15 minutes checking them, and that is the cutting phase.
If that trade is not worth $80 to you, that is a fair position and I am not going to argue you out of it. You just want to be honest that you are still paying for it. You are paying with your time.
I am not going to relitigate the $100 case here, because it has its own page: is Claude Max worth it. The point for this decision is narrower. The thing that breaks Pro is not code. It is production.
The two-week test
So, the actual thing to do before you spend the extra $80.
Do not measure your usage. You will not learn anything from a percentage, and you will talk yourself into an upgrade you do not need.
Instead, change one habit for two weeks:Stop treating it as a chat. Once a week, hand it a job with an output, not a question. A real one from your actual work. Edit this. Build this. Process these forty files.
Route the routine work down. Do not run the most capable model on things that do not need it. This is the single biggest lever and it is free.
Start fresh sessions. A conversation you never restart forces Claude to re-read an ever-growing history every single turn. That is the most expensive habit there is, and it is a habit, not a plan problem.
Then wait for pain. Not a warning. Actual pain: stopped, mid-task, on something with a deadline, more than once.If two weeks of that never stops you, Pro was enough and you just saved $960 a year.
If it stops you repeatedly, you now know exactly what for, which is the only useful reason to upgrade.
Let the pain tell you
That is my whole rule for this ladder, and it is how I climbed it without regretting a step.
I went free, then the $20 Claude plan, then ChatGPT at $20 as well, then cancelled one because at that point it felt expensive. Back and forth for a while. Eventually $100. Eventually $200, because I kept running out of credits and got tired of waiting.
Not one of those upgrades came from a comparison table. Every one came from a wall.
The same rule governs the step above this one, and I have written out how to apply it at the $100 to $200 line on Claude Max 5x vs 20x.
The step that was clearly wrong
Since I am telling you to let pain decide, here is the one time I did not, and it cost me.
I went to an annual plan on Codex, OpenAI's coding tool, and it was too quick. I had not earned it.
Because I had committed, I spent the whole time trying to justify it, which meant burning the highest available model on every single task. When I ran Claude, I would have Codex running alongside as a sparring partner, permanently. Not because the work needed it, but because I had so much allowance that spending it felt free.
It still does that job sometimes. But now we are deliberate about which model gets which work.
The lesson is not "avoid annual." It is that buying capacity ahead of need does not just waste money. It teaches you expensive habits, and those habits are what make the next tier feel necessary too.
Which is also why I would tell anyone new to Claude to expect to be inefficient for a few weeks. You will burn far more than you need to, then you will learn the habits. Do not buy a tier to paper over a learning curve you are about to climb anyway.
The cheaper levers, in order
Before $80 a month, all of these are free or cheaper:Model routing. Run the cheap model for routine work by default. Most people have this backwards.
Session hygiene. Fresh sessions rather than dragging one bloated conversation through the whole day.
A project instructions file so you stop re-explaining yourself every session.
A second $20 subscription somewhere else instead of one $100 plan. Two providers at $20 gives you a working day when one meter closes. One provider at $100 does not. If AI does real work in your week, a single vendor is a single point of failure.So, is Pro enough?
For most people reading this: yes, and the fact that you are asking rather than being stopped is itself the answer.
The exception is not "you are a developer." It is "you have started making things, at volume, and the waiting has begun costing you real time." When that is true you will not need a blog post to tell you. You will be annoyed.
At that point the real question is not Pro versus Max, it is which Max. For non-coders, that is is Claude Max worth it.
Still deciding whether to pay at all? That is is Claude Pro worth it. The tier-by-tier gates are on Claude Pro vs Max, and the two-meter mechanics are on Claude usage limits explained. More at Claude at Work.
Published September 10, 2026. Plan prices, the annual-billing gap and the feature parity between the two Max tiers were verified for Claude Max 5x vs 20x against Anthropic's Max plan support article and claude.com/pricing. The 15%-of-a-weekly-quota figure is Engadget's reporting of an allegation in a filed complaint, linked above, not a finding and not my measurement. The meters are my own Max 20x account. The video editing hours are my own workload, not a benchmark. Anthropic publishes no message count for any tier, so this page does not quote one. I looked for non-developer accounts of hitting the Pro wall to include here and could not verify any with a real name and a live link, so that section was left out rather than filled with anonymous paraphrase.
Quick answer: Grok Bot is for standing jobs, not questions. If a task happens on a schedule, produces something you check later, and does not need files on your laptop, it belongs to a bot. If you would rather just ask and read the answer, use the chat.
Everything ranking for this query is a list of 35 or 50 hypothetical use cases assembled from launch-week hype. Useful for ideas. Nobody behind them has run one for a month.
I have. Here is my actual roster, what each job produces, the one that flopped, and the one I would cancel first.
First, the price thing, because it is wrong everywhere
The most repeated claim about Grok Bot is that it requires a $300 SuperGrok Heavy subscription. Whatever was true at launch, it is not true today, and I can show you the page.xAI's pricing page, captured September 10, 2026. "Grok Bot access" is a listed feature of the $30 SuperGrok tier.
And the feature matrix on the same page has a row that nobody has written about:Same page, same day. Read the Grok Bot row across all seven columns.
Two things fall out of that row:Grok Bot starts at $30, on SuperGrok. Not $300.
Grok Bot is not listed on Business or Enterprise. The two plans a company would actually buy are the two that do not get the agent product. Grok Build, the coding tool, is on every tier including Free. Grok Bot is not.If you are evaluating this for a team, that is the fact to check before anything else on this page.
I pay for SuperGrok Plus at $100 myself, for the headroom rather than for access.
The one-line test for whether a job belongs to a bot
Before the list, the filter I use. A job goes to a bot when all three are true:It repeats on a schedule you could write down.
It does not need files on my laptop. The bot lives on its own cloud machine.
I can check the output later instead of watching it happen.Fail any one of those and it is not a bot job. That third one is the one people skip, and it is why half the use-case lists on the internet describe things nobody could actually verify.
The three folders
I group everything into growth, publishing, and intelligence. Not because the product asks for folders, but because those are the three things I need to happen without me.My account today, September 10, 2026. Inbox Hawk runs three passes a day on a routine, then pings me with only the items that need a human. The message body is blurred because it is my actual inbox.Chief of Staff on the same day. The other bots report into it, and I ask it one question instead of reading each feed.
One mechanic worth knowing before you build a roster: every bot on your account shares one cloud virtual machine, along with its logins and its files. You authenticate once and the whole team inherits it. Durable work belongs under /workspace, because files elsewhere can disappear during recovery.
That is convenient and it is also the security story. My read of that architecture, and it is a read rather than something xAI documents: your bots are not meaningfully isolated from each other. Whatever one of them can log into, you have effectively granted to the roster. Scope what you connect accordingly.
Folder 1: Growth (the ledger)
These exist so I stop posting blind. They scrape my own analytics roughly three times a day and write to a ledger.Bot
What it doesShorts Hawk
Watches short-form performanceSubstack Hawk
Watches newsletter performancePlan Watch
Tracks changes to AI plans and pricingBank Pull
Pulls financial numbers on a scheduleComments
Tracks comments across platforms and tells me when someone is talking to meLinkedIn Hawk
Watches LinkedIn performanceAll six scrape on a schedule, roughly three to four times a day, and write what they find into a ledger.
The point of this folder is not the dashboards. It is that the wins and the losses get logged rather than felt. Before this, I was guessing about what worked from memory, which is the worst analytics system ever built.
Folder 2: PublishingBot
What it doesYouTube long-form uploader
Takes a finished video and gets it publishedSubstack Notes publisher
Posts notes without me opening the appThis is the folder that does something I genuinely could not do before. Not "did it faster." Could not do at all, because it happened at a time when I was doing something else.
Folder 3: IntelligenceBot
What it doesChief of Staff
Oversees the other bots, surfaces what mattersInbox Hawk
Checks and manages my inbox, starting at 9am, about three times a dayNews Scout
Scrapes X for anything that drops in my spaceDistribution Scout
Watches videos I follow and summarizes them my wayNews Scout is the one that changed how I work. Yesterday it told me OpenAI had shipped a new image generation model, which I agreed was a big deal, and we are now looking at integrating it into our YouTube thumbnails. That is a bot changing what I did with my afternoon.
Distribution Scout is the newest one and it has earned its keep fastest. It watches long videos for me and gives me summaries in the format I want, which is not the format any summarizer app gives you.
What is actually good about this, versus a normal automation
I have built the same kinds of jobs before with conventional workflow automation, and the difference is not capability. It is persistence.
A normal automation is a cron job wearing a costume. When it breaks, it tends to break silently, and you find out three weeks later that a thing stopped. A Grok Bot has a good enough sense of its own job that it keeps going, and when it trips, it can pick itself up or escalate.
The way I keep describing it: with a normal workflow, you have to move the thing yourself, and the moment you stop moving it, it dies. It will occasionally get up and do a boring chore, and sometimes it falls over. Grok Bot plays on its own, and when it falls it can get up, or call one of the others for help.
That is a feel, not a benchmark. But it is the reason I stopped rebuilding these as workflows.
The tradeoff is control. A hand-built workflow is customizable to its exact purpose, which is a real advantage, and also its problem: it becomes a machine assembled from scraps, some good, some generic, none of it fluid. Grok Bot is one system with one heart. More unified, less yours.
The honest part: what is not working
I would rather you get this from someone who pays for it than from a listicle.
One flopped outright. In week one I built a lead scout and it produced nothing. The product moved a boundary. It did not become magic.
Shorts Hawk is buggy. It is in the roster and it is not reliable yet.
News Scout is a little rough. Not broken, just inconsistent enough that I read its output with suspicion.
LinkedIn Hawk is the one I would kill first. If xAI doubled the price tomorrow, that is the one that goes. It is fun to watch and I could not honestly tell you what it is doing for me. Writing this page was a decent audit trigger, which is its own lesson: building the plumbing is the easy part, and then you let it run and stop asking whether it earns anything.
Chief of Staff and Inbox Hawk are the ones that stay. Inbox Hawk is not impressive and it has held up. Chief of Staff is managing everything else, which is the job I actually needed.
And the billing risk is real. Each eligible plan includes a weekly Grok Bot allowance, and extra usage is reported to be billed from model and token cost, with no Grok Bot specific spend cap yet, per eesel's pricing writeup. A scheduled agent with no ceiling is a category of risk that a fixed subscription does not have. Watch it closely for the first month.What I actually pay xAI. SuperGrok Plus, for the headroom.
What everyone else runs
The community inventories are genuinely useful for ideas, and I would rather point you at them than pretend I invented the category:Matt Van Horn's 30-day sweep of what people across X, Reddit and YouTube actually use it for.
Sid Saladi's 50-use-case inventory, the most-cited list out there.The categories that recur across the community lists generally, beyond what I run: inbox triage and cleanup, meeting prep, supplier and vendor outreach, dispatch and logistics coordination, and permit or paperwork chasing for trades businesses.
I am deliberately not retelling the individual operator stories that circulate with those lists. Several are compelling and I could not trace them to an original post I could link, so they stay off this page. If you find them, judge them yourself.
How to steal this
If you want one bot by Friday, do this:Pick the job you do at the same time every day and resent. Inbox triage, most likely.
Write the job in three sentences. If you cannot, the bot will not do it well.
Give it a schedule, not a trigger. Standing jobs beat clever ones.
Make it write somewhere you will see. A ledger, a doc, a message. Output you never read is a bot that does not exist.
Audit it in 30 days. Ask what it produced. Kill it if the answer is thin. Mine would not all survive that question, which is the point.The mistake is not building the wrong bot. It is building twelve and never asking any of them what they did.
Deciding whether the product is for you at all? What is Grok Bot is the week-one field test, Grok Bot vs Claude Cowork is the assistant comparison, Grok Bot vs Claude Code is the one for people who already run an agent, and Grok usage limits explained covers the meters. The equivalent collection on the other side is Claude Cowork use cases. More at Claude at Work.
Published September 10, 2026. The plan tiers that include Grok Bot, including its absence from Business and Enterprise, were read and captured that day from xAI's pricing page, screenshotted above. The one-VM-per-account mechanic and the connector catalog were verified for my week-one field test on August 28, 2026. The roster, the bots and the verdicts are my own paid SuperGrok Plus account. The overage-billing point is secondary reporting and is labeled as such. Community operator anecdotes I could not trace to an original source were left out rather than repeated.
Quick answer: there are two clocks and two separate pools, and OpenAI does publish numbers. Per its help center, GPT-6 Astra's estimated messages per five-hour window are 5 to 45 on Plus, 25 to 225 on Pro 5x, and 100 to 900 on Pro 20x. A weekly limit sits on top of that, and you need allowance in both to keep going.
Every page that ranks for this query copied the same leaked numbers with no date and no source. So I went and screenshotted the actual table.
The finding that matters most is not any single number. It is the ratio sitting inside the table: Astra costs you at least double what GPT-5.6 Sol costs you, at every tier.
The numbers OpenAI actually publishesOpenAI's help center article "Managing usage with GPT-6 Astra in Work and Codex," captured September 10, 2026. The page header showed it had been updated two days earlier.
The caveat OpenAI prints directly above that table deserves to be read out loud, because every page quoting these numbers drops it:The table below shows estimated local messages per five-hour period. These are not fixed message limits. Actual usage varies by task, model and settings, and weekly limits may also apply.So the honest version of "how many Astra messages do I get" is: somewhere between 5 and 45 on Plus, and which end of that range you land on is decided by what you ask it to do.
Now read the same table for the thing nobody pointed out:Model
Plus
Pro 5x
Pro 20xGPT-6 Astra
5 to 45
25 to 225
100 to 900GPT-5.6 Sol
10 to 100
50 to 500
200 to 2,000GPT-5.6 Terra
25 to 200
125 to 1,000
500 to 4,000GPT-5.6 Luna
250 to 2,000
1,250 to 10,000
5,000 to 40,000Astra is about half of Sol on every row. The bottom of each range is exactly half: 5 against 10, 25 against 50, 100 against 200. The top of each range is a little under half: 45 against 100, 225 against 500, 900 against 2,000.
That is the whole cost story in one line: the same conversation costs you at least twice as much on Astra, and up to about 2.2 times as much.
And if you drop to Luna for the routine stuff, you get roughly 45 to 50 times the messages you would get on Astra.
One thing the column headers will not tell you: Pro 5x is the $100 plan and Pro 20x is the $200 plan. OpenAI labels them by multiplier, your credit card labels them by price, and nobody maps the two for you.
Two clocks, and you need both
This is the mechanic that makes limits feel random when they are not.Same article, same capture date. The shared-pool sentence and the two-window rule, in OpenAI's own words.Depending on your plan, usage limits may apply over a five-hour window and a weekly window. Where both apply, you need allowance remaining in both to continue.The five-hour limit determines how much usage is included in each window. A new window starts when you send your first message in Work or Codex after the previous one ends. It is not a fixed clock on the wall; it starts when you do.
The weekly limit controls how much included work you get over the weekly usage period.The consequence, in OpenAI's words: "You may reach the five-hour limit before five hours have passed, even if you still have weekly usage remaining."
That is why the wall feels arbitrary. You still see plenty of weekly allowance in Settings, so you assume you are fine, and then a single heavy task closes the five-hour door.
Two pools, and Chat is not one of them
The second confusion is which meter you are draining.Surface
Pool
Who gets Astra thereChatGPT Work
Shared Work and Codex allowance
Plus, Pro, Business Standard and PremiumCodex
Same shared Work and Codex allowance
SameChatGPT Chat (as GPT-6 Pro)
Separate Chat message limits
Pro, Business, Enterprise. Not PlusOpenAI's wording: "Work and Codex share the usage allowance included with your plan." And separately: "In Chat, GPT-6 Pro is powered by Astra and is available on eligible Pro, Business and Enterprise plans. Its message limits are separate from the shared Work and Codex allowance and vary by plan."
So running Astra hard in Codex does not eat your Chat allowance. It does eat your Work allowance, because those two are the same bucket.
One more rule that catches people mid-task: "Switching models does not restore allowance in a shared usage pool." Dropping from Astra to Luna after you have hit the wall does not open the door. It only slows the next burn.
The setting that matters more than the model
Here is the paragraph I think is the most valuable thing OpenAI has published about Astra, and it is filed under a settings heading where nobody will read it.Same article, same capture date. Read the bolded sentence twice.Lower effort does not mean lower capability across models. For example, Astra at Low effort can outperform Sol at High effort. If you've been getting good results with Sol at High, try Astra at Low or Medium as a starting point.And the warning attached to the top setting:Higher effort can use more of your allowance and does not always produce a better result.If you take one action from this page, it is that one. If Astra feels like it is draining you, check whether you are running maximum effort on tasks that never needed it. That was my problem.
My own advice lands in roughly the same place. The way I put it: use the top setting to plan the hard thing, then drop down to execute it. I hedge on how far down, somewhere between low and high, because I genuinely do not know.
The reasoning earns its price during planning. During execution it usually does not.
Did OpenAI cut the limits in September?
You will find headlines saying allowances were cut by up to 4x for heavy users, and you will find OpenAI saying Astra consumes 3 to 4 times less subscription usage on long-tail workloads. Both claims are circulating in the same week.
I am not going to pick the scarier one. I did not capture the old table, so I cannot show you a before and after, and neither can any of the pages ranking for this query.
What I can show you is the table as it stands today, which is above, and the ratio inside it: Astra is priced against your plan at somewhere between two and two and a bit times GPT-5.6 Sol. If your allowance feels like it evaporated after the rollout, that ratio is a sufficient explanation on its own. You did not need a secret cut to burn twice as fast. You just needed to start using the new model for everything, which is exactly what everyone did.
I certainly did. My first week I ran maximum effort on questions that did not deserve it.
Resets, including the ones OpenAI handed out
There are three kinds, and they behave differently.
Banked resets. Saved to your account until used or expired. Per OpenAI: "A full banked reset refreshes your five-hour and weekly Work and Codex limits and changes your weekly reset date." You use one from Settings then Usage in the desktop app, where it shows up as something like "1 reset available" or "Full reset."
Automatic or global resets. Applied for you, announced by OpenAI, and they do not appear as a saved reset in Settings.
Purchased instant resets. Available to eligible personal Plus and Pro accounts depending on billing country. They refresh immediately at checkout, cannot be saved, and they pull your whole weekly period forward: your new week starts at your next Work or Codex request, and the next automatic weekly reset is seven days after that.
That last mechanic is worth understanding before you buy one in a panic. You are not adding allowance to this week. You are starting next week early.
And the launch giveaway was real, which is why so many people had spare resets in early September:The launch reset schedule, in OpenAI's own words. Captured September 10, 2026.
Three separate grants: banked resets on September 3 and September 4, and an automatic refresh on September 7 that did not bank anything. OpenAI attributes them to "the broader Astra launch delay," and states that individual access problems after launch do not earn you another one.
Also worth knowing before you get your hopes up: "A reset restores your allowance. It does not permanently increase your plan's limits or change how much allowance a task uses."
What a real meter looks like
Here is my own account today. ChatGPT Pro at $100 a month.My Settings then Usage and billing pane, captured September 10, 2026. Sidebar cropped.
Reading it left to right, because this is what the pages quoting leaked numbers cannot give you:Weekly usage limit: 68% left, resets September 16. That is six days of runway on a meter that read 30% left five days earlier, before the September 7 automatic reset landed.
Credits balance: $0, automatic reload off. No hidden top-up propping this up.
One banked "Full reset" left, expiring October 5. That is what remains of the launch giveaway.On a heavy prompt I can watch five to ten percent of that meter disappear at once. Not over a session. On one prompt.The same pane five days earlier, September 5, 2026: 30% left, resetting September 7. That September 7 reset is the automatic one OpenAI documents above.
I have burned two full resets since Astra shipped. The first one was pure stupidity: I ran the highest effort setting on everything because it was new. The second went to a weekend attempt at building a game, which ate an entire allowance and produced something ugly.
I have not hit a wall since. Partly because I had banked resets to fall back on rather than sitting and waiting, and partly because I stopped spending like that.
What to do when you actually hit it
In order, and be honest with yourself at step one:Check which clock stopped you. Settings then Usage tells you. If it is the five-hour window and your weekly meter is healthy, you are not in trouble, you are just early.
Wait, if the job can wait. Five hours is not a crisis. Most jobs can wait five hours.
Move the job down a model, not sideways. Luna gives you roughly 45 to 50 times the messages of Astra. Most of what you were doing did not need Astra.
Spend a banked reset if you have one and the deadline is real. Check the expiry first.
Move the job to another tool. This is the one nobody says out loud. I run Claude and Grok Bot alongside ChatGPT, and when one meter closes I do not sit and stare at it. Grok Bot in particular has been generous with resets.That last point is the real answer to "how do I stop hitting limits," and it is not a limits answer. There is no AI that does it all, and if you are running a business on this you should not rely on one. A single subscription is a single point of failure.
The one-line version
Two clocks, two pools, and Astra costs at least double GPT-5.6 Sol per message. Turn your effort setting down before you turn your spending up.
Comparing the meters across vendors? Claude usage limits explained and Grok usage limits explained are the same breakdown for the other two. If you are deciding between the $100 and $200 ChatGPT tiers, that is ChatGPT Pro $100 vs $200, and what GPT-6 Astra actually is covers the naming. More at Claude at Work.
Published and last reviewed September 10, 2026. Every allowance number, the two-window mechanic, the Work-versus-Chat pool split, the effort-level guidance, the reset types and the launch reset schedule were read and captured that day from OpenAI's help center article "Managing usage with GPT-6 Astra in Work and Codex," screenshotted above.
The reported September allowance cut and OpenAI's long-tail efficiency claim are both secondary reporting and are labeled as unverified in the text. I did not capture the previous table and cannot show a before and after. The meters are my own paid Pro account with the sidebar cropped, on two dates five days apart. OpenAI's figures are estimated ranges, not entitlements.
Quick answer: GPT-6 Astra is OpenAI's most capable model, released September 3, 2026, with a context window over one million tokens. It is a model, not a new app. And there is a good chance the word "Astra" does not appear anywhere in your ChatGPT, which is not a bug in your account so much as a naming decision.
I pay $100 a month for ChatGPT Pro. When I opened the model list today to write this page, Astra was not in it.
More on that in a minute, because it turns out to be the most useful thing on this page.
What it is, without the benchmark table
Astra is the model OpenAI now points at its hardest work. From OpenAI's launch post: "state-of-the-art on computer use, browsing, software engineering, cybersecurity, science, and professional work," with a context window of more than one million tokens.
OpenAI's own help center is blunter about when to reach for it. It calls Astra "Our most capable model for coding, research, analysis, and complex problem-solving," and the examples it gives are investigating a difficult bug and working through an unfamiliar problem.
Read that list again as a normal person. It is not "answers your questions better." It is: reads a lot, goes and does things, sticks with a hard problem.
That is the honest one-line version. Astra is for the jobs that take a while.
For me it is research and a sparring partner. It is genuinely good at pushing back on Claude Fable when I want a second opinion on something I already believe, and it is smart enough that the pushback is worth reading. I have also started using it to build a game, which I will get to.
What I do not open it for is my daily content work. That still goes to Fable, because that is the one that feels like a partner rather than a very good search.
One model, three names
This is the part that turns a five-minute question into a forty-minute one, and nobody explains it.Where you are
What it is called thereChatGPT Work, and Codex (the developer surface, ignore it if you do not code)
GPT-6 AstraThe ChatGPT Chat tab
GPT-6 ProThe API
gpt-6-astraOpenAI's help center states the split directly: "In Chat, GPT-6 Pro is powered by Astra and is available on eligible Pro, Business and Enterprise plans. Its message limits are separate from the shared Work and Codex allowance and vary by plan. Plus includes Astra in Work and Codex, but not GPT-6 Pro in Chat."
So if you are on Plus, hunting the Chat model picker for the word "Astra," you will hunt forever. It was never going to be there.
Here is my own Chat picker on the $100 Pro plan a few days ago. No Astra anywhere in it.My own Chat picker, captured September 5, 2026, on ChatGPT Pro at $100 a month. The model is in the account. The name is not in the list.
Why you cannot find it, in the order worth checking
Now the live one. This is my Work tab today, September 10, with the model list open.My Work model list on September 10, 2026, on a Pro plan that is entitled to Astra. Default, 5.6 Sol, 5.6 Terra, 5.6 Luna, 5.5, 5.3 Codex Spark. No Astra.
I am not showing you that to claim OpenAI pulled the model. I am showing it because it is the single most common thing that happens to real people in week two, and OpenAI documents the fix in its own help center under a heading called "Update your ChatGPT desktop app."
Their instruction, which is more specific than the usual "try restarting":If Astra is missing, check for updates even if you recently installed one. Your app may need another update and a full restart.OpenAI's help center, captured September 10, 2026. A three-step procedure that exists because this happens to a lot of people.
So the checklist, in the order that actually resolves it:Update the desktop app, then fully quit and reopen. Then check for updates again. OpenAI explicitly says a second round may be needed.
Check which tab you are in. Work and Codex offer Astra by name. Chat offers GPT-6 Pro instead.
Check your plan against the surface. Plus gets Astra in Work and Codex only.Worth saying plainly: my app had crashed to an error screen minutes before I took that shot, and the recovery still came back without Astra listed. Updating is a real step, not a brush-off.The state my app was in shortly before the model list above, September 10, 2026. Note which button OpenAI puts first.
When it is there, it looks like this. Same account, five days earlier, running a real job in Work:The same account on September 5, with the Work composer chip reading GPT-6 Astra Ultra. Third name, same model.
Do you have it? The short versionFree: no.
Plus: yes, in Work and Codex. Not in the Chat picker.
Pro $100 and Pro $200: yes, in Work and Codex, plus GPT-6 Pro in Chat.
Business and Enterprise: yes, with seat-level rules.I am deliberately not rebuilding the full surface-by-plan table here, because it already exists: which ChatGPT plans get Astra is the page for that question. This one is about what the thing is.
The part nobody warns you about: it eats your plan
Astra burns your included usage fast. I am on the $100 plan and I have already burned through two full resets.
That is not a vibe. OpenAI publishes estimated messages per five-hour window, and on every plan tier Astra's range is about half of GPT-5.6 Sol's. Same messages, at least twice the burn. The full table and the two clocks behind it are on GPT-6 Astra usage limits explained.
The number worth carrying around: on Plus, the bottom of OpenAI's own estimated range for Astra is five messages in a five-hour window. So "a couple of messages" is genuinely the right mental model for a heavy Astra session.
Which is why the setting nobody touches matters more than the model you pick. OpenAI's guidance here surprised me, and it is worth quoting:Lower effort does not mean lower capability across models. For example, Astra at Low effort can outperform Sol at High effort. If you've been getting good results with Sol at High, try Astra at Low or Medium as a starting point.OpenAI's help center, captured September 10, 2026. The most useful paragraph OpenAI published about Astra, and it is buried under a settings heading.
I had heard that claim secondhand and half believed it. It is real, and it is in their documentation.
My own rule lands in the same place from the other direction. In my words: use the top setting to plan the hard thing, then drop down to execute it. I hedged on how far down, somewhere between low and high, because I genuinely do not know.
Planning is where the reasoning earns its cost. Execution mostly is not. Mapping my "top setting" onto OpenAI's effort control is my reading, not their instruction.
The full mechanics of the two clocks and three pools are in GPT-6 Astra usage limits explained.
What people are actually building with it
The word-of-mouth is not benchmark tables. It is people posting what they made. These are the ones that made me open the model picker, with the original posts linked so you can judge the source yourself.Tom Krcha gave Astra an old drawing of a steam train and asked for it in Blender. His post says it came back as 3,295 editable objects in a few minutes. Original post on X.Roberto Nickson fed it nine phone photos of his studio. It worked out the spatial layout and returned a walkable 3D model. Original post on X.
Three more worth your time, no screenshots because they are threads and video:KP's roundup of the early demos, compiled the day after launch with credit to each builder. The fastest way to see the range.
Min Choi's ten examples: games, 3D scenes, character models, all from prompts or images.
Riley Brown's take: his biggest moment was not visual at all. He gave Astra context on his whole business and let it plan. That matches my own use more than the game clips do.One caveat on all of it. Every one of these people is on a Pro or Business plan or the API. None of them are doing this on the $20 plan, and the next section is why.
The game thing, honestly
The hype about people building games with Astra is real. I have seen shooters. I wanted in, so I spent a weekend trying to build a Pokemon-style game with it.
The graphics it generated genuinely stunned me. These looked different.Three screens Astra generated for me over the weekend of September 6. Original creatures, original names, a battle system and an overworld. This is what "impressive graphics" meant.
The game did not get built. It burned two full usage resets and what I ended up with was ugly.
I am fairly sure that is me, not the model. I did not ask the computer the right way, and the learning curve is a real curve. But I want that on the record next to the demo videos, because the demo videos are all successes and the honest week-two experience includes an expensive failure.
If you want to build something serious with it, my read is you want the $200 plan. At $100 you feel the ceiling, and I never felt it with GPT-5.6 Sol.
What I would do this week
If you are on Plus: open the Work tab, not Chat, and give it one hard thing you have been avoiding. A long document you never read, a mess you have not untangled. That is the shape of task Astra is for. Do not burn it on questions.
If you are on Pro: try the thing OpenAI suggests and drop your effort setting a notch. You may find you were paying double for reasoning you did not need.
And if you cannot find it at all, update the app twice before you conclude anything about your plan.
Astra is a real step up on hard, long, multi-step work. It is not a better chatbot, and treating it like one is the fastest way to spend a month's allowance on a Tuesday.
Working out which subscription this affects? Which ChatGPT plans get Astra has the surface table, GPT-6 Astra vs Claude Fable 5.1 has the benchmark read, and GPT-6 Astra vs Grok Bot covers the model-versus-teammate question. More of how I run all of this at Claude at Work.
Published September 10, 2026. The effort-level guidance, the Chat-versus-Work split and the update procedure were read and captured that day from OpenAI's help center article "Managing usage with GPT-6 Astra in Work and Codex," and the two screenshots of it above are from that capture. The allowance figures are summarized here and screenshotted in full on the limits post linked above. The launch description and context window come from OpenAI's September 3 announcement, cited as of September 4 when this site first captured it. The account screenshots are my own paid Pro plan, sidebar cropped. My missing-Astra screenshot documents one account on one day and one app build. It is not a claim about OpenAI's rollout.
Quick answer: If your boss is offering to pay, take the money. Then spend it on capability per person rather than on an admin console. For two to five people, three individual Pro plans cost less than three Team seats, and a Team Standard seat treats Fable exactly the way Pro does (on usage credits), so the extra $5 a seat buys admin, not model access. Team is worth buying when billing, access control and shared projects have become a real problem, not before.
Two things the internet still gets wrong about this, both verifiable in thirty seconds: Team does not require five seats, and Premium seats are not $150.
Here is the math nobody on the first page of results has actually done.
The short answer, by who you areYou are
Buy this
WhyTwo or three people, no IT department
Individual Pro plans
$5 per seat cheaper, same model accessSomeone whose boss will expense it
Individual plans, and ask for more
Spend on capability, not seatsHandling client material under contract
Team
Model training is none by default, not opt-outChasing expense reports around the office
Team
Central billing is the actual productNeeding shared projects across people
Team
Individual plans do not have project sharingA whole department, 10 or more
Team, mixed seats
This is where the admin genuinely pays for itselfThe math, with today's numbers
Three people. Every figure from Anthropic's pricing page, read September 5, 2026.Option
Monthly
Annual equivalent
Fable treatment3 × Pro
$60/mo
$51/mo
Usage credits3 × Team Standard
$75/mo
$60/mo
Usage credits, same as Pro3 × Max 5x
$300/mo
monthly only
Included, up to 50% of weekly limits3 × Team Premium
$375/mo
$300/mo
Included, same as MaxLook at the first two rows. Three individual Pro plans cost $15 a month less than three Team Standard seats, and the model access on both is identical.
That is the whole contrarian case on this page, and it is just subtraction.
What the internet still gets wrong
The top Reddit thread on this query is a good record of what people believed, and most of it has since changed. Worth reading in that spirit rather than as current fact.
The original poster, u/Abeck72, talked himself out of Team like this:"At first I thought, 'easy, I'll just get a Team plan,' but then I realized the Team plan doesn't include Claude Code. To get access, you need Premium seats... Given that, wouldn't it make more sense to just get three individual 5× accounts?"He reached a reasonable conclusion from a premise that is no longer true. Anthropic's pricing page today lists "Includes Claude Code and Claude Cowork" directly under the Team plan, and the Team column of the feature table marks Claude Code as available.
A reply from u/kondadotm captured the other half of the folklore:"It is absurd that Claude Code is behind a SECOND paywall on the Team plan. It is absurd, too, that the MINIMUM for the Premium seats is 5 users. Why the f? It's like 'I don't want your money if it isn't a lot of money.'"Both halves of that are now wrong. Anthropic's page reads "For teams of 2 to 150." Two.
Here is Anthropic's Team card as it actually renders today. Three of the four corrections below are visible in this one image.Anthropic's Team pricing, captured September 5, 2026. "For teams of 2 to 150", both seat prices, and "Includes Claude Code and Claude Cowork" are all in frame.
So here is the corrected scoreboard, as of September 5, 2026:The claim you will read
What Anthropic's page says today"Team requires 5 seats"
"For teams of 2 to 150""Premium seats have a 5-user minimum"
No minimum stated beyond the 2-seat floor"Premium seats are $150"
$125 monthly, $100 annually"Team doesn't include Claude Code"
"Includes Claude Code and Claude Cowork""You pick one seat type for everyone"
"Mix and match seat types"If you were talked out of Team by any of those, the reason you were given has expired. You may still not want it, but you should decline it for a current reason.
What Team actually buys you: administration
Here is the thing to hold onto, and it is my rule for this one: Team does not make Claude better. It makes Claude manageable.
The capability rows are identical. Same models, same 200K context window, same Claude Code, same Cowork. What Team adds sits in a different category entirely:Central billing and administration, instead of five people expensing $20
Single sign-on and domain verification
Admin controls for remote and local connectors
Usage analytics across the organisation
Enterprise search across your team's content
Organisation-wide skills deployment
Project sharing and collaboration, which individual plans do not have
Adding seats midtermThat project-sharing row is easy to miss and matters more than it looks. On individual plans, projects are yours alone. Team is where a project becomes something a colleague can open.
And one row that a consultancy should read twice: model training is "None by default" on Team, against "Opt-out" on individual plans. If you handle client material under contract, that difference is worth more than the seat price, and it is the strongest argument on this page for buying Team early.
Fable on Team: Anthropic's own two pages disagree, and one is clearer
This one is worth slowing down for, because it is easy to get wrong and I nearly did.
On the pricing page's feature comparison, the Fable row shows a bare No in the Team column. Read alone, that says Team seats cannot use Anthropic's most capable model at all.
That is not what happens.
Anthropic's Fable plan article is more precise, and it sorts by seat type rather than by plan:Pro plans and standard seats on Team plans: Fable "aren't included in your plan's usage limits. You can use them with usage credits."
Max plans and premium seats on Team plans: Fable is "included as a standard part of your plan," up to 50% of weekly usage limits.So a Team Standard seat behaves exactly like Pro on Fable, and a Team Premium seat behaves exactly like Max.
Which actually makes the math cleaner rather than messier. A Team Standard seat costs $5 a month more than a Pro plan for identical model access. You are paying that $5 for administration, and nothing else.
If you see a page telling you Team cannot touch Fable, it is reading the pricing grid and not the help article.
Take the money. Just spend it well.
Now the part I actually believe, because I want to be clear I am not telling you to turn down budget.
If someone told me their boss would pay for three people, I would absolutely take it. That is a no-brainer. Tooling that makes people more productive is the easiest yes in business.
To be clear about what I actually said: that answer is about getting the team onto Claude at all, not about which SKU to buy. The "buy individual plans instead of Team seats" conclusion comes from the seat math above, not from running a Team plan myself, which I have not.
The reason is not "AI is good." It is that you can take one person and multiply them.
Here is what that looks like in my own work. If I handed an editor my video editing skill, they would save enormous time on every cut. But the saving is not the point. The point is what the saved time becomes.
Say they are editing by themselves and a long-form video takes them three to four hours. Now it takes one to two.
They just got two hours back.
And now those two hours can go into watching the analytics and working out which visuals and which hooks are actually working, and why. That is the superpower. Not "edits faster." One person who now also does the thinking nobody had time for.
So yes, spend your boss's money. The only argument this page is making is about the vehicle: three individual plans deliver that same multiplication for less than three Team seats, and you can move to Team the day the admin becomes the bottleneck.
When Team is genuinely the right call
I have spent most of this page arguing the other way, so here is the honest case for buying it.
You have more than five people. The expense-report tax is real and it compounds. Somewhere around five or six seats, central billing stops being a nice-to-have.
Somebody has to leave the company one day. With individual plans, that person's account and its history walk out with them. With Team, an owner manages the seat.
Your contracts say something about training data. None-by-default beats opt-out when a client asks you to put it in writing.
People need to share work. Project sharing and collaboration is a Team row and there is no individual-plan workaround.
You need SSO or connector controls. These do not exist below Team at any price.
Above Team sits Enterprise, at $20 per seat plus usage billed at API rates, which adds SCIM, audit logs, role-based access and custom data retention. If you are reading this page you are almost certainly not there yet.
One user in that thread, u/frythan, pointed at the commercial terms specifically, noting that they call out that you own your content. That instinct was right, and it is the underrated reason to buy Team early if you do client work.
If you do move, use the mix-and-match rule: put your two heaviest users on Premium seats and everyone else on Standard, rather than buying the expensive tier for the whole room.
How I checked this
Every price, seat band, feature row and the Fable gap come from claude.com/pricing, read on September 5, 2026, including the Team feature comparison table. The seat math is that page's numbers multiplied by three.
I have never run or administered a Claude Team plan. I pay for Claude individually, so nothing on this page is a first-hand report of the admin console, seat management or how Team usage behaves in practice. What I can tell you is that my recollection of the structure, that each person gets their own account and someone gets an admin layer over the top, checks out against the current feature table.
The Reddit quotes are dated and I have flagged them as historical rather than current. Anthropic changed the seat floor and the Premium price after those threads were written.
So should you buy it?Two or three people, no compliance pressure: individual Pro plans. $5 per seat cheaper for the same model access.
Your boss offered to pay: say yes immediately, then buy individual plans and ask for the difference in something else.
You handle client data under contract: Team, for the none-by-default training terms alone.
You need shared projects or SSO: Team. There is no workaround below it.
More than five people: Team, mixed seats, heaviest users on Premium.
Someone needs Fable inside their plan limits rather than on credits: that person needs Max, or a Team Premium seat. Both cost $100.If individual plans are the answer, the next question is which one: is Claude Pro worth it covers the $20 decision, Claude Pro vs Max covers the $20-to-$100 step, and is Claude Max worth it covers whether the top tier earns it. If limits are the reason you are shopping, start with Claude usage limits explained.
Current plans and prices sit on the AI plan tracker.
This post is part of Claude at Work, the hub for using Claude at your job without code.
Published September 5, 2026. Seat prices, the 2-to-150 seat band, Claude Code inclusion, the Fable row and the model-training terms all checked that day against Anthropic's pricing page, linked inline. Anthropic changes these quietly, so open the live page before you buy.
Quick answer: Probably not at $200, and maybe at $100. Pro is worth it when waiting has started costing you more than the money does, and not one day sooner. There are two Pro tiers now and the honest question is not yes or no, it is which rung. I paid $200, dropped to $100 to fund something else, and did not miss it.
Every page ranking for this question is answering the $200 version. Nobody is going to win that argument, because for almost everybody the answer is obviously no.
The $100 tier is the one that actually changes the math, and nobody has written about it.
The short answer, by who you areYou are
Buy this
WhyAsking questions, drafting, summarising
Plus, $20
You will not hit the ceiling. Save your moneyHitting limits occasionally and you can wait
Stay on Plus
Annoying is not the same as expensiveGetting stopped mid-task on work with a deadline
Pro $100
5x Plus usage. This is the real first stepStill stopped on $100, running agent work daily
Pro $200
20x Plus. The top of the ladderWant GPT-5.6 Sol Pro
Pro, either tier
The one thing Plus genuinely cannot buyTempted by Pro but also want a second AI
Two $20 plans
Two allowances, two models. Often the better tradeThere are two Pro tiers, and the whole SERP is a year behind
This is why the other pages are useless to you.OpenAI's own Pro tiers help article, captured September 5, 2026. The tier split is stated in plain language and most ranking articles still do not mention it.
From OpenAI's Pro tiers article, verbatim:"Both Pro tiers include the same core capabilities. The main difference is usage allowance: Pro $100 unlocks 5x higher usage than Plus, while Pro $200 unlocks 20x usage than Plus."And on whether the expensive one changed:"No. The $200 Pro plan remains the highest usage tier. The $100 plan simply adds another option."So when someone writes "is ChatGPT Pro worth $200," they are answering a question about the ceiling of a product whose entry point is now $100. The step up from Plus is $80, not $180.
What the $100 actually gets you
I will give you the unglamorous answer, because it is the true one.
It is mostly more usage. You get to use the strongest models for longer, and room to experiment without watching a meter. The genuine additions are the two Pro-only models, the Extra High reasoning setting, a bigger context window and expanded Codex, and those matter to fewer people than the headroom does.
That sounds like a letdown until you have been stopped mid-task. Then it is the only thing you want.
The precise version, from OpenAI's pricing comparison read on September 5, 2026:Plus, $20
Pro $100
Pro $200Usage vs Plus
baseline
5x
20xGPT-6 Astra in Chat
No
Yes, as GPT-6 Pro
Yes, as GPT-6 ProGPT-6 Astra in Work and Codex
Yes
Yes
YesGPT-5.6 Sol
Yes
Unlimited*
Unlimited*GPT-5.6 Sol reasoning levels
Medium and High
+ Extra High
+ Extra HighPro models (GPT-6 Pro, Sol Pro)
No
50/week, shared across both
200/week GPT-6 Pro, 170/day Sol ProInstant context window
54K
128K
128KReasoning context window
256K
400K
400KCodex
Yes
Expanded
ExpandedAnnual billing
none
none
noneNotice what is identical between the two Pro tiers. Everything except the size of the bucket.
And notice what you already have on $20. You get a lot on the $20 plan. You get Sol, you get Codex, you get deep research, and you get Astra in ChatGPT Work and Codex, just not in the chat box. The main difference is mostly not what you can do, it is that you are going to run out.
The one row that is a genuine capability gate is the Pro models. Those come with published message counts rather than an unlimited badge.
On $100 it is one shared allowance of 50 messages a week across GPT-6 Pro and GPT-5.6 Sol Pro; switching between them does not give you more. On $200 they are metered separately: 200 GPT-6 Pro messages a week, a separate 170 a day for Sol Pro, and a combined ceiling of 200 a day. Worth knowing before you upgrade for the frontier model.
The receipt: I went down, not upMy ChatGPT app, September 5, 2026: Pro plan at $100 a month, and 70% of the week's general usage already gone with two days to the reset. This meter is the whole decision.
Here is my actual history with this, because I have been on all three rungs and the direction of travel is the part nobody writes about.
I had the $20 plan and I ate my own cooking. I used it until it ran out. It was painful, and I sat there thinking I either need to figure out how to make do with the twenty or I need to upgrade.
I eventually upgraded. I went to the $200 plan.
What pushed me there is worth telling properly, because it was not a feature announcement. I had a video editing skill that was not working, and I asked Codex to review it. It went and built something. That was my first time using Codex and honestly I did not get the hype at first; I did not think it was good enough.
Then Sol came out and it started getting impressive. It was actually doing the thing. Actually building the skill.
I got a lot of use out of that on a $20 plan. Then I hit the limit, repeatedly, and that is when it was time to move.
Then I scaled it back down to $100. Not because Pro disappointed me. Because Grok Bot came out and I needed it, and the money had to come from somewhere.
I have not regretted it since.
That is the part I want you to take: the right tier is not the best tier, it is the one that leaves budget for the other tool you actually need. You have to figure out your use cases.
The pain test, made specific
My rule for this is simple: let the pain tell you. Here is what that means in a normal week, because "upgrade when you need to" is useless advice.
You will not hit the limit asking it things. Rewrite this email, what should I wear, summarize this article, help me think through a decision. That is what the $20 plan is sized for and you can do it all day.
You start burning real credit when you ask it to build something. Help me make a website. Help me edit a video. Make me a video. That is a different class of work and it drains the allowance fast.
So the moment to watch for is specific: you are sitting there unable to continue on something you wanted to finish, and waiting is costing you.
Not a warning banner. Not a percentage. Just you, stopped, on a weekday, on work that mattered.
If waiting two hours is merely annoying, you are not ready. If waiting two hours means a client does not get their thing today, you are.
When Pro is definitely not worth it
The sharpest disqualifier I have seen on this whole question came from u/positive on r/ChatGPTPro:"If gpt-5 thinking is already pretty good for your tasks, gpt-5 pro will be either the same or better but slower. If gpt-5 thinking is not good at your task, gpt-5 pro won't be any better."Sit with that, because it kills the most common reason people upgrade. If the model is already handling your work, paying more makes it slower, not smarter. If the model is failing at your work, paying more will not rescue it.
Pro buys capacity. It does not buy comprehension.
The speed cost is real too. u/petermalik01, who is otherwise a fan and says the top tier is worth $200 for genuinely complex legal problems, adds the catch in the same breath: the top model is "slow" and runs "~10–15 min per reply" in their experience.
Three other things Pro will not do for you.
Support cannot reset your limits. OpenAI states it flatly: "OpenAI Support does not reset ChatGPT or Codex usage limits." Nobody writes this down and it is the most useful operational fact on the page.
"Unlimited" has an asterisk. Individual models keep separate allowances, and those allowances differ between the $100 and $200 tiers. A model can go temporarily unavailable on Pro.
There is no annual discount. Per the same help article, OpenAI does not support annual billing on Go, Plus or Pro. You cannot commit your way to a lower rate.
Where Plus genuinely wins
Plus is the right answer for most people reading a page called "is ChatGPT Pro worth it," and I would rather say that than sell you an upgrade.
It has Astra. It has Sol. It has legacy models, Codex, deep research, projects and scheduled tasks. It is not a crippled tier, it is the same product with a smaller bucket.
And there is an option the comparison pages skip entirely: $20 here plus $20 somewhere else often beats $100 in one place. Two separate allowances, two different models, two sets of strengths.
That is roughly the shape my own stack has taken. I run ChatGPT with Codex, Claude, and Grok with Grok Bot, and the $200 ChatGPT tier lost its slot to the third of those rather than to anything OpenAI did.
How to try Pro cheaply
The no-annual-billing rule cuts in your favor here, so use it.
You switch tiers in Settings then My Plan. Upgrades take effect immediately with billing adjusted automatically. Downgrades take effect at your next renewal and you keep your current plan until then.
That means one month of Pro $100 costs you $80 to find out for certain. Nothing locks. If it was not the problem, you drop back at renewal.
Eighty dollars is cheaper than another month of guessing. It is also cheaper than what I did, which was jump straight past the middle rung to $200 because I assumed I would need it.
How I checked this
The two Pro tiers, the usage multiples, the support policy, the billing rules and the switching mechanics all come from OpenAI's Pro tiers help article, read September 5, 2026. Models and context windows come from chatgpt.com/pricing as it rendered in a browser the same day, since OpenAI serves those values client-side.
The tier history is mine: $20, then $200, then down to $100, where I still am. I have not run a controlled comparison of $100 against $200. OpenAI publishes no message counts for ordinary chat, so I am not going to give you a number for how much of a week either tier absorbs; anyone who does is guessing. The Pro-model allowances above are the exception, and they are quoted exactly as OpenAI states them.The 60-second version, from my Shorts: My AI limit stopped running out (the routing loop).So is it worth it?Your work is questions, drafting and everyday tasks: no. Stay on Plus and stop reading upgrade pages.
You are building, editing or researching hard and getting stopped: yes, at $100. That is what the 5x is for.
You are stopped even on $100 and AI is doing real work in your day: yes, at $200.
The model is already failing at your task: no. More capacity will not fix a comprehension problem.
You want Pro and a second AI tool: buy the second tool first. That is the trade I made and I would make it again.If the comparison you actually want is the tier below, that is ChatGPT Plus vs Pro. If Pro is settled and only the price is not, that is Pro $100 vs $200. And if you are weighing this against the other side of the market, Claude Max vs ChatGPT Pro is the $100-against-$100 version.
Current plans and prices live on the AI plan tracker.
This post is part of Claude at Work, the hub for using AI at your job without code.
Published September 5, 2026. Tiers, usage multiples, models, context windows and billing rules checked that day against OpenAI's Pro tiers help article and chatgpt.com/pricing, both linked inline. OpenAI changes these often, so open the live pages before you buy.
Quick answer: These are not rivals, they are two different purchases. Gemini is a bundle you probably already own through a Google account, Workspace or a phone plan.
Grok is a $30 agent you deliberately buy, and what you are buying is Grok Bot running jobs while you are not there. So the real question is not "which is better." It is "do I need to add Grok."
Full disclosure before you read another line: I pay for Grok and I do not use Gemini.
I am not going to pretend that is neutral. What I will do is judge Gemini on its published facts and on the one real session I ran, show you both receipts, and tell you exactly what would make me switch.
The short answer, by who you areYou are
Pick
WhyAlready inside Google all day, Gmail and Docs
Gemini, and check what you already have
You may be paying for it and not knowWant better answers to questions you ask
Either, honestly
Both are good at this now. Not worth $30Want research reports without babysitting
Gemini free tier first
Deep Research is on the free planWant jobs that run on a schedule without you
Grok, $30 SuperGrok
This is the actual product differenceChasing the top of the benchmarks
Neither, on principle
It changes every six weeks. Pick on workflowThe two ladders, side by side
Both price lists in one place, read September 5, 2026.Grok (xAI)
Gemini (Google)Free
$0, web and X search, voice
$0, Gemini 3.6 Flash, Deep Research, Canvas, LiveEntry paid
SuperGrok $30, unlocks Grok Bot
AI Plus $4.99, Gemini in Gmail and VidsMid
SuperGrok Plus $100
AI Pro $19.99, Docs, Flow, 5 TBTop
Lite and Heavy, no published price
AI Ultra $99.99 (5x) / $199.99 (20x)The thing you are really buying
Agents that run jobs on a cloud computer of their own
A model already wired into your documentsAlready included with
nothing
Google One, Workspace, some carrier plansThat table is the whole point. Google's ladder starts lower and is often already paid for. Grok's starts at $30 because the $30 is the agent.
The fork nobody writes about: bundle versus deliberate buy
Every other page on this search runs a benchmark table. That is the wrong axis.
Gemini arrives. It is in your Gmail, in Docs, in your Google account, and it is bundled into Workspace plans and some carrier deals.
I have it in a browser tab right now because it comes with my Verizon subscription. I did not choose it. It showed up.
Grok does not arrive. You go to x.ai and you pay $30.
That one difference decides more than any benchmark. One of these you are already carrying, and the other has to justify a new line on your card every month.
What $30 of Grok actually gets you
From x.ai's pricing page, read September 5, 2026.Plan
Price
What it addsFree
$0
Real-time web and X search, voice mode, connectorsSuperGrok
$30/mo
Grok 4.6, Grok Bot access, connectors, higher rate limits, Expert, image and video generationSuperGrok Plus
$100/mo
Everything above, plus 1080p video, significantly higher usage across Chat, Imagine, Voice and Build, priority accessSuperGrok Lite
not published
Appears in the comparison table with no price anywhereSuperGrok Heavy
not published
SameTwo things worth flagging.
The Lite and Heavy tiers have no published price. They sit in xAI's own comparison table with buttons that take you to a checkout. That is genuinely under-reported and you should know it before you assume $30 and $100 are the whole ladder.
Grok Bot runs on a cloud computer of its own, not your laptop, and its usage is listed as its own included allowance. xAI's Grok Bot page lists "Grok Bot's own computer" and "Weekly Grok Bot usage included" as line items, and says Bots keep working "24/7, even when your laptop is closed." Note the singular: per xAI's docs all of your Bots share that one computer rather than getting one each. I have seen the separate-pool claim repeated a lot; what xAI actually publishes is the line above, so that is what I will stand behind.
Two smaller things that confuse people searching "Gemini vs Grok" in either order.
The $8 X Premium thing is not this. X Premium raises your Grok usage limits inside X. It is not a SuperGrok plan and it does not include Grok Bot. The direct plans are the $30 and $100 ones above.
Grok Build is separate from Grok Bot. Build is xAI's app-building surface; Bot is the agent that runs jobs. My SuperGrok covers both, and Build is the one I have not properly played with yet.
Also worth knowing: x.ai now trades as SpaceXAI LLC. Same Grok, same pricing page.
What Grok actually does in my week
This is the part I can speak to first-hand, and it is why the $30 stays on my card.
I do not use Grok as a chat window. Just having another chat AI is not useful to me anymore. It is fine on the go, but for that I use ChatGPT, and sometimes Claude. For actual workflows I use Grok Bot, and that is a different product wearing the same brand.
I run three folders: growth, publishing, and intelligence.
Growth holds the scrapers. There are agents I named Shorts Hawk, Substack Hawk and Plan Watch, and their whole job is to scrape my own analytics three to four times a day. Shortly after a post goes out, then again through the day. They watch whether anything is popping, whether something is over-indexing, and they track it across the week.
Then they drop it in a ledger.
That ledger is the actual point. It is not "here are your numbers," it is "Chris, this is working, do more of this next week." I stopped posting into the wind and started running experiments I can read back.
Publishing does something I could never automate before: it uploads my long-form YouTube videos for me. Same for my Substack notes.
Intelligence runs a news scout that scrapes X, plus an agent that watches YouTube videos from people I follow and summarizes what is worth applying.
I authenticated once and they have been running since. I had assumed each sub-agent got its own machine; per xAI's Grok Bot docs that is not how it works. "All of your Bots use the same persistent cloud computer. They share files, browser sessions, and app logins," and the computer is scoped to your account rather than to a single Bot.
Which is exactly why authenticating once was enough. Worth knowing before you assume the Bots are walled off from each other, because they are not.
It feels like a little mini swarm. A small team doing its thing.
I would not hand that to Gemini, because Gemini does not do it. That is not a knock on the model. It is a different shape of product.
For what it is worth, this is roughly where the people who have tested both land too. VKTR's editorial team, who ran all three of ChatGPT, Gemini and Grok head to head, framed the choice the same way I do: as a question of which one fits the work rather than which scores highest.
What Gemini gets you, at every price
From Google's subscription page, verified September 5, 2026.Plan
Price
What you getFree
$0
Gemini 3.6 Flash, varying access to 3.1 Pro, image generation, Deep Research, Gemini Live, CanvasGoogle AI Plus
$4.99/mo
200 Flow credits, Gemini in Gmail and Vids, custom tool creation, 400 GB storageGoogle AI Pro
$19.99/mo
Gemini in Gmail, Docs and Vids, Flow with 1,000 credits, 5 TB, YouTube Premium LiteGoogle AI Ultra
$99.99/mo
5x higher limits than Pro, first access to Deep Think and Gemini Spark, Project GenieGoogle AI Ultra 20x
$199.99/mo
20x higher limits than ProThe correction that saves people money: Deep Research is on the free tier. Google lists it right there in the free plan's feature set. If research reports are what you wanted Gemini for, you may not need to pay anything at all.
The correction that costs people money: Deep Think and Gemini Spark are Ultra features. They are listed under the $99.99 plan as "first access," not under the $19.99 one. A lot of people upgrade to AI Pro expecting Deep Think and do not get it.
I ran one real session, and it was not bad
I do not use Gemini, so I am not going to review it. But I did sit down and actually use it for a few minutes rather than write about it from a spec sheet.
I gave it a setup prompt: structure the visual hierarchy and site architecture for a cookie showcase website, outline the pages, navigation flow and content blocks needed for launch.My actual prompt and the answer that came back, running on Flash. It answered fast.
Honestly? A little verbose. It gave me a page map table when I would have taken three lines.
Then I asked it to draw it.Same session, one follow-up. It generated an actual site mockup with the page flow drawn between screens.
That is pretty impressive. And it was fast.
So no, it is not a bad model. I want to be clear about that, because "I do not use it" and "it is not good" are different statements and the internet keeps confusing them.
I do not use it because I do not have a specific use case for it. I already have Codex, Claude and Grok. That is genuinely too many options for one person to run well, and adding a fourth without a job for it is how you end up paying for four things and getting good at none.
My wife uses Gemini. It is fine.
What would actually make me switch
I want to answer this properly rather than leave it as a shrug, because it is the only honest version of a verdict I can give you.
I would switch if it became more powerful than Fable, or just as powerful for less money. That is the whole condition. Not vibes, not a benchmark chart, not a launch video.
And I would not bet against it. I watched Grok go from a thing I did not use and thought was laughable to one of my main drivers. Google has the ability to come back the same way, and the benchmarks are trending in the right direction.
I also genuinely like some of their other products. Notebook LM is good. I just do not reach for it often.
So am I opposed to Gemini? No. I have too many options right now, and that is a real constraint, not a verdict on the model.
Where Gemini genuinely wins
Let me make the case I am not naturally inclined to make.
It is already paid for. If your employer has Workspace, or you have Google One, or it came with your phone plan the way it came with mine, the marginal cost of using Gemini today is zero. Nothing on this page beats free.
It is where your documents already are. Grok is not in your Gmail and it is not in your Docs. For a lot of office work, being one click from the file beats being a slightly different model.
Deep Research is on the free tier. Grok has nothing at $0 that competes with that.
Notebook LM is a genuinely good product and it has no real equivalent on the Grok side.
If you already have Gemini and your work is documents, email and research, the correct answer to "should I add Grok" is probably no.
How I checked this
Grok prices, tiers and the Grok Bot inclusion come from x.ai's pricing page, read September 5, 2026. Gemini prices and the tier placement of Deep Research, Deep Think and Gemini Spark come from Google's subscriptions page, read the same day.
The Grok Bot workflow is my own setup, running since late August 2026. The two Gemini screenshots are one session I ran on September 5, 2026, with the picker set to Flash. The account comes bundled with my Verizon plan, so I have not checked which tier it resolves to. That session is the only Gemini use behind this page. I have not tested Gemini's paid tiers, I have not benchmarked either product, and I have not compared answer quality across them at any scale. Where I have an opinion, it is labeled as mine.The 60-second version, from my Shorts: Grok builds the app from one sentence and hands you the link.So which one should you buy?You already have Gemini through Google or a carrier: use it, and do not add anything until it fails you at a specific job.
You want research reports for free: Gemini's free tier. Deep Research is included.
You were about to buy AI Pro for Deep Think: stop. That is a $99.99 feature.
You want jobs that run without you: Grok, $30. That is what Grok Bot is and nothing on Google's ladder is the same shape.
You have three AI subscriptions already: do not add a fourth until one of them stops doing a job you need.If Grok is the side you are leaning toward, what is Grok Bot covers the agent properly and is Grok Pro worth it covers the money. For the cross-vendor version of this decision, try Grok vs Claude or Claude vs Gemini.
Every plan and price I track sits on the AI plan tracker.
This post is part of Claude at Work, the hub for using AI at your job without code.
Published September 5, 2026. All prices and tier contents checked that day against x.ai and Google's own subscription pages, linked inline. Both companies change these often, and two Grok tiers still have no published price, so open the live pages before you buy.
Quick answer: Use Sonnet. Anthropic's own guidance says that if you are not sure which model to pick, start there, and for writing, analysis and everyday multi-step work it is the right call.
Save Opus for problems you have already watched Sonnet struggle with. The question is not which model is smarter. It is which is the cheapest one that still gets your job right.
And the question itself is out of date. There are not two models. There are four.
That matters more than it sounds, because picking the heavy one by reflex is the single most common reason people hit a limit and think Claude is broken.
The short answer, by the job in front of youThe job
Pick
WhySummarize this, pull the dates out, quick lookup
Haiku 4.5
Instant, and the lightest on your limitWrite it, analyze it, work through it, most things
Sonnet 5
Anthropic's stated default. Start hereYou already tried Sonnet and it missed things
Opus 5
Reasoning specialist. Costs more of your limitLong project, many connected steps, few check-ins
Fable 5.1
Heaviest, slowest, and on Pro it costs creditsYou are hitting limits constantly
Go down a model
Not up. This is almost always the fixSonnet vs Opus: start with Sonnet, and there are four models, not two
Whether you searched "Opus vs Sonnet" or "Sonnet vs Opus", the answer is the same, and Anthropic states it outright: start with Sonnet.
The other half of the answer is that the two-model question is out of date. Here is Anthropic's own table, read September 5, 2026.Model
Rate limit use
Anthropic's stated best forHaiku 4.5
Lightest
"Quick answers, summaries, and simple extraction"Sonnet 5
Moderate
"Coding, writing, analysis, and multi-step workflows... your versatile default"Opus 5
Heavy
"Deep research and complex reasoning that genuinely needs sustained thinking"Fable 5.1
Heaviest
"Your largest, most critical projects: long, complex tasks"Read the middle column, not the right one. That is a price list.
Every ranking page treats this as a quality ranking where Fable is the good one and Haiku is the compromise. Anthropic does not describe it that way. It describes four tools with four costs, and it names Sonnet as the one to reach for when you have not thought about it.
There is a fifth Claude model, Mythos, built for cybersecurity and biology research. It is restricted to vetted organisations through Anthropic's trusted access programmes, so it will not appear in your picker and it is not part of this decision.
Why you keep hitting your limit
This is the section I would send to most people instead of the rest of the page.
Anthropic says it directly: "if you use Opus or Fable on a task that Sonnet or Haiku could handle, you may be using more of your limit unnecessarily."
Your limit is not a message count. It is a token budget running across two windows at once, a rolling five-hour session and a weekly cap.
A heavy model on a light task spends more of that budget for an answer that was not better. Anthropic does not publish how much more, and I am not going to invent a multiplier. The one number it does print is on the Effort control below.
The cleanest version of this I have seen came from a non-coder on r/claude. u/TeachMeThings3209067 wrote:"Yeah I was using Opus 4.6 frequently and hitting my usage limits. Then I realised sonnet was actually good enough for what I wanted to do which was just regular analysis... I am literally just analysing transcripts from meetings."They fixed a limit problem by going down a model. Nothing else changed.
The original poster in that thread, u/LinkDaSquid, landed in the same place:"I even asked Claude itself if I should use Opus instead of Sonnett, and it said to just keep using Sonnett. As a result, my weekly limit barely gets above like 15%."The control nobody mentions: Effort
Open your model picker and look under the model name. There is a second setting, and it changes your limit consumption on the model you already chose.My own Claude picker, September 5, 2026. Note the warning label on Max. Anthropic prints the cost right there and almost no comparison page mentions this setting exists.
Five levels: Low, Medium, High as the default, Extra, and Max. The app flags Max with "1.5x or more usage" in orange, which is Anthropic telling you the price before you pay it.
That gives you a cheaper move than switching models. If Sonnet is close but not quite getting there, raise Effort before you jump to Opus. If Opus is working but draining you, drop Effort before you drop the model.
My actual routing rule
This is how the work gets assigned in my week. It is not a recommendation for your setup, it is what I do.
Fable is for complex skills. Editing my video skill, generating content for my website, running my SEO work. Anything long and genuinely complicated, that is Fable, hands down.
Opus is my everyday driver. Making shorts for YouTube, planning, thinking something through. When I want to plan something out, that is Opus. No reason to spend Fable's budget on it.
Sonnet and Haiku I do not personally use much anymore. I want to be straight about that rather than pretend to a routing rule I do not run.
One caveat that matters if you are on Pro: I am on Max 20x. Running Fable and Opus as freely as I do is affordable at $200 and would not be at $20.
But the history is the useful part. Before Fable existed, my split was Opus and Sonnet: Opus did the planning and the main execution, and Sonnet did all the grunt work, the research, the fetch-and-summarize. That split was good. There are just so many models now that I stopped reaching for the light end.
If you are on Pro and watching your limit, that old split is better advice than what I currently do.
When the heavy model is the wrong tool
Here is the scar, and it is a cheap one to avoid.
Say you are writing a script. You are not going to use Fable for that. It might nail it, but it is burning tokens for no reason, and it is not even going to be faster.
It will probably be slower.
That is not a feel. Anthropic says the same thing in its own model guide: Fable "takes time to think through problems before answering, so responses take longer, and it uses the most of your rate limit." You pay twice, in waiting and in allowance, for an answer a lighter model would have handed you.
One genuine exception worth knowing: for biology and security topics, Anthropic states that "Claude answers these topics with Opus even if you've picked Fable." If that is your work, picking Opus yourself is simpler than being quietly rerouted.
Where Opus genuinely earns it
I have spent this page talking you down the ladder, so let me be fair to the top of it.
Opus is not Sonnet with a bigger bill. Anthropic describes it as built for problems that need sustained thinking over time, and its own worked example is analyzing complex research papers: long specialised documents, methodology critique, conclusions you will act on.
The best description of the gap came from u/roselan on r/claude:"Everyone can make a good sandwich, but Opus is the 3 Michelin stars Chef."Which is exactly right, and also exactly why you should not order from him every day.
u/Far-Pomelo-1483 gave the compressed version: "Only use opus if sonnet can't do it."
The test that settles it in ten minutes
Do not take my routing rule or anybody else's. Anthropic publishes a method and it is better than an opinion.Pick a task you already know the answer to. A report you have actually read, a document you wrote.
Run it on the lighter model first.
Start a fresh chat and run the identical prompt on the heavier one.
Compare where the answers differ, not how long they are.That last instruction is Anthropic's own, and it is the part people get wrong. A longer answer feels better and usually is not. You are checking whether the lighter model missed anything you would have caught yourself.
If it did not miss anything, that task is Sonnet-shaped forever. Bank it and stop paying for the upgrade.
What changes on Free, Pro and Max
Model access is gated by plan, and one row surprises people.Plan
Haiku
Sonnet
Opus
FableFree
Yes
Yes
No
NoPro, $20
Yes
Yes
Yes
Usage credits onlyMax 5x and 20x
Yes
Yes
Yes
Included, up to 50% of weekly limitsTeam standard seat
Yes
Yes
Yes
Usage credits onlyTeam premium seat
Yes
Yes
Yes
Included, up to 50% of weekly limitsThe Fable row is the one to notice. On Pro you can use it, but it is not inside your plan limits, so it bills separately on top of the $20. On Max it is included up to half your weekly allowance. Verified against Anthropic's pricing page on September 5, 2026.
How I checked this
The model names, rate-limit ordering, stated use cases, the Effort guidance and the biology and security behavior all come from Anthropic's model selection guide, read September 5, 2026. Plan gating and the Fable rows come from claude.com/pricing, read the same day. The picker screenshot is my own machine on that date.
I have not benchmarked these models against each other and I am not going to pretend otherwise. Anthropic publishes no speed or quality numbers for chat, so anything you read that calls this a measured contest is somebody's impression dressed up as data. The routing rule above is mine and I have labeled it as mine.The 60-second version, from my Shorts: Claude's 4 models in plain English (and the one you're wasting).So which one should you open?You are not sure: Sonnet. That is Anthropic's answer and it is the right one.
You do not write code and your work is documents, analysis and writing: Sonnet, and you will rarely need more.
Sonnet tried and visibly missed things: Opus, for that task specifically, not as your new default.
You are hitting your limit constantly: go down a model or drop Effort before you spend anything.
You are running one long complex build with few check-ins: Fable, and check whether your plan includes it before you start.If the limits themselves are the real problem, Claude usage limits explained covers the two ceilings, and Claude Pro vs Max covers whether more headroom is worth buying. If you have not paid for Claude at all yet, and Opus and Fable are the reason you are considering it, start with is Claude Pro worth it. If you are new to the current lineup, Claude 5 explained covers the models themselves.
This post is part of Claude at Work, the hub for using Claude at your job without code.
Published September 5, 2026. Model lineup, rate-limit ordering, plan gating and the Effort control checked that day against Anthropic's model guide and pricing page, both linked inline. Anthropic ships new models often, so check the picker before trusting any list of four.
Quick answer: Plus is $20 a month. Pro starts at $100. What the extra $80 buys is mostly room, not brains: OpenAI says the $100 tier gives you 5x the usage of Plus.
It also unlocks two Pro-only models and a bigger context window. If you are not currently hitting walls, none of that is worth $80 to you.
I pay for the $100 tier. I got there by hitting the ceiling on $20 over and over, not by reading a feature list.
So this page answers "ChatGPT Plus vs Pro" and "ChatGPT Pro vs Plus" the way I would answer it for a friend: what actually changes, what does not, and the one test that tells you whether to move.
The short answer, by who you areYou are
Get this
WhyAsking it questions, drafting, summarising
Plus, $20
You will not hit the ceiling doing this. Ordinary chat is cheapBuilding things with it: sites, video edits, long research
Watch your limits
This is the work that burns the allowanceGetting stopped mid-task, waiting for resets, on a deadline
Pro $100
5x the usage of Plus, per OpenAIRunning agent work all day and still hitting the $100 wall
Pro $200
20x Plus. This is the top of the ladderYou want GPT-5.6 Sol Pro specifically
Pro, either tier
The one thing $20 genuinely cannot buyAlready decided on Pro and choosing between the two prices? That is Pro $100 vs $200.
Pro is two prices, and most comparison pages only know about one
This is the single biggest reason the other pages on this search are wrong.
There are two Pro tiers, not one. From OpenAI's Pro tiers page:"Both Pro tiers include the same core capabilities. The main difference is usage allowance: Pro $100 unlocks 5x higher usage than Plus, while Pro $200 unlocks 20x usage than Plus."Read that carefully, because it is doing a lot of work. The difference between the Pro tiers is capacity only. Same models, same features, different ceiling.ChatGPT's pricing page, captured September 5, 2026. OpenAI renders these numbers in the browser rather than in the page source, which is why so many articles quote stale prices.
Almost every page ranking for this query compares $20 Plus against a single $200 Pro. That is a $180 gap. The real first step up is $80.
What the extra money actually buys
Here is the honest inventory, from OpenAI's own pricing comparison as it rendered on September 5, 2026.Plus, $20/mo
Pro, from $100/moUsage
baseline
5x Plus at $100, 20x Plus at $200GPT-5.6 Sol reasoning
Medium and High only
Medium, High and Extra HighPro models (GPT-6 Pro, Sol Pro)
Not included
Included, with a message allowanceGPT-6 Astra in Chat
No
Yes, as GPT-6 ProGPT-6 Astra in Work and Codex
Yes
YesGPT-5.6 Luna
Yes
Unlimited*Legacy models
Yes
YesInstant context window
54K
128KReasoning context window
256K
400KInstant input maximum
~40 pages of text
~250 pages of textReasoning input maximum
~320 pages of text
~680 pages of textCodex
Yes
ExpandedAnnual billing
none
noneThree things worth pulling out of that table.
Pro models are the real gate. GPT-6 Pro and GPT-5.6 Sol Pro are the two rows where Plus says no and Pro says yes. Everything else is a bigger portion of something Plus already has.
Plus tops out at High on the reasoning slider. Extra High is a paid-tier-above-Plus setting. If you have been wondering why your reasoning options look shorter than someone else's, that is why, and almost nobody writes about it.
The context window grows more than the token numbers suggest. Instant goes from 54K to 128K, which is about 2.4x on tokens, but OpenAI's own input maximums for those same rows read ~40 pages against ~250. Those are OpenAI's figures, not a conversion I did, and they do not scale linearly with each other. If your job is dropping long documents in and asking questions about them, read the pages row rather than the tokens row.
The Astra question, answered per surface
This is where almost every page on this search goes wrong in one direction or the other, and the honest answer needs one extra word: where.
Availability is per surface, not per plan.
From OpenAI's GPT-5.6 and GPT-6 Pro article, updated the day before this page:"GPT-6 Pro, powered by GPT-6 Astra, is rolling out in ChatGPT for Pro $100, Pro $200, Business and Enterprise plans... Plus plans include GPT-6 Astra in ChatGPT Work and Codex as it rolls out."So both halves of the internet argument are wrong.
Plus is not locked out of Astra. It reaches it through ChatGPT Work and Codex.
But Plus does not get it in the chat box. In Chat, Astra is branded GPT-6 Pro, and that is a Pro, Business and Enterprise thing.
That distinction is the actual $80 question, and we have a whole page on it: which ChatGPT plans get Astra has the plan-by-surface table.
Pro's headline models come with a message countMy ChatGPT app, September 5, 2026: Pro plan at $100 a month, and 70% of the week's general usage already gone with two days to the reset. This meter is the whole decision.
Here is the fact that reframes the two Pro tiers, and it is the one thing OpenAI does put a number on.
GPT-6 Pro is not unlimited on either tier. From the same help article:Plan
GPT-6 Pro allowance in Chat
How Sol Pro shares itPro $100
50 messages per week
One shared 50-per-week allowance across both Pro modelsPro $200
200 messages per week
Separate 170 per day for Sol Pro, both capped at 200 per day combinedRead the $100 row twice before you buy it.
Fifty messages a week, shared across GPT-6 Pro and GPT-5.6 Sol Pro. Switching between the two models does not give you more once that allowance is gone.
That is roughly seven top-model messages a day. If the reason you are upgrading is the frontier model specifically rather than general headroom, the $100 tier is a much smaller purchase than "5x usage" makes it sound.
On $200, hitting the GPT-6 Pro weekly limit drops you automatically to GPT-5.6 Thinking at Medium rather than stopping you.
The upgrade trigger: let the pain tell you
This is my actual rule, and it is the only part of this page I would defend in an argument.
Let the pain be the reason you move up. Not the spec sheet. That is my rule for every AI tier I have ever climbed.
Here is how it went for me. I started on the free tier and used it until waiting for the reset became unbearable. Then I paid $20 and used that until waiting became unbearable again. Then I moved up. Every step was forced by friction I actually felt, never by a feature I read about.
What that looks like in a normal week is the useful part.
You are not going to hit the limit just asking it things. Should I eat this with my steak, what is the weather like, what should I wear, help me rewrite this email. The $20 plan handles that all day long. If that is your usage, you are shopping for a solution to a problem you do not have.
Where it starts burning is when you ask it to make something. Can you help me build a website. Can you help me edit a video. Can you create a video. That is when you start burning real credit, and that is when this becomes a serious workflow question instead of a chat question.
For me the differentiator was always testing new things. I am tinkering and documenting all day, so I outgrew tiers faster than a normal person would. That is my job, not a general truth about the plans.
What Pro does not fix
Two things people expect from the upgrade and do not get.
Support cannot rescue you. From OpenAI's Pro tiers article: "OpenAI Support does not reset ChatGPT or Codex usage limits." If you hit a wall you wait, or you switch models. There is no setting to bypass it and no favor to ask for.
There is no annual discount to soften the jump. OpenAI states it does not support annual billing or multi-month prepay on Go, Plus or Pro. You cannot commit your way to a better rate.
That second one cuts in your favor, though, and it is the reason I would tell anyone to just try it. There is no lock-in. You switch tiers in Settings then My Plan, billing adjusts automatically, and a downgrade takes effect at your next renewal.
One month of Pro costs you $80 to find out. That is a cheaper experiment than a week of guessing.
Where Plus genuinely wins
Not "wins on price." Wins, full stop, for most people reading this.
Plus has GPT-5.6 Sol, legacy models, Codex, deep research, projects and scheduled tasks, and Astra in Work and Codex. The feature list is not meaningfully shorter. It is the same product with a lower ceiling and one locked door.
Real users say this better than I can. On r/ChatGPTPro, u/alexoff put it plainly:"For everything else - ChatGPT Plus is usually enough. Also, try better prompting and you will see much better results in my opinion that would be enough even on a Plus plan."That last clause is the one to sit with. A lot of people who think they need more compute actually need a better prompt.
The counterweight, from the same thread, is u/peraltz94 explaining what made Pro worth it for them:"Having access to better models, more context, more juice, legacy models, and deep research and agent mode limits, were my factors. If I can get improved responses and save me time, and I can quantify to hours then it has enough value."Note the test they applied. Not "is it better." Can I quantify the saving in hours. That is the right question.
How I checked this
Everything factual on this page comes from OpenAI's own pages, read on September 5, 2026: the pricing comparison for models, context windows and prices, and the Pro tiers help article for the tier split, the support policy and the billing rules. The screenshot above is that pricing page as it rendered in a browser, because OpenAI serves those prices client-side and page scrapes miss them.
The upgrade rule and the usage story are mine, from paying for these tiers and moving between them. I have not run a controlled test of $100 against $200.
One caveat on numbers: OpenAI publishes no message counts for ordinary chat usage, so any general "X messages per day" figure is guesswork. The Pro-model allowances above are the exception. Those are published, and they are quoted here exactly as OpenAI states them.The 60-second version, from my Shorts: One workday of AI chat ate 31% of a monthly quota.So which one should you buy?You use it for questions, drafting and everyday work: stay on Plus. You will not hit the ceiling, and you already reach Astra through Work and Codex.
You are building, editing or researching heavily and getting stopped: Pro $100. That is the wall the 5x is for.
You are stopped even on $100: Pro $200, and at that point you are not asking this question anymore.
You need GPT-5.6 Sol Pro: Pro, either tier. It is the one thing $20 cannot buy.
You are not sure: stay where you are until waiting costs you something real. Then move. There is no annual lock-in, so the experiment is one month cheap.If you are weighing this against the other side of the market, Claude Pro vs ChatGPT Plus is the $20-against-$20 version, and Claude Max vs ChatGPT Pro is the $100-against-$100 one. If you have not paid for anything yet, start at is ChatGPT Plus worth it.
Every current plan and price I track lives on the AI plan tracker.
This post is part of Claude at Work, the hub for using AI at your job without code.
Published September 5, 2026. Prices, models, context windows and billing rules checked that day against chatgpt.com/pricing and OpenAI's Pro tiers help article, both linked inline. OpenAI changes these often and renders prices in the browser, so open the live page before you buy.
Quick answer: the two ChatGPT Pro tiers stopped being proportional. $200 is twice the money and four times the frontier-model messages, plus a fallback ladder the $100 tier does not have.
$100 gets you 50 GPT-6 Pro messages a week. $200 gets you 200.
I dropped from $200 to $100. I do not regret it, and I will show you the math that decides whether you should.
The allowance table
Straight from OpenAI's help center, checked September 4, 2026.Pro $100
Pro $200GPT-6 Pro in chat
50 messages / week
200 messages / weekGPT-5.6 Sol Pro
Shares the same 50/week
Separate 170 / dayCombined daily cap
n/a, one weekly pool
200 / day across bothAt the cap
Shared pool is gone; switching models adds nothing
Auto-switches to GPT-5.6 Thinking at Medium, and Sol Pro stays selectable while its daily allowance holdsModels available
Identical
IdenticalAnnual billing
No
NoFor context on the rung below: Business Standard gets 15 GPT-6 Pro messages a month, and Business Premium gets 50 a week. If you were assuming a work plan covers you, check that number.
The mechanic that actually bites
Everyone compares 50 against 200 and stops. The part that actually changes how the $100 tier feels is the word shared.
On $200, GPT-6 Pro and GPT-5.6 Sol Pro have their own allowances. Burn through your weekly GPT-6 Pro messages and OpenAI says ChatGPT automatically switches you to GPT-5.6 Thinking at Medium, and Sol Pro stays selectable while its daily allowance holds.
On $100, both Pro models drink from one 50-message weekly pool.
Switching between them does not buy you anything. There is no second Pro bucket.
That is the real gap between the tiers, and it is worth more than the raw 4x. You are not just buying more messages at the top, you are buying a second Pro allowance to fall back on.
To be precise about what OpenAI does and does not publish: the automatic switch to GPT-5.6 Thinking at Medium is documented for Pro $200. For $100 the help page says only that switching between the two Pro models does not add messages once the shared allowance is exhausted. Separately, it says that on any plan, hitting a GPT-5.6 reasoning limit means ChatGPT "may continue with another available reasoning model." So $100 is not a dead end, it just has no second Pro tier to drop to.
One more line from the same page that people learn the hard way: OpenAI Support does not reset ChatGPT or Codex usage limits. If you are out, you wait or you buy credits. Astra usage is included in your existing subscription allowance, and extra usage is purchasable on top.
Why I dropped to $100
Not because $200 was bad. Because my stack changed.My ChatGPT app, September 5, 2026: Pro plan at $100 a month, and 70% of the week's general usage already gone with two days to the reset. This meter is the whole decision.
Worth being straight about that screenshot: it shows Pro, not which Pro tier. ChatGPT does not surface the tier there. The $100 part is my word.
I now run three AI subscriptions instead of two, and I wanted budget for the third. Claude is my main engine, and honestly about 70% of the heavy work in my pipeline runs there. ChatGPT is where I do research, automations, and QA, and it is a very good sparring partner when Claude and I are building something.
Moving $100 from one subscription to a third tool bought me more than the extra ChatGPT capacity would have.
Since the downgrade I have hit the $100 wall a few times. On $200 I never did. That is the honest trade, and it was still worth it for how I work.
Now that GPT-6 Pro is 50 a week on my tier, do I regret it?
No. 50 focused Pro messages a week is more than I use, because I am not doing frontier-model chat all day. And Astra turned out to be far lighter on the plan than I expected. I ran real tasks on it and watched it move my usage by a couple of percent.
Worth saying plainly: my sense that OpenAI has been unusually generous with limit resets lately is my observation, not an OpenAI commitment. Do not budget around it.
The rule for who should be on $200
I get asked this constantly and my answer is the least satisfying one possible: let the pain tell you.
Do not upgrade because a launch made the top tier sound necessary. I bought the $200 plan once on exactly that logic and it was capacity I did not use.
The upgrade is right when all three of these are true:You are hitting the ceiling repeatedly, not once during a launch week.
There is no reset coming that solves it.
The waiting is actually costing you working time, and you can feel it.That third one is the real test. Everyone hits a limit occasionally. The signal is when you find yourself sitting there unable to continue and unwilling to wait.
That is exactly how I climbed the Claude ladder, by the way. Every step up happened because I kept hitting the limit and did not want to wait for the reset. The pain made each decision, not a spec sheet.
Cheaper moves to try before you upgrade
Three of these solve the problem for most people at a fraction of $100 a month.
Move agent work off chat. Work and Codex have usage rules separate from chat. If the thing draining your weekly Pro messages is long agentic work, running it in Codex does not touch your chat allowance at all. This is the single most underrated fix.
Buy credits instead of a tier. Astra usage is included in your allowance, and OpenAI sells credits for additional usage. A one-off overflow is cheaper than a permanent $100 a month.
Spend the $100 on a second provider. Two subscriptions on two stacks means two entirely separate allowances and two toolsets. Splitting my own work across two providers is what stopped me hitting walls, and the $20-tier version of that trade is priced out in Claude Pro vs ChatGPT Plus.
Upgrade for one month, then drop back. Neither tier has annual billing, per OpenAI's Pro tiers help page, which cuts both ways. Go up for a heavy build month, come back down. Nothing is lost.
Verdict
$100 is the right tier for almost everyone who is not living inside frontier-model chat all day. You get the same models the $200 tier gets, at half the price, with a weekly ceiling most people never touch.
$200 buys capacity and a safety net: four times the GPT-6 Pro messages, a separate Sol Pro allowance, and an automatic fallback so your work does not stop dead.
Pick $200 when the wall is costing you hours. Pick $100 until it does.
Comparing across providers instead of within ChatGPT? Claude Max vs ChatGPT Pro is the same question one level up, and which plans get GPT-6 Astra covers the surface gate. If $20 is the real budget, start at Is ChatGPT Plus worth it. More systems at Claude at Work.
Quick answer: these two models now cost exactly the same and are good at different things. GPT-6 Astra wins computer use, math, cybersecurity and long context. Claude Fable 5.1 wins the two broadest "how smart is it" boards, including one published by OpenAI itself.
Both are $10 per million input tokens and $50 per million output. The sticker price stopped being the decision.
Every headline yesterday said Astra swept. OpenAI's own launch table does not say that.
The rows nobody quoted
OpenAI published a comparison table on the Astra launch page with Claude Fable 5.1 in a column. Two lines in that table go to Claude.
Every number below is transcribed from that page, checked September 4, 2026. Where OpenAI's footnotes qualify a Claude score, that is flagged in the next section.Benchmark (OpenAI's own table)
GPT-6 Astra
Claude Fable 5.1Artificial Analysis Intelligence Index v4.1.1
61.2
65.7Humanity's Last Exam (with tools)
57.2%
65.0%Terminal-Bench 4.0
57.9%
55.8%Terminal-Bench Science 0.1
64.6%
52.6%FrontierMath Tier 4 (v2)
97.6%
87.8%GPQA Diamond
96.0%
93.7%AutomationBench
41.4%
31.4%DeepSWE v1.1
74.1%
67.4%FrontierCode 1.1 Main
53.3%
50.9%ARC-AGI-2
95.0%
90.0%Computer use safety, internal (lower is better)
2.4%
9.5%The Artificial Analysis Intelligence Index is the closest thing the industry has to a single general-capability number. OpenAI put it in its own launch table, and it hands Claude a 4.5 point lead.
Artificial Analysis, quoted independently:Sits beside GPT-5.6 Sol in Intelligence: GPT-6 Astra scores equal to GPT-5.6 Sol in the Index at 61. This is 5 points lower than Claude Fable 5.1 (max with fallback).That is not a small footnote. On the broadest board, the newest OpenAI model landed level with the previous OpenAI model.
Read the footnotes before you read the table
OpenAI ran these numbers. That does not make them wrong, but the footnotes change how you should read four rows.BenchCAD: OpenAI's footnote 5 says Claude's scores reflect three modifications to the eval.
ScreenSpot-Pro and ExploitGym: footnote 17 says the Fable scores reported are actually from Mythos, described as "Fable with fewer safeguards." That is a different model.
HealthBench Professional: footnote 11 says OpenAI independently evaluated the Claude models itself, using GPT-5.4 as grader, with Opus 5 substituted when Fable 5.1 refused.
Three science evals: footnote 12 says Claude Fable 5 and 5.1 are excluded from LifeSciBench, GeneBench Pro and MedChemBench because they refuse the majority of questions.None of that is scandalous. Vendors benchmark their own launches. But "state of the art across the board" is a press summary, not what the table says.
The 99.9% asterisk
The number that traveled furthest was Astra saturating ARC-AGI-3 at 99.9%. It deserves the least weight of anything in the launch.Two problems, both disclosed by the people who ran it:Harness. Per the ARC Prize blog, the 99.9% came from OpenAI's custom "Provider Adapter harness" at about $19K. The default ARC-AGI harness scored 62.7% at about $26K. OpenAI's own footnote 1 confirms it ran a modified responses API harness. That is a 37 point spread depending on plumbing.
No opponent. Claude Fable has no published ARC-AGI-3 result. The headline comparison of the launch is not a comparison.The Provider Adapter, in ARC Prize's words, "preserves opaque reasoning state between requests and uses compaction for longer conversations, allowing the model to reuse prior work." That is a genuinely interesting engineering result. It is not a raw intelligence score.
Simon Willison, writing the day it shipped, put the whole launch in one honest sentence: Astra "appears to score higher than Fable on most of OpenAI's self-reported benchmarks."
Self-reported is doing a lot of work in that sentence.
Where Astra genuinely pulls ahead
Strip the noise and Astra's wins are real, and they cluster.
Computer use. This is the headline capability, not the benchmark. Astra scores 59.3% on Agents' Last Exam against 55.5% for Claude Opus 5, and OpenAI reports it hitting 72.6% on OSWorld 2.0 in roughly 47% less time per task than GPT-5.6 Sol.
Math and science. FrontierMath Tier 4 at 97.6% against 87.8%. Terminal-Bench Science at 64.6% against 52.6%. These are not close.
Cybersecurity. 100% on ExploitBench, 88.0% single-attempt on SRE-Bench binary reverse engineering. OpenAI says this crosses the Critical threshold in its own Preparedness Framework, and it has restricted the model accordingly at launch.
Long context. 100% on OpenAI's eight-needle test at 256K to 512K tokens, and 96.3% at 512K to 1M.
Token efficiency. This one matters more than the scores. On Agents' Last Exam, OpenAI reports Astra using about 65% fewer output tokens than Opus 5 at the highest-scoring settings.
Where Claude Fable 5.1 holds
The broad boards. Intelligence Index and Humanity's Last Exam, both from OpenAI's own table.
Coding, closer than the headline. Astra edges Fable 5.1 on FrontierCode 1.1 Main, 53.3% to 50.9%. But look one column over in OpenAI's table: Fable 5 scores 53.5%, ahead of Astra. On the Artificial Analysis Coding Agent Index, Astra posts 67.0 against 67.2 for Fable 5 and 68.1 for Opus 5. Coding is a wash.
Cost per task at equal quality. Artificial Analysis again: "Per task, the model is less than half the cost of Claude Fable 5, for the same score." That is the one place Astra's efficiency turns into a real advantage, and it is worth more than the trophy rows.
What I actually threw at it on day one
I pay for both stacks, and Astra is one day old as I write this.I did not run a benchmark. I ran my actual work, which is the only test I trust.
The one that changed my mind about Astra was a prospecting task. I asked it to find companies in a specific niche worth reaching out to, and it went and searched forums and threads on its own without me telling it where to look. What came back was roughly a month old, where that kind of query usually surfaces threads from many months back.
That is the real upgrade. Not the score, the autonomy.Another Astra run from the same week, on my Pro account. I asked it to plan two weeks of paid-work experiments around my real numbers; it pulled live listings on its own and priced the options. The top of the answer has my personal figures in it, so it is trimmed. The chip in the composer is the part worth noticing: on this surface it is labeled GPT-6 Astra Ultra, not GPT-6 Pro.
Then I did the thing I would actually recommend: I ran the output back through Claude Fable to QA it, Fable rewrote the prompt, and I sent it back at Astra. The results from that second pass were promising enough that it is now how I would run this by default.
The honest other side: I have used Fable more, and Fable still feels more familiar and more dependable to me for building. Fable 5.1 has also done some genuinely dumb things this week, mostly losing track of tools it definitely had access to. Minor, but real.
My read on the benchmark inversion: I assumed the scoreboard said Astra beat Fable outright. It does not, and that matches my hands-on impression rather than contradicting it. Fable still feels better at the deep work. Astra feels newer at getting things done by itself.
The decision rule I would give someone today
Do not switch providers over this. Route work instead.The job
OpenDeep building, writing in your voice, long careful work
Claude Fable 5.1Anything the AI should do on its own across apps and the web
GPT-6 AstraResearch and volume where you need lots of runs
GPT-6 AstraReviewing and QAing another model's output
Claude Fable 5.1Math, science, anything with a checkable right answer
GPT-6 AstraYou can only pay for one, and you are a generalist
ChatGPT (usage and resets decide it, not the score)The thing I keep coming back to: smartest is no longer the flex. A model that burns tokens to win a board is worth less to me than a leaner one that finishes the task. That is the direction both companies are being pushed, and this launch is the clearest evidence of it so far.
Verdict
Same price, split by job, and the split is the useful output. Astra is the better agent. Fable is still the better thinker on the broadest measures OpenAI itself published, and it is my daily driver for building.
If you were about to cancel Claude because of a headline, read the table first.
Working out which subscription this actually affects? Which ChatGPT plans get Astra, Claude Max vs ChatGPT Pro, and Claude vs ChatGPT for the everyday call. More of how I run both at Claude at Work.
Published September 4, 2026, one day after GPT-6 Astra shipped. Every benchmark figure was transcribed that day from OpenAI's launch page, and the ARC-AGI-3 harness figures from Simon Willison's write-up citing the ARC Prize blog, both linked above. These are vendor-run numbers on a vendor page and OpenAI's footnotes qualify several Claude scores; read the footnotes section before quoting any row. Benchmarks and prices move fast and this page will need re-checking.
Quick answer: if you pay $20 for ChatGPT Plus, you do get GPT-6 Astra. You just do not get it in the chat box.
OpenAI's help center puts Astra on Plus inside ChatGPT Work and Codex. In the normal chat model picker it is branded GPT-6 Pro, and that is listed for Pro $100, Pro $200, Business and Enterprise.
So the thing half the internet is saying today, that Plus got locked out of GPT-6, is wrong. And the thing OpenAI's own announcement implies, that everyone has it, is not true yet either.
Availability is per surface, not per plan. That is the whole story, and it is the table below.
The plan by surface table
This is the part every write-up skipped. Same plan, different answer depending on where you open it.Plan
Chat picker
ChatGPT Work
CodexFree
No
No
Terra onlyGo
No
No
Terra onlyPlus
No
Yes (rolled out Sept 4)
Yes (rolled out Sept 4)Pro $100
Yes, as "GPT-6 Pro"
Yes
YesPro $200
Yes, as "GPT-6 Pro"
Yes
YesBusiness
Yes
Yes
YesEnterprise
Yes, off by default
Yes
YesSources: OpenAI's GPT-6 Astra announcement (September 3, 2026) and the help center article GPT-5.6 and GPT-6 Pro in ChatGPT, checked September 4, 2026.
One more source, because it is fresher than the help page: on the evening of September 4, OpenAI's Codex and ChatGPT lead Thibault Sottiaux posted on X that "Astra is now rolled out to all Plus and Business users too." That post does not name a surface. Read together with the help page, it means the Work and Codex rollout on Plus is done, not pending. If the chat picker changes for Plus, it will show up on the help page first and this table gets updated the same day.
Two more details from the same help page that decide real cases:Rollout is gradual, and OpenAI says availability can differ between Chat, Work and Codex on the same account.
Usage and credit rules in Work and Codex are separate from chat. Running Astra in Codex does not eat your chat allowance.The picker slot that decides it
Here is the mechanical reason a Plus account cannot show you GPT-6 Pro.
ChatGPT's picker now has reasoning levels: Instant, Medium, High, Extra High, and Pro. GPT-6 Pro lives under the Pro option, next to GPT-5.6 Sol Pro.
Plus does not have the Pro option at all.That table is about GPT-5.6 Sol reasoning levels, but it is the same door. No Pro slot means no GPT-6 Pro, no matter how long you wait. This is a plan gate, not a queue.
Why everyone got this wrong
Because OpenAI published two sentences that read differently, on two different pages, one day apart.
The announcement, September 3:GPT-6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users.The help center, updated September 4 (and OpenAI staff on X that evening confirmed the Plus rollout was complete):GPT-6 Pro, powered by GPT-6 Astra, is rolling out in ChatGPT for Pro $100, Pro $200, Business and Enterprise plans. Plus plans include GPT-6 Astra in ChatGPT Work and Codex as it rolls out.Both are true. The announcement is talking about Astra the model reaching all paid plans. The help center is talking about which surface it reaches on each one.
Every news write-up quoted the first sentence and stopped.
The gap between those two OpenAI pages is exactly why so many people spent today thinking their account was broken.
Independent write-ups landed in the same place. Simon Willison, covering it the day it shipped, quoted the announcement line verbatim and added: "I've not tried it yet myself, so I don't have a great deal to say about it yet." Nobody had the surface split on day one.
It is not called Astra in ChatGPT
If you are scrolling the Chat model picker hunting for the word "Astra," you will not find it there.
In the Chat picker it is GPT-6 Pro, per OpenAI's help center. In the API it is gpt-6-astra. And on my own Pro account the Work composer chip reads GPT-6 Astra Ultra, a third label for the same model. One model, three names depending on where you are standing, and the naming split is doing real damage to people trying to answer a simple question about their own account.
While you are in there: GPT-6 Pro and GPT-5.6 Sol Pro are two different models that share an allowance on some tiers. They are not the same thing, and on the $100 tier they draw from one pool. That is its own decision, and its own page.
What I actually see, from the tier that has it
I pay for ChatGPT Pro at $100. I dropped down from the $200 tier to free up budget for another tool, and I have run Astra on the $100 plan on a few real tasks since it landed.That is my Usage & billing screen in the desktop app, September 5: Pro at $100 a month, with 30% of the weekly limit left.And that is the Chat picker on the same account. The model list under the effort slider reads Latest, GPT-5.6 Sol, GPT-5.5. Nothing in this Chat menu says Astra, which is the point of the section above: in Chat you find a Pro level, not the word. Switch to the Work surface on the same account and the chip reads GPT-6 Astra Ultra. Same model, different label.
The thing that surprised me: it barely moved my usage. I expected a frontier model to drain the plan and it used a couple of percent across those tasks. My impression is that Astra is a bit more token-efficient than what I was running before, though that is a feel from a handful of runs and the published benchmarks are the place to check it.
I have also seen OpenAI reset limits generously lately. That is my observation, not an OpenAI commitment, so do not budget around it.
One day of use on one account is not a benchmark. But it changes the upgrade math below.
Should you jump from Plus to Pro just to get it in chat?
No. Not for that reason alone.
Here is my rule, and it has not changed for any model launch: let the pain tell you. Do not upgrade because a launch made you feel behind. Upgrade when you are actually hitting a wall that costs you time.
Three specific reasons the jump is usually wrong right now:You already have Astra. Plus includes it in Work and Codex. If your work is research, document work or anything agentic, that is where you want it anyway.
The chat allowance is small. Pro $100 is 50 GPT-6 Pro messages a week, shared with GPT-5.6 Sol Pro. That is a real ceiling, not an unlimited upgrade.
$80 a month is a lot to pay for a picker entry. If you want a second frontier model, $20 on a second provider buys a whole separate allowance instead of one row in a menu. That trade is priced out in Claude Pro vs ChatGPT Plus.The honest exception: if you are already hitting Plus limits every week and waiting on resets, you were going to upgrade anyway. Astra just made the upgrade more interesting.
Before you conclude you are missing out
Run this list first. Most "I don't have GPT-6" posts are one of these.Check the surface. Open ChatGPT Work or Codex, not the chat box. On Plus that is where Astra lives.
Update Codex. OpenAI says Astra needs Codex CLI 0.153.0 or newer. An older CLI hides it completely.
Update the desktop app. Same help page asks for the latest ChatGPT Desktop build.
On Enterprise, ask your admin. OpenAI says Enterprise access is off by default at launch and depends on your workspace's model-access permissions.
On Pro, give it a day. Rollout is gradual and can differ per surface on the same account.If you have done all five on a Pro plan and still see nothing, you are in the rollout queue, not excluded.
The one-paragraph version
Plus gets GPT-6 Astra in ChatGPT Work and Codex, not in the chat picker. In chat the model is called GPT-6 Pro and it is gated to Pro $100, Pro $200, Business and Enterprise, with Enterprise off by default. Free and Go get GPT-5.6 Luna and nothing newer. Nobody is broken, availability is just per surface, and no plan change is worth making on launch-day panic.
Deciding between the top tiers instead? That is Claude Max vs ChatGPT Pro. Everything I run and pay for lives on Claude at Work.
Quick answer: you connect them through Zapier MCP, and it is click-through, not code. Create a server at mcp.zapier.com, add only the app actions you want Codex to have, connect it once, and Codex can now act in your other apps: add sheet rows, create Notion pages, send a Slack ping. Ten minutes, one paste.
But know the direction before you start: Codex can reach into Zapier. Zapier cannot start Codex. That single fact sorts out most of the confusion around this search.
I use Codex a lot, including in my demos. This page is the setup written for a normal person, because the only real how-to out there was written for engineers shipping pull requests.
Codex Can Act In Your Apps; Zapier Cannot Start CodexYou want
Do thisCodex to act in your other apps (Sheets, Notion, Slack, Gmail)
Zapier MCP, steps belowThe fastest route inside Codex
The official Zapier plugin, two commandsA Zap to kick off a Codex run
Does not exist as of Aug 28, 2026; Zapier's app directory returns 404Automations that run on their own schedule without you
Different tool conversation: Codex vs n8nTo build with the Zapier SDK in code
You are not the reader this page is for, and that is fineFirst, the direction, because nobody says it
Every article assumes the connection is symmetrical. It is not.
Codex reaching into Zapier: works, and it is the whole point. Through MCP, Codex gets hands in 9,000+ apps during a session you are running.
Zapier starting Codex: not a thing. Checked August 28, 2026: Zapier's Codex app pages return 404, and that can change. There is no trigger called "run Codex."
The only workaround is hosting your own webhook receiver, which is real engineering. If your dream was "every morning at 8, Zapier wakes Codex up," that dream is currently a workflow-software job, not a Zapier one.
That is the Ship Lean rule for this pair: Codex reaches out, Zapier never reaches in.
The Whole Setup Is Five Clicks In The Zapier Dashboard
This is Zapier's own documented flow, translated from engineer to human:Go to mcp.zapier.com and sign in with your Zapier account (free to create).
Click "+New MCP Server" and choose Other as the client.
Click "+Add tool" and pick your first action: choose the app (say, Google Sheets), then the specific action (say, "create spreadsheet row"), then connect that account.
Repeat for each action you actually want. This is the safety step hiding in plain sight: Codex can only ever use the actions you added. If you never add "send email," it cannot send email. Scope tight; add more later.
Click "Connect" at the top of the dashboard and follow the instructions to add the server to your Codex account. On the desktop app this is a single paste.One nice mechanical fact from OpenAI's MCP docs: the ChatGPT desktop app, Codex CLI, and IDE extension share the same MCP configuration. Set it up once, it carries everywhere. ChatGPT on the web is the exception; there it rides plugins instead.
The shortcut, if a terminal does not scare you: Zapier ships an official Codex plugin. Two commands, codex plugin marketplace add zapier/marketplace then codex plugin add zapier@zapier, and you are done. If typing commands is a wall, skip it; the dashboard route above needs none.Zapier's MCP-for-Codex guide, captured August 14, 2026. Their examples are all GitHub and Jira; the steps work the same for Sheets and Notion.
What I would wire first, and it is not what Zapier suggests
Zapier's guide demos GitHub issues, Jira tickets, and deploy logs. Useful if you are an engineer; useless for most people. The pattern that would actually earn this setup, for work shaped like mine, is Codex-does-the-thinking, app-gets-the-result:Draft lands where you work: Codex writes or edits something in a session, then files it into Notion as a page instead of leaving it in the chat.
Numbers land in the sheet: you have Codex pull together your weekly stats, and the MCP action appends the row to the spreadsheet you already track.
The ping when it is done: long task finishes, one Slack or email action tells you.Notice the shape: the intelligence stays in Codex, and Zapier is just hands. The moment you catch yourself designing multi-step Zap logic around Codex, stop; you are rebuilding workflow software the hard way.
When you should skip Zapier entirely
Honest counter-position, because our house view has not changed: a capable agent usually needs fewer connectors than you think. Codex already reads and writes files, browses, and runs work in place.
If your "integration" is really "get the answer into a doc," ask Codex to produce the file and skip the plumbing. Add Zapier MCP when a result must land inside a specific app you live in, not as a reflex.
And if what you actually want is the always-on scheduled kind of automation, the decision tree for that lives in Codex vs Zapier and Codex vs n8n. This page is the how; those are the whether.
If you are: already living in Notion or Sheets and want Codex results to land there, do the MCP setup above. Comfortable in a terminal, use the plugin. Wanting something to run at 8am without you, skip this page and read Codex vs n8n.
How I know what is on this page
Setup steps: Zapier's own MCP guide, the official Zapier plugin repo, and OpenAI's MCP docs, all read August 28, 2026, quoted at the claim. The missing-connector fact: checked live the same day against Zapier's app directory, and it can change, so re-check if you are reading this months later. The judgment calls, what to wire and when to skip it, are mine as a regular Codex user, labeled as such. I have not wired this exact Zapier MCP setup myself; the setup steps come straight from the docs above. Where Zapier publishes no limits or pricing for MCP, this page says so instead of guessing.
Quick answer: stop asking which one wins. They sit in different rooms of your life. Grok Bot is a teammate with its own computer in xAI's cloud that you text from your phone: monitoring, scraping, routines, the always-on stuff. Claude Cowork is the worker at your actual desk, inside your real files and projects, where careful work happens. I pay for both, $100 SuperGrok Plus and a paid Claude plan, and I use both daily. Here is the split as I actually live it.
The Ship Lean split, if you want it in one line: Grok Bot carries your pocket. Cowork carries your desk.
Published August 28, 2026. Last reviewed August 2026 against xAI's and Anthropic's live pages.
Quick decisionYour situation
PickYou want an assistant you text from your phone that works while you sleep
Grok BotYour work lives in documents, folders, and real projects on your machine
Claude CoworkYou bounced off agent tools because setup was too technical
Grok BotYou already pay for Claude Pro or Max
Cowork first; it is already included$30 is your whole experiment budget
Grok Bot on SuperGrokYou are actually choosing between the chatbots
Different page: Grok vs ClaudeGrok Bot is a cloud teammate with its own computer and its own usage pool
Launched August 11, 2026, early beta. Per the launch post: "your team of always-on agents. They have their own computer, work inside tools and apps like you do, and keep working 24/7." The computer is a real machine in xAI's cloud, so jobs do not stall when you step away. You teach a bot by doing a workflow once while it watches; it saves it as a routine and reruns it on its own.
Two facts every comparison page misses, both official:Entry is $30, not $200. xAI's pricing page lists "Grok Bot access" on the SuperGrok tier as of August 28. The enterprise-paywall story ranking around the web is stale.
Its usage is separate. The launch post states bot work "won't count against your existing usage." That matched my week: bots running, chat usage untouched. Details: Grok usage limits explained.Access rides SuperGrok or Cursor plans; the closest OpenAI equivalent is ChatGPT Work, covered in Cowork vs ChatGPT Work.
Claude Cowork is the worker inside your own files, and it now has its own browser
Per Claude's Cowork page: give it a goal and "it works across your files and tools," then delivers work for review. It ships on every paid Claude plan from $20 Pro up, on macOS and Windows desktop, web, and iOS and Android.
And the objection everyone still repeats is dead: Cowork now has a browser built into it, separate from your own logins and tabs, and scheduled tasks run unattended with your laptop closed. If your mental model of Cowork is six months old, it is wrong. Where it sits next to Claude's coding tool is its own question: Claude Cowork vs Claude Code.
Where Grok Bot wins: the pocket and the clock
My real setup after week one: a bot called Partner manages a small team, with each bot on its own computer with its own skills. They read my channel analytics and come back with specific calls. Another bot scrapes AI news for me three times a day. I built all of it in plain English, and I check on it from my phone. (Full week-one story with the roster screenshot: What is Grok Bot.)
This is the OpenClaw dream with the pain removed. OpenClaw was the first "AI on the go" moment and I loved it, but it was buggy and daunting for most people. Grok Bot is that idea with an actual dedicated computer, a stronger model, a genuinely good mobile app that syncs with desktop, and a dedicated usage pool. For the people who could not survive OpenClaw setup, this is the answer, hands down.The pocket side of the split: my real bot team in the Grok Bot macOS app, captured August 28.
Where Cowork wins: the desk
Here is a real Cowork session from my week, not a hypothetical:Cowork watched a 48-second screen recording of a workflow I do by hand, then drafted a reusable version of it and put it up for my approval. It also caught a conflict between two of my own saved preferences and told me which one it followed.
Read what is happening in that screenshot, because no feature grid captures it. Cowork watched me work for 48 seconds, understood the outcome rather than the clicks, proposed a permanent automation, and flagged a contradiction in my own preferences. That is judgment inside my actual system, with my actual files.
That is the lane where, in my week, Grok Bot lost to Claude. When the task was deep, multi-step, and touched things I care about, I did not even consider handing it to a bot in someone else's cloud. Claude is stingier with usage, and that is a real cost I feel, but one 48-second recording turning into a reusable skill is work no phone bot did for me all week.
New to running Claude as a work tool? Start here: Claude at work.
The honest costs on both sides
Grok Bot's honest costs: it is early beta and acts like it. My lead-scout experiment produced nothing in week one and I may kill it. And its whole model means your logins live on a cloud computer that acts as you; think before you hand it accounts that matter.
Cowork's honest costs: Claude usage. Anthropic is the stingy one of the two, and even the current 50% usage boost is an extension, not a promise. Meanwhile my Grok limits have reset weekly, sometimes more than once, and bot work does not touch the chat pool. If your bet is "who gives me more machine per dollar," xAI is playing that game harder. If it is "who does the most careful work," that is Claude.The receipt for the usage claim: my SuperGrok Plus billing screen, August 28, "Reset Available" badge showing.
The comparison creators are circling the same fight: Simon Scrapes titled his review "Did Grok Bot Just Overtake Claude?" and Riley Brown ran a three-way super-app test. Watch both; notice that the disagreement is always about which room of your life the tool sits in.
Claude Cowork vs Grok Bot: which one to start withYou already pay for Claude: open Cowork today; it is in your plan. Start with one folder and one recurring task.
You do not pay for anything yet and want the assistant feeling: SuperGrok at $30. It is the cheapest real agent on the market right now.
You are me, running a one-person operation around a full job: both, split by room. Pocket work to Grok Bot, desk work to Cowork, and neither subscription resents the other.How I know what is on this page
Product facts: xAI's launch post and pricing page, and Claude's official Cowork pages, all read August 28, 2026, dated on my AI plan tracker. The verdicts: my own two paid accounts and my own week, including the Cowork session screenshot above and the bot team on the Grok side. Where a claim is mine and not official, like how my own bots behave day to day, I have said so.
Quick answer: Grok Bot is xAI's AI teammate product, launched August 11, 2026, still labeled early beta. Each bot gets its own computer in the cloud, signs into your tools, and keeps working after you close your laptop. It starts at $30/month, and its work does not count against your Grok chat usage. I pay for it, I built a small team on it in week one, and it is the first agent product I would hand to a non-technical person, beta caveats included.
The Ship Lean read, in one line: Grok Bot rents your assistant a computer, Claude works in your files, and the $30 tier is the whole story.
One naming correction before anything else, because it burns people in search: this is not the @grok bot that replies to tweets. Same company, different product. This one you hire.
Grok Bot Is an Assistant With Its Own Cloud Computer, From $30Question
AnswerWhat is it
AI teammates with their own cloud computers, from xAILaunched
August 11, 2026, early beta, per the launch postCheapest way in
SuperGrok at $30/month lists "Grok Bot access" on xAI's pricing page, checked August 28Where it runs
Desktop app (macOS .dmg) plus iOS; it syncs between themUsage
Separate pool; bot work does not drain your Grok chatWho it is for
People who want an assistant that works while they sleep, without setup pain"Its own computer" is the whole product
Every explainer repeats the phrase and never translates it. Here is the translation.
Your bot lives on a machine in xAI's cloud. It stays awake when you step away, signs into websites and apps there, and works "including platforms with no clean API or MCP" in xAI's own words. You teach it by doing a workflow once while it follows along; it saves the routine and runs it on its own next time.
Bots can even message each other and coordinate.
That is also the honest part nobody says out loud: your logins end up on a cloud computer that acts as you. HN user wiradikusuma asked the exact right question: "How does it work with login-walled sites like LinkedIn then? And what does 'own computer' mean?" The answer is: it works because the computer is real, and you should think before handing it accounts you care about.
I Built Three Bots in Week One, and One Flopped
I am on the $100 SuperGrok Plus plan, and I spent the week attacking my own gaps. Three builds:My actual bot roster in the Grok Bot macOS app, captured August 28. Partner runs the team. Each bot has its own skills and access to my data, and per xAI's docs they all share one persistent cloud computer scoped to my account, which is why a single login covers the whole roster.
A mini marketing team that watches my analytics. My gap is that I make content constantly but only review the numbers for one channel. So I built Partner, a co-founder bot that manages a small team: Friday, Tweet, Steal, and Coach. They read my actual channel data and come back with specific calls: this dropped, steal this video, this flopped because of that.
A mini team hawking my analytics, telling me what to improve every week.
A news scraper. Three times a day it collects the latest AI drops for me. It already cleaned up some of my lists. Simple, boring, exactly what an assistant should do.
A lead scout, which flopped. I pointed a bot at my services page and told it to find my ideal clients and start conversations. First week was a dud. I might drop it. I am telling you this because it is an experiment, not a system, and any page that only shows you wins is selling something.
The thing that actually surprised me: the usage math
I hit my usage limit on the $100 plan this week. Then it reset. In my experience the resets land weekly, sometimes more than once, though xAI does not publish that schedule anywhere. Meanwhile the launch post says Grok Bot "comes with its own usage, separate from your Grok and Cursor plans."
Two pools, one price. The careful version of that claim, with the official quotes and where the docs go quiet, is in Grok usage limits explained.The other half of the subscription: Grok chat running a live "what's trending" search on my desktop, captured August 28. Chat and Bot are separate products on separate meters.
This is the OpenClaw idea, shipped properly
If you remember OpenClaw, that was the first "AI on the go" moment: amazing when it worked, buggy and daunting to set up. Grok Bot is that idea as version 2.0: an actual computer instead of your session, a stronger model (Grok 4.6), a real mobile app that syncs with the desktop, and a dedicated subscription instead of duct tape.
I am not alone in that read. HN user thenbrent put it as: "Right now Grok Bot looks a lot easier to get started and maintain with a simpler UI (arguably better), but OpenClaw and Hermes give you more configurability and choice." That is the honest trade: easier and more polished, less control.
It Starts at $30, Not the $200 Reddit Says
The launch had real pricing confusion, and stale answers are still ranking. Reddit threads claim it needs $120 per seat or a $200 to $300 plan. As of August 28, xAI's pricing page lists "Grok Bot access" on the $30 SuperGrok tier, with SuperGrok Plus at $100 for much higher usage. During the beta, the Bot docs and the pricing page have not always agreed on tiers, and HN user gexla called that out directly: "This is super confusing."
So the practical rule: read the pricing page the day you buy, not a blog post from launch week, mine included. I keep the receipts dated on my AI plan tracker, and my own billing screen is on the worth-it breakdown:My actual billing screen, August 28: SuperGrok Plus, $100, paid August 21. Note there is no published price for Lite or Heavy anywhere on it.
Who should get it, who should not
Get it if you want a personal assistant, not a coding partner. The pitch that lands with normal people is: for $30 budgeted deliberately, you get an assistant with its own computer, on your phone. Tailored news pulls, monitoring, reminders, the boring recurring stuff. For people who bounced off agent tools because setup was too technical, this is the answer, hands down.
Skip it if your work lives in your files. My deep work still happens in Claude, where the agent works inside my actual projects on my actual machine. That comparison gets its own page: Grok Bot vs Claude Cowork. If you are deciding between the model chatbots themselves, that is Grok vs Claude and Grok vs ChatGPT. And for where any of this fits into an employed person's week, start at Claude at work.
Wait if beta bugs scare you. It is labeled early beta, and it earns the label sometimes. My lead scout produced nothing in week one. The product moved a boundary; it did not become magic.
How I know what is on this page
Product and pricing facts: xAI's launch post, bot page, and pricing page, all read August 28, 2026, dated on the tracker. The builds and the usage story: my own paid account, with the bot roster screenshot above from my actual desktop app. The skepticism: Hacker News users quoted verbatim with links, because the confusion is part of the truth of a two-week-old product.
Quick answer: I never thought I would write this sentence, but here it is: for research, ChatGPT wins - even against Fable 5, which I consider the strongest model available. ChatGPT's Sol Ultra tier is phenomenal at digging into a question. Claude wins the moment research has to become something real.
The Ship Lean split: ChatGPT wins the research. Claude wins the build. The handoff is the skill.
My credentials for this one are just receipts: I pay $200/month for each. Both tools earn it, at different points of the same project.
Quick decisionResearch job
PickOpen-ended digging on a hard question
ChatGPT (Sol Ultra tier)Talking a problem through until it makes sense
ChatGPT, voiceReading your own live dashboards and analytics
ChatGPT's agent, in the browserTurning findings into a document or system
ClaudeResearch inside your own files and projects
ClaudeCited quick answers on current events
Honestly, free PerplexityA real session, start to finish
The clearest way to show the split is a session I actually ran, not a hypothetical.
I needed to figure out why my YouTube shorts were slipping and what to change. Here is how it went:OpenAI's own plan documentation, captured August 21, 2026. The Sol Ultra reasoning tier this post keeps praising lives at the top of this ladder.
Phase 1 - research, in ChatGPT, by voice. I opened voice and said, roughly, "look at my analytics." Its agent works in a real browser, so it logged into my YouTube dashboard - no API, no exports - and read the backend with me. Huge unlock, and most people have no idea it exists.
We walked the numbers together and saw where shorts were dropping. Then we brainstormed. This is the part that felt genuinely new: it was not question-answer, it was partner-mode. What about this? What about that? For 30 to 60 minutes we went hard at it, until I had actual clarity instead of a pile of facts.
Phase 2 - the plan. With clarity reached, I had it draft the plan properly. It looked at my existing setup and produced a proposal I understood and believed in. That last part matters: I did not want a plan handed down, I wanted the plan we had just reasoned our way to, written up.
Phase 3 - the build, in Claude. We started building in ChatGPT's coding side, and honestly, for this job it was not the right fit. So I took the finished plan to Claude and built it there. Execution is where Claude is strongest right now - it has the model for it and the working surface for it. Plan in, working system out.
That workflow - research and clarity in one tool, execution in the other - beats either tool alone, every time I run it.
(ChatGPT vs Claude for research, if you searched it that way - same split, same page.)
ChatGPT wins the research phase on voice, browser, and Sol UltraSol Ultra is phenomenal. OpenAI's top reasoning tier - the one that ships with the Pro plans - digs deeper on open questions than anything else I have used, my own beloved Fable included. It pains me slightly. It is true.
Voice makes research a conversation. Thinking out loud is how humans actually reason through problems. ChatGPT is the only tool where that experience is good enough to use daily - on walks, in the car, on the go.
The browser agent reads your real data. Logging into your own analytics and interrogating live numbers, no exports, no API keys, is a different category of research session.
Better interface for the meandering middle of research: branching, revisiting, riffing. It feels like a workspace, and OpenAI keeps investing in exactly this - the business-facing side of ChatGPT is growing fast.Claude wins the moment research becomes a buildExecution. When findings must become a document, a system, or code that runs, Fable 5 is the strongest builder I have used. The research phase produces a plan; Claude ships it.
Your own material. Research over your files, past work, and ongoing projects lives naturally in Claude's Projects and Cowork.
Instruction-holding on long outputs. The final artifact follows the spec - structure, voice, constraints - with less drift. In research-to-deliverable work, that is the difference between one pass and four.The 60-second version, from my Shorts: Stop Claude and ChatGPT Contradicting Each Other on Your Work.Concessions, both directions
ChatGPT's research win is not total: for cited quick lookups, free Perplexity is often the faster tool, and for source-grounded work over a fixed pile of documents, Google's NotebookLM is a sleeper pick. And Claude's build win is not total either: plenty of light execution - emails, summaries, quick drafts - is fine in ChatGPT, and switching tools for it would be ceremony.
No one model takes it all. That is not a hedge; it is the actual finding from paying top dollar for both, month after month.
How I compared these
Both tools: daily paid use at the $200 tiers, across real research-to-build projects like the session above. Plan and pricing facts on this page and the tracker are read from official pages, dated August 21, 2026. Verdicts are mine and labeled as experience, not benchmarks.
Three setups, three movesOne subscription, research-heavy work: ChatGPT - my full worth-it verdict. Add free Perplexity for citations.
One subscription, build-heavy work: Claude - sized via Pro vs Max.
Both, or planning to get both: steal the workflow above verbatim: voice-research in ChatGPT until clarity, plan on paper, build in Claude. The general comparison lives at Claude vs ChatGPT, and the top-tier money question at Claude Max vs ChatGPT Pro.Last reviewed August 2026.
Quick answer: Claude 5 is Anthropic's new model family, and Fable 5 is its first model - a new tier that sits above Opus in capability. In plain English: the ceiling moved. Tasks that used to almost work now just work. That sentence sounds like marketing until it happens inside your own workflow, so this post gives you mine, before and after.
The honest compression: Opus 4.8 needed a plumber. Fable 5 just did the job.
The plain-English version of what launched
Anthropic's Claude 5 family launched June 9, 2026, led by Fable 5 - with a bumpy start: access was suspended June 12 and redeployed July 1, all logged openly on Anthropic's own announcement. The parts that matter to a normal person:Anthropic's own announcement, captured August 21, 2026. Launch June 9, access suspended June 12, redeployed July 1 - the timeline is right on the page.Fable 5 is a new class of model above Opus - Anthropic calls the tier Mythos-class. Fable 5 and Claude Mythos 5 share the same underlying model; Fable is the version everyone can use, with additional safety measures, while Mythos goes only to approved organizations.
Opus 5 is still around and still excellent. The standard lineup did not get worse; the ceiling above it got higher.
The plan gotcha: since July 20, Fable 5 is not part of Claude Pro's normal usage - it bills as separate usage credits there. On Max plans it is included, capped at 50% of weekly limits, per Anthropic's pricing page. I track every plan detail like this, dated, on my AI plan tracker.My before-and-after, with receipts on YouTube
Here is the change measured in my actual work, not benchmarks.
Before, on Opus 4.8: I was trying to build a visual-aid system for my short videos - the thing that adds graphics, titles, and proof screenshots onto a talking-head clip automatically. Complex, multi-step, lots of judgment. On 4.8 I could get close. There was always some plumbing left, and what came out was a bit Frankenstein: parts bolted together, me holding the wrench.
After, on Fable 5: it just did it. I rebuilt the skill with Fable, and over the following weeks it got reliable enough that it now runs my shorts post-production. I record a talking head; the system does the visuals.My actual shorts page, captured August 21, 2026. The price overlays and proof cards on these videos are generated by the pipeline Fable 5 rebuilt - the thing Opus 4.8 could only almost do.
You do not have to take my word - the receipts are public. Go to my YouTube shorts and compare this week's videos against ones from four months ago. Night and day. Part of that is me - lighting, style, topics evolving - but the visual layer you see on recent shorts is the Fable-built pipeline doing work that the previous generation could not finish.
And it is not just me. My feeds during the launch window were full of people shipping fully polished videos from scratch, even avatar versions of themselves. Fable moved a boundary: complex creative-technical pipelines went from expert projects to things a determined non-coder can have built for them, in plain English.The 60-second version, from my Shorts: There are 4 different Claudes. Here's which one does what.The plan decides how much Fable you actually getPlan
Fable 5 access
Who it fitsFree
Not included
Everyday questions on standard modelsPro ($20/mo)
Pay-per-use usage credits
Full work surface, Fable on demandMax ($100/$200)
Included, up to 50% of weekly limits
AI doing real work every dayFree and everyday use: you do not need Fable. The standard models are genuinely strong, and most people would not feel the difference on normal tasks.
Pro at $20: you get the full working surface - Claude Code, Cowork, Projects - on the standard models, with Fable available as pay-per-use credits. My full verdict on that tier: Is Claude Pro worth it.
Max at $100/$200: Fable inside your plan up to half your weekly limit. If your work leans on AI all day, this is where the new ceiling is actually usable daily - sizing guide in Claude Pro vs Max.
A practical tip from daily use: Opus 5 on high effort is phenomenal and covers most real work. I reach for Fable when the task is genuinely hard, not by default. That habit stretches any plan a long way.My picker, September 2026. Fable 5.1 is one tap away under More models, but Opus 5 stays the default, and the Effort dial is what actually moves your meter.
The honest limits
No clean sweep here. For research workflows, I still find ChatGPT's Sol Ultra phenomenal - in my experience sometimes better than Fable at digging through a question (that full comparison is here). And the competition is genuinely hot: xAI is shipping faster than anyone and their agent products are moving weekly (my read on that).
No one model is going to take it all. Claude is the design-and-build freak, ChatGPT the research generalist with the better everyday app, Grok the fast-moving bet. The Claude 5 launch did not end that race; it raised Anthropic's ceiling in it.
How I know what is on this page
Model-family facts: Anthropic's own announcement, linked above. Plan and pricing facts: read off Anthropic's official pages August 21, 2026, dated on the tracker. The before-and-after: my own production pipeline, with the output public on my channel - linked so you can check the receipts yourself.
Three seats, three movesCurious non-coder: ignore the model-name noise. Get any paid Claude plan, use the standard models, and let one real chore be your benchmark.
Pro user hitting the ceiling on hard tasks: try Fable via usage credits on your hardest workflow before upgrading plans. If it finishes what Opus almost finishes, the Max math starts making sense.
Deciding between labs at the top: Claude Max vs ChatGPT Pro is the $200 decision, written from paying for both.More plain-English Claude guidance: Claude at Work.
Last reviewed August 2026. Plan facts verified against official pages on August 21, 2026.
Quick answer: hell yes. And if $20 a month feels expensive for what ChatGPT Plus does, I am going to be straight with you: you are not evaluating the tool, you are avoiding it. AI is here to stay. You will use it eventually - the only question is whether you start while it is an advantage or after it becomes a requirement.
If $20 feels expensive, you are pricing the tool. Price the hours.
Where I am coming from: I pay for ChatGPT at the top tier and for Claude Max at $200. I am not an OpenAI fan defending the home team - Claude runs my business. And ChatGPT Plus is still the first $20 I tell normal people to spend.
Quick decisionYou are
Do thisNew to paid AI
Plus at $20, todayLight chatting only
Go at $8, eyes open about ads and no Sol or AstraWork gives you Copilot, no personal use yet
Stay free, revisit in a quarterHitting Plus limits weekly
The Pro tiers, or the $200 lab decision belowPlus buys the reasoning model and the deep features
Verified against OpenAI's own plan article, August 21, 2026 (prices and plan facts also tracked, with dates, on my AI plan tracker):GPT-5.6 Sol, the reasoning model - the single biggest thing Free and Go do not have
Deep research, scheduled tasks, record mode, developer mode, expanded Codex
Custom GPTs, projects, memory, the best voice mode in consumer AI
54K context on instant chats, 256K on reasoning
$20/month, monthly billing only - OpenAI does not offer annual on PlusOpenAI's own Plus plan article, captured August 21, 2026. The pricing pages render prices in the browser now; the help center is where the plan facts actually live.
And the part the spec sheet cannot show: OpenAI is generous with usage in practice. Limits reset frequently - lately it has felt like multiple times a month - and the practical mileage per dollar is the best in consumer AI right now.
Why ChatGPT first, even from a Claude guy
People assume I would say Claude here. For a first subscription, I do not, and this is the honest reasoning: you will get more out of the easier tool.
ChatGPT has the better app, the better voice, the smoother on-the-go experience. I use it on my phone constantly - talking through ideas while walking, research on the go. I never do that with Claude; the desktop is where Claude earns its $200 from me. A newcomer lives on their phone. Start where you will actually use the thing.
Claude has the strongest model in Fable 5 and the serious work surface - Cowork, Claude Code - and when your output starts to depend on AI, that comparison tips differently. First $20, though: ChatGPT.The 60-second version, from my Shorts: 300 Million People Use ChatGPT Every Week and Most Have No Idea What It Can Do.Who should NOT buy PlusYou only chat, lightly. Go at $8 covers unlimited everyday text chats. Know the trade: you get Think on GPT-5.6 Luna but not Sol or Astra, and OpenAI lists Go as ad-eligible.
Your AI use is entirely inside your job's tools. If your company gave you Copilot and your AI life is company data, use the approved tool for that and revisit when you have personal use.
You are already deep in another paid ecosystem that is working. Switching costs are real; a working Gemini setup beats a neglected Plus.The two things that would make me cancel
Worth stating in advance so the recommendation stays honest:Strict usage limits. If Plus started cutting real sessions short - use it a little and it bangs out - the value story dies. That is not today's situation, but it is the thing to watch.
Ads for paying users. The moment a paid surface shows ads, I start planning my exit toward local models. Go's docs already carry the "may include ads" line; if that language ever crawls up to Plus, cancel-worthy.Neither has happened on Plus. That is precisely why it is still an easy yes.
How I judge "worth it"
Not benchmarks - replacement value. Plus replaces: a research assistant for planning and digging, a writing partner, a voice thinking-partner on walks, a second opinion on everything. If it saves you two hours a month, it paid for itself at any wage worth automating. For me it clears that bar without trying, which is exactly why the $20 tier is the least interesting bill I pay.
Where you land, and what to doNever paid for AI: ChatGPT Plus, today. Give it one real job this week - planning, a document, research - not just questions.
Budget-tight: Go at $8, with eyes open about what the $12 gap buys.
Power user outgrowing Plus: the two Pro tiers at $100 and $200 are "5x" and "20x" Plus usage per OpenAI's docs - and at that spend, read Claude Max vs ChatGPT Pro first, because the real decision at $200 is between labs, not tiers.
Wondering about the other $20: Claude Pro vs ChatGPT Plus, tier for tier.Last reviewed September 5, 2026. Correction made that day: this page listed ChatGPT Go at $12 and said Go has no reasoning model. Go is $8, read off OpenAI's rendered pricing page, and it does reason via Think on GPT-5.6 Luna. What Go lacks is GPT-5.6 Sol and GPT-6 Astra. Other plan facts verified against OpenAI's official pages on August 21, 2026.
Quick answer: keep Perplexity on the free tier and give the $20 to ChatGPT. Perplexity is a great search engine. ChatGPT is a whole workbench. Only one of them is worth twenty dollars a month.
I remember when this was a real debate. A year or two ago Perplexity was the clever pick - "why pay for all the subscriptions when Perplexity searches everything?" I said versions of that myself. It is not the same fight anymore, and pretending otherwise would be nostalgia, not advice.
My setup, for honesty: Perplexity is wired into my own research tooling and I use it regularly. I also pay for ChatGPT at the top tier. This is a comparison of two tools I actually touch, not a spec-sheet summary.
Quick decisionYour situation
PickYou want cited answers to current-events and research questions
Perplexity, free tierOne $20 budget for an everyday AI
ChatGPT PlusYou mostly write, plan, and think with AI
ChatGPTYou want an agentic browser-computer thing
Watch Perplexity Max's Comet lane, but see the caveatZero budget
ChatGPT free + Perplexity free is a strong stackThe prices, verified
Read off the official pages August 21, 2026, tracked with a dated changelog on my AI plan tracker:Perplexity: Free at $0 ("good for limited daily usage"), Pro at $20/month with 10x the free tier's web answers, uploads, and asset generation, Max at $200 with "maximum Computer usage" and their frontier-model council, per Perplexity's pricing page.
ChatGPT: Go at $8/month, Plus at $20, Pro at $100 and $200, per OpenAI's plan docs. No annual billing.Perplexity's own pricing page, captured August 21, 2026. The $20 Pro tier is the one this post argues most people should skip.
(ChatGPT vs Perplexity, searched the other way around, gets the same answer.)
Perplexity still wins cited search
Cited search, still. This is the product. Ask a question, get a synthesized answer with sources you can actually check. For "what is the current state of X" questions, it remains one of the best interfaces on the internet, and it is the reason Perplexity keeps a seat in my own research stack.
Honesty of format. Answers with receipts train you to check receipts. That habit alone is worth keeping the free tier around.
The agentic swing. Their Comet computer-use direction at the Max tier is genuinely ambitious - a browser-level agent doing tasks, not just answering. I find it more impressive than I expected. It is just priced against the two best products in AI.The 60-second version, from my Shorts: Perplexity's Leaked System Prompt Changes How You Should Use It.ChatGPT wins almost everything else
Almost everything else, and that is the problem for Perplexity Pro. The same $20 that buys you Perplexity Pro buys you: the best consumer AI app, the best voice mode, projects and custom GPTs, scheduled tasks, deep research, image generation, and Codex if you are even slightly technical. Paying $20 for Perplexity gets you maybe 10 to 50 percent of that surface, depending on how you count. I cannot make the math work, however they bundle it.
The limits keep moving in your favor. OpenAI has been resetting usage generously - multiple times a month lately. The practical mileage per dollar on a $20 ChatGPT subscription is the best in the industry right now.
Search caught up enough. ChatGPT's own web search with citations closed most of the gap for everyday questions. Perplexity's edge is real but narrow now, and narrow edges do not win $20 decisions.
The reframe most comparisons miss
This is not really model vs model, it is depth vs breadth. Perplexity went deep on one job: answer questions with sources. ChatGPT went broad: be the everything-tool. Deep-on-one-job products are wonderful as free companions and hard sells as subscriptions when the broad tool does 80% of the job.
Which is exactly why the answer is not "Perplexity is bad." It is: use it where it is deep, pay where the breadth is.
How I compared these
Prices from the official pages, August 21, 2026, dated on the tracker. Usage: Perplexity as a working part of my research tooling; ChatGPT as a daily paid driver at the top tier. Judgments labeled as mine.
Four readers, four callsResearcher-brain, love citations: free Perplexity forever, and pay ChatGPT when you need a workhorse. That is my setup.
One $20 slot: ChatGPT Plus - my full verdict lives at Is ChatGPT Plus worth it. If $20 stings, Go vs Plus is the $12 conversation.
Considering Claude instead: different comparison, written here. For research work specifically, Claude vs ChatGPT for research is the deeper cut.
Eyeing Perplexity Max at $200: you are in Claude Max vs ChatGPT Pro territory - read that first.Last reviewed September 5, 2026. Correction made that day: this page listed ChatGPT Go at $12. It is $8, read off OpenAI's rendered pricing page. Other prices verified against the official pages on August 21, 2026.
Quick answer: Claude, by a mile, if the question is which tool is better. But that is rarely the real question. The real question is what to do when your job hands you Copilot for free, and the answer to that one is: use Copilot for work data at work, and pay for your own AI for everything that is actually yours.
The Ship Lean line: Copilot is the tool your job gives you. Claude is the tool you would choose.
My receipts here are unusually direct. I use Copilot at my day job, daily. I pay for Claude Max at $200/month on my own dime and run my whole operation on it. This comparison is lived on both sides.
Quick decisionYour situation
PickCompany documents, emails, meetings
Copilot. It is the approved tool.Your own writing, thinking, projects, career
ClaudeYou want an agent that does real work (files, code, chores)
Claude (Cowork / Claude Code)Your company enabled almost nothing in Copilot
Claude, and stop blaming yourselfBudget is zero
Copilot at work + Claude's free tier at homeWhat Copilot is actually like, from inside a job that uses it
I can tell you confidently, because I live it: it is rough. The best thing about our Copilot is that it can run GPT-5.6-class models under the hood - the model itself is smart. And it is still clunky in ways that matter every single day.
It reminds you of a plain chat window. Sometimes the web features are just not there, because your IT team decides what is switched on. It is built as enterprise software first, and it succeeds at exactly that: compliance, tenant data, admin control. Nobody has ever told a friend "you have to download Copilot." You use it because your job has it.
That last sentence is the whole product review.
To be fair to Microsoft, there are two different products called Copilot - the free chat most employees get versus the paid add-on that can actually see your email and files - and I broke that split down in Copilot vs ChatGPT. If your company never bought the add-on, half of what the ads promised does not exist for you, and no amount of prompting will fix it.
(Copilot vs Claude, if you searched it the other way - same page, same verdict.)The 60-second version, from my Shorts: The AI you already pay for is hidden inside Docs, Word, and Notion.Claude wins on model, surface, and ownership
The model. Fable 5 is the strongest model I have used. Even Opus 5 on high effort is phenomenal. When output has to be right and has to follow instructions exactly, this is the tool. Copilot's underlying models are capable; Claude's frontier is simply ahead, and you feel it on hard tasks.
The working surface, and it is a landslide. Claude Cowork points the model at your regular office chores. Claude Code builds actual systems. Projects hold context across real work. Copilot, as most employees experience it, is a chat box. The distance between "chat box" and "agent that does the chore" is the biggest capability gap in this entire comparison.
It is yours. Your goals, your finances, your side project, your job search - none of that belongs in your employer's tenant, and your employer's logs. A tool you personally pay for is the only honest place for your own life. This is the argument I care most about, and it has nothing to do with benchmarks.
Copilot genuinely wins inside the tenant
Compare, do not co-sign - Copilot has real wins:It is the allowed tool. For company data, it is often the only place you are permitted to paste anything. That is frequently the entire decision.
It is inside the apps your company lives in - Word, Excel, Outlook, Teams - if your company bought the paid add-on.
It costs you nothing personally, and for meeting summaries and email cleanup on work data, the free tier is genuinely enough for a lot of people.
It is better than nothing, by a lot. If Copilot is your only option, use it hard. A clunky frontier model still beats no model.The prices, verified
Read off the official pages August 21, 2026, tracked on my AI plan tracker:Claude: Pro at $20/month ($17 annual), Max at $100 and $200, per Anthropic's pricing page.
Copilot, consumer side: here is the August 2026 surprise - the standalone Copilot Pro plan is gone. Microsoft folded consumer Copilot into Microsoft 365: Personal at $9.99/month, Family at $12.99, Premium at $19.99 with the "highest" AI usage tier, per Microsoft's pricing page. The work version is a separate per-seat license your employer buys.Microsoft's own individual-plans page, captured August 21, 2026. Copilot now lives inside Microsoft 365 tiers - the standalone Copilot Pro consumer plan is gone.
How I compared these
Copilot: daily use at my actual day job, on a real corporate deployment - which means my experience partly reviews my employer's configuration, a caveat every honest Copilot comparison owes you. Claude: daily paid use at the $200 tier for my own business. Prices: official pages, dated, on the tracker.
Four seats, four movesEmployee with free Copilot, AI-curious: use Copilot for work things, then get Claude's free tier tonight and give it one real personal task. The gap will make the decision for you.
Deciding whether your own $20 is worth it: Is Claude Pro worth it is my full answer. If ChatGPT is also in the running, start with Claude vs ChatGPT.
Employer bought the full Copilot add-on: genuinely use it for the work data - that grounding in your email and files is its one unique power - and keep your own tool for your own life.
Manager picking a direction: the honest read is that employees quietly route around Copilot when it is locked down. Ask what is actually enabled before concluding anything about AI.More on using a serious AI tool inside a normal job: Claude at Work.
Last reviewed August 2026. Prices verified against the official pages on August 21, 2026.
Quick answer: if your life already runs on Google - Gmail, Docs, an Android phone, a family plan - Gemini is genuinely fine and I am not going to pretend otherwise. If AI is where your actual work happens, Claude is the stronger tool, with the stronger model, and it is not close.
The Ship Lean rule for this one: if your life runs on Google, Gemini is enough. If your work runs on AI, it isn't.
My receipts: I pay for Claude Max at $200/month and it is the center of my work. And I am not a Gemini stranger - I used it for image generation in my content pipeline for a long time, and my household literally pays for it: the subscription came bundled with the phone, and my wife uses it daily. So this is not a Claude guy dunking on a tool he never opened.
Quick decisionYour situation
PickAI for everyday life: questions, planning, personal budgeting
Gemini, keep itDeep in Google Workspace at your job
Gemini first, it is already in your appsWriting, documents, or building things that must be right
ClaudeYou want an office agent or coding agent
Claude (Cowork / Claude Code)Heavy image generation on a budget
Gemini earns itResearch from a pile of sources
Try NotebookLM before paying anyoneThe prices, verified
Read off the official pages August 21, 2026, tracked with a dated changelog on my AI plan tracker:Claude: Pro at $20/month ($17 annual), Max at $100 and $200 for the 5x and 20x tiers, per Anthropic's pricing page. Since July 20, Fable 5 runs on separate usage credits on Pro; Max includes it up to 50% of weekly limits.
Gemini: AI Plus at $4.99/month with "2x higher usage limits than Free," AI Pro at $19.99 with 4x, AI Ultra from $99.99 with "up to 20x higher usage limits in Gemini than the Pro plan," per Google's plan page.Google's own plan page, captured August 21, 2026. Plus, Pro, Ultra - with storage bundled into every tier.
The quiet money detail: Google's plans bundle storage (5 TB on AI Pro) and even YouTube Premium tiers. If you already pay for Google One and YouTube separately, Gemini's effective price can drop toward zero. Nobody else in this market can do that.
(Gemini vs Claude, if you searched it that way around - same fight, same split.)
Gemini wins on everywhere-ness and the bundle
It is already there. In Gmail, in Docs, in Sheets, on the phone in your pocket. The best AI is partly the one with zero setup, and for millions of people that is Gemini by default. If you have been using it for months and it is working, keep using it. I mean that.
Everyday life. For general questions, planning, summarizing, personal budgeting - it handles all of it at a level that is honestly fine. Feed it a messy life problem and it does a decent job.
Image generation. Google's image models are phenomenal bang for the buck. This was my lane for a long time: I ran Gemini image generation inside my own content pipeline and it delivered. If visuals are your main AI use, weigh this heavily.
NotebookLM. The underdog of the whole Google lineup. Source-grounded research over your own documents, and one of the best AI products nobody talks about enough.The NotebookLM homepage, captured August 21, 2026 - Google now brands it Gemini Notebook. Still the sleeper pick for source-grounded research.
The bundle math. Google's paid AI tiers carry storage and even YouTube Premium along for the ride, and the ladder starts at $4.99 - a real paid tier at a price nothing else in this comparison can touch.The 60-second version, from my Shorts: 7 free Google AI tools you should already be using.Claude wins the work that pays you
The model. Fable 5 is the strongest model I have used, and the gap matters exactly when the work matters: long documents, careful instruction-following, code, anything where a plausible-but-wrong answer costs you real time. Gemini's models are good. Claude's frontier is better, and the labs pushing hardest right now are Anthropic, OpenAI, and xAI. Google is in the race; it is not setting the pace.
The working surface. Claude Cowork points the model at your actual office chores. Claude Code builds real things. Projects hold context across a real piece of work. This is the difference between an assistant you ask and a tool you work in. When people ask why I pay $200 instead of $19.99, this is the answer.
Instruction-following under pressure. The boring superpower. When output has to follow a spec - a template, a voice, a structure - Claude drifts less. That is worth money in any job where you ship words or systems.
The part both sides get wrong
The comparison pages frame this as model vs model. For most readers it is actually ecosystem vs tool. Gemini's real pitch is "your Google life, now with AI." Claude's real pitch is "your work, done better." Those are different products, and that is why "which is smarter" misses.
It also means the honest answer can be both: Gemini for the Google-shaped parts of life, Claude for the work that pays you. That is roughly my house's split, and neither subscription feels wasted.
How I compared these
Prices from the official pages, August 21, 2026, dated on the tracker. Gemini experience: my own image-generation use in a production pipeline plus a household subscription in daily use. Claude experience: daily paid use at the $200 tier. Opinions labeled as mine.
Four situations, four honest callsHappy Gemini user, AI is life-admin: stay. Spend the $20 difference on coffee.
Google-native but work is getting AI-heavy: keep Gemini, trial Claude Pro for a month on your heaviest work task. My read on that tier: Is Claude Pro worth it.
Choosing a first serious work tool: Claude, sized via Pro vs Max. If ChatGPT is also on your list, that comparison is here.
Forced onto a different tool at work entirely: that is usually Copilot, and I wrote that one from my day job.More on making Claude earn its keep: Claude at Work.
Last reviewed August 2026. Prices verified against the official pages on August 21, 2026.
Quick answer: Claude, for real work, today. It has the strongest model in Fable 5, and the working surface around it - Claude Code, Cowork, Projects - is something Grok does not match yet. But I will say something I never expected to write: Grok is the first tool that made me want a third $200-class subscription.
Claude runs my business today. Grok is the bet I want to place. Budget is the only referee.
Where I am coming from: I pay for Claude Max at $200/month and it sits at the center of how I work. I have used Grok through the trial - the research agent, spinning up bots that talk to each other - not as a daily driver. This page is honest about that split: my Claude takes are lived, my Grok takes are a taste plus a hard look at what xAI is shipping.
Quick decisionYour situation
PickWork is documents, thinking, building things you rely on
ClaudeYou want a coding or office agent with training wheels off
Claude (Code / Cowork)You want always-on agents running jobs on their own computer
Grok Bot laneYou live on X, want frontier AI for $8
Grok via X PremiumOne subscription, most complete everyday product
Honestly, neither is my first pick for a newcomerThe prices, verified
Read off the official pages August 21, 2026, and kept current on my AI plan tracker:Claude: Pro at $20/month ($17 annual), Max at $100 (5x) and $200 (20x), per Anthropic's pricing page. One wrinkle worth knowing: since July 20, Fable 5 sits outside Pro's plan usage as paid usage credits, while Max includes it up to half your weekly limit.
Grok: SuperGrok at $30/month, SuperGrok Plus at $100, free tier with "generous limits," per x.ai's pricing page. X Premium at $8/month bundles increased Grok limits.xAI's own pricing page, captured August 21, 2026.Anthropic's own pricing page, captured August 21, 2026. The Fable-5-on-Pro usage-credit wrinkle lives in the fine print here.
(If you searched Claude vs Grok, same page, same verdict.)
Claude wins the work that has to be right
The model, for now. Fable 5 is the strongest model I have used, and I use it all day. When work has to be right - a system that touches my files, a long document, code that ships - Claude is the one I trust with it.
The working surface. This is the underrated gap. Claude is not just a chat window: Claude Code for building, Cowork for pointing the same brain at your regular office work, Projects for keeping context. Grok is still mostly a very good chat with very good search. For the depth of stuff you can hand over, Claude wins by a mile.
Anthropic is being pushed, and users benefit. Competition is doing its job: Anthropic has been running 50% higher weekly usage limits on paid plans since May and keeps extending the boost, currently through August 31 with talk of making it permanent. My read is that pressure from OpenAI's generous resets and xAI's pricing is exactly why. Good.
Grok wins on velocity and the agent bet
Velocity, and it is not subtle. xAI has poured money into data centers and hardware at a scale that even competitors respect, and the model progress shows it. Grok went from a curiosity to a top-five model in about a year, and the pace keeps closing gaps that looked permanent six months ago. A year ago this comparison would have been a joke. It is not a joke now.
Grok Bot is a new category. Launched August 11: always-on agents with their own cloud computers that sign into your tools, run multi-step jobs unsupervised, learn routines by demonstration, and coordinate with each other in threads. Here is why that lands for me personally: I run a self-hosted AI operator on a Mac mini, and it took real tinkering to build. Grok Bot is that idea as a product anyone can turn on. If you ever wanted that setup and bounced off the complexity, this is your door in.
The mission is easy to bet on. This is a taste call, labeled as one: I believe the xAI story - the SpaceX pedigree, the hardware aggression, the shipping pace. Some tools you pay for what they are. Some you pay for where they are going.
Price of entry. $8/month via X Premium against Claude's $20 floor. For a taste of a frontier model, Grok is the cheapest seat in the building.
The honest wall: my budget, probably yours too
I already spend $400/month on AI subscriptions. Grok wants another $30 to $100. I missed the $99-for-three-months promo, and the honest truth is I have not found the revenue line that pays for seat number three yet.
That constraint is the actual comparison for most people. Not "which lab is winning" - which tool earns its slot this month. Claude has a deep stack of receipts in my workflow. Grok has a trajectory and a trial.
How I compared these
Prices and plan facts: official pages, read August 21, 2026, dated changelog on the tracker. Grok Bot claims: xAI's launch coverage, linked above. Experience: Claude from daily paid use at the $200 tier; Grok from trial use, labeled as such throughout. No affiliate stake in either.
Four seats, four callsBuilding or running anything real on AI: Claude. Start at Pro vs Max to size the plan; my full read on the $20 tier is here.
Wanted a self-hosted agent, never built one: watch Grok Bot closely. This is the productized version.
On X daily: Premium at $8 and just use Grok there. Cheapest experiment in AI.
Comparing Grok against the other giant instead: Grok vs ChatGPT.More on making Claude earn its keep at a normal job: Claude at Work.
Last reviewed August 2026. Prices verified against the official pages on August 21, 2026.
Quick answer: if you are picking one subscription today, pick ChatGPT. It is the more complete product and the $20 earns its keep every day. But Grok is not the punchline it was a year ago. It went from a top-ten contender to a genuine top-five model, xAI is out-building everyone on data centers, and Grok Bot just shipped the most interesting agent product of the summer.
ChatGPT earns today. Grok is a bet on next year. That is the Ship Lean read, and the rest of this page is me showing my work.
Full disclosure up front: I pay for ChatGPT and Claude at $200 each, every month. I have used Grok through the trial, not months of daily driving. So this is not a fake "I tested both for 30 days" post. It is an honest comparison of a tool I run my business on against a tool I am actively trying to justify adding.
Quick decisionYour situation
PickFirst AI subscription, want the most product for $20
ChatGPT PlusOn a budget, mostly writing and everyday questions
ChatGPT Go at $8, or Grok's free tierYou live on X and want AI where you already are
Grok via X Premium at $8You want always-on agents doing jobs while you sleep
Grok Bot is the one to watchYou already pay for ChatGPT or Claude and have $30 spare
SuperGrok is the most interesting add-on of 2026What each one costs, verified
I read these off the official pages on August 21, 2026, for my AI plan tracker:ChatGPT: Free at $0, Go at $8/month, Plus at $20, and two Pro tiers at $100 and $200. Per OpenAI's own plan docs, Pro $100 is "5x higher usage than Plus" and Pro $200 is 20x. No annual billing on any of them.
Grok: Free at $0 with "generous limits," SuperGrok at $30/month, SuperGrok Plus at $100, per x.ai's pricing page. The side door: X Premium at $8/month includes "increased usage limits on Grok," and Premium+ at $40 goes higher.xAI's own pricing page, captured August 21, 2026 - note the SpaceX logo in the corner. Free, $30, and $100 tiers.
Prices on this page age. The tracker link above is where I keep them current, with a dated changelog.
(Searching this the other way - ChatGPT vs Grok - lands you the same answer: same fight, same split.)
ChatGPT wins on the product around the model
The product around the model. This is the whole case, and it is a strong one. The app is the best in the category, voice is the best in the category, and the feature list is deep: deep research, scheduled tasks, custom GPTs, Codex for anyone technical, and record mode. I use ChatGPT on my phone, on the go, constantly. It is the AI I talk to out loud.
The limits keep resetting in your favor. OpenAI has been unusually generous lately, resetting usage multiple times a month. When I hit a wall, the wall usually moves before I do. For a $20 product, the mileage is absurd.
It is the safe recommendation. If a stranger asks me "which AI should I pay for first," I say ChatGPT without hesitation. Not Claude, not Grok. The experience is just easier, and a newcomer gets more value from a polished generalist than from a specialist tool.
Grok wins on velocity, agents, and price of entry
Velocity. This is the part that changed my mind about xAI. They have invested in data centers and hardware harder than anyone, and it shows in how fast the models improve. Grok went from an afterthought to a top-five model in about a year. Betting on the team that ships fastest is a real strategy, and it is the Ship Lean bet: xAI is the SpaceX of this race.
Grok Bot. Launched August 11, 2026: always-on agents that get their own cloud computer, sign into your tools, finish multi-step jobs unsupervised, and coordinate with each other in group threads. If you have wanted a self-hosted AI operator setup but could not stomach the tinkering, this is that idea as a product. In my trial the multi-bot part was the thing that stuck with me: open a research agent, spin up a second bot, and watch them hand work to each other.
The X integration. Real-time search over X is genuinely unique. No other assistant has that data, and for news, sentiment, and what-is-happening-right-now questions, it shows.
The bundle math. $8/month for X Premium with increased Grok limits is the cheapest way into a frontier model that exists right now.
The honest budget wall
Here is the part the affiliate comparisons never write. I want Grok. I missed their $99-for-three-months promo and I am still annoyed about it. But I already pay $400/month across Claude Max and ChatGPT Pro, and another $30 to $100 has to be earned by revenue, not enthusiasm.
That is the real decision most readers face, just at a different scale. The question is never "which model is smarter." It is "which subscription earns its spot this month." ChatGPT has years of proof in my workflow. Grok has a trial, a great trajectory, and a promo I missed.
If your budget is one subscription: ChatGPT. If it is zero: Grok free tier plus ChatGPT free tier is a genuinely strong stack now.
How I compared these
Pricing and limits come from the official pages, read August 21, 2026, and tracked with dates here. Product claims about Grok Bot come from xAI's launch and its coverage, linked above. The experience calls are mine: ChatGPT from daily paid use across a year-plus, Grok from trial use only, and I have labeled them that way in the text.
Pick your seat: four situations, four callsNewcomer with $20: ChatGPT Plus. Do not overthink it. If $20 stings, Go vs Plus is the $12 question.
Power user with subscriptions already: keep what earns, put Grok's free tier in your rotation, and watch Grok Bot. That is where I am.
X native: Premium at $8 is the cheapest frontier-model seat in the game.
Claude person wondering about the other side entirely: I wrote the Grok vs Claude version of this decision too, and the wider Claude vs ChatGPT call.Last reviewed September 5, 2026. Correction made that day: this page listed ChatGPT Go at $12. It is $8, read off OpenAI's rendered pricing page. Other prices verified against the official pages on August 21, 2026.
Quick answer: every feature is identical across Claude Max 5x and Max 20x. Same Fable 5 and Fable 5.1 access, same Opus 5 default, same priority access, same 200k default context window with 1M on Fable 5.1, Opus 5, and Sonnet 5. The only thing $200 buys over $100 is capacity.
That makes this a simpler decision than the SERP suggests. You are not choosing a better Claude. You are choosing how often you want to wait.
Which leaves one real question: how much capacity is "20x" actually?
Right now there are three answers to that on the table, and they do not agree.Source
What 20x is worthAnthropic, published
20x Pro usage, per five-hour sessionA filed class action, alleged
roughly 6x ProMe, paying for 20x since June
feels like 10x to 15xNobody writing about this plan shows you all three at once. Most pages print Anthropic's number and stop. A few print the lawsuit and get angry. None of them tells you what to actually do.
I started on 5x and moved to 20x. I mined nine weeks of my own logs on the 20x tier, and the honest finding points down the ladder, not up.
Every feature gate is the same. Here is the proof
Most pages imply 20x is a better product. It is not.Max 5x ($100/mo)
Max 20x ($200/mo)Usage per 5-hour session
5x Pro
20x ProFable 5 and Fable 5.1
50% of weekly limits
50% of weekly limitsOpus 5
Default model
Default modelContext window
200k default; 1M on Fable 5.1, Opus 5, and Sonnet 5
200k default; 1M on Fable 5.1, Opus 5, and Sonnet 5Higher output limits
Yes
YesPriority access at peak
Yes
YesEarly access to features
Yes
YesAnnual billing
No
NoSources: Anthropic's Max plan page for the multipliers and priority access, Claude's pricing page for the Fable 5 and context rows, and the Opus 5 announcement for the default model.
You do not have to take my table for it. Here is Anthropic's own comparison grid, re-checked on September 10, 2026:claude.com/pricing, captured September 10, 2026. Read the Max 5x and Max 20x columns against each other, row by row. They are the same column twice.
The plan cards above that grid say the rest of it out loud.claude.com/pricing, September 10, 2026. Max is "From $100," per month, and what it sells is "Choose 5x or 20x more usage than Pro."
Not better models. More usage. And still no annual option on either Max step, while Pro has one.
One row deserves attention because nobody prints it: neither Max step offers annual billing. Claude Pro does, at $17 a month billed as $200 up front. Max is monthly only, at both steps. If you were hoping to soften the jump by committing for a year, that option does not exist.
What the multiplier officially means, and what it does not
This is where the community argument lives, so let us be precise about what is published and what is not.
Anthropic's wording is per session: Max 5x gives "five times more usage per session than the Pro plan," and Max 20x gives "20 times more usage per session." The session window resets every five hours.
On top of that sits a separate weekly limit across all models, which resets at a fixed time each week assigned to your account.Anthropic's Max plan documentation, captured August 14, 2026. Every listed benefit apart from the usage multiplier applies to both steps identically.
What Anthropic does not publish: how the weekly limit scales between the two steps. Or any numeric token or message quota at all, on either step.
That gap is exactly what the threads are arguing about. u/cavebreeze on r/ClaudeAI put the question well on August 9: "I heard several people claim 5x Max isn't a 5x for the weekly limits, but 5x for the 5 hour session limits and around 3x for the actual weekly limits, and that 5 individual pro accounts yield more usage. Has anyone tried this?"
I want to be careful here.
That 3x figure is a user estimate, repeated between users. It is not an Anthropic number and I have not been able to verify it. What is verifiable is that the official multiplier describes sessions, and that the weekly cap is a separate mechanism nobody outside Anthropic has the numbers for.
Anthropic is being sued over exactly this
That unpublished gap is not just a forum argument any more. It is in a filed complaint.
A proposed class action filed June 14, 2026 alleges Anthropic misled Max subscribers about usage limits. I want to be careful with the wording here, because care is the whole point: these are allegations in a complaint, not findings, and no court has ruled on any of them.
What the complaint alleges, per the reporting:Max 5x delivers roughly 3.5x Pro's Sonnet hours, not five.
Max 20x delivers roughly 6x, not twenty.
Lead plaintiff Karl Kahn of Washington upgraded to 20x for heavy coding, and one five-hour session consumed 15% of his weekly quota.
Claims are false advertising and unfair business practices, with an amount in controversy over $5 million, covering US purchasers of either Max tier since the plans launched in April 2025.Reported by Engadget, Quartz, and tracked at OpenClassActions.Quartz's coverage of the June 2026 filing, captured September 5, 2026. The dispute is about the gap between a per-session multiplier and an unpublished weekly ceiling.
Now the part where I have to be fair to Anthropic, because most coverage is not.
Anthropic's published wording has always described the multiplier per five-hour session. That is on the support page, it is quoted higher up this article, and it is not hidden.
What the reporting does not say is what the complaint's figures are measured against, and I have not read the complaint. So I am not going to reconstruct the plaintiffs' arithmetic for them.
What I can say is that two things can be true at once. The wording can be technically accurate and the average subscriber can reasonably have read "20x" as "twenty times the usage," because nobody buys a plan thinking in five-hour windows.
That is the honest shape of the dispute. It is a grey area created by what is not published, and as a consumer you are allowed to be annoyed about the greyness without anyone having lied to you.
What a real 20x meter actually says
So here is the third number, from the only account I can show you.My own Max 20x meters, captured September 5, 2026. Session 32%, all models 54%, Fable 51%. Read the notice in the middle honestly: my limits were temporarily boosted that week, with the weekly Claude Code limit 50% higher through September 13. I cannot separate that boost from the all-models bar, so treat this reading as flattering to me, not damning of Anthropic.
I use this thing most days and I run agents. My honest read, as the guy paying $200 a month, is that it is not truly 20x. It might be 10x to 15x.
That is a feel, not a measurement, and I am labeling it as one. I have no more access to Anthropic's weekly formula than the plaintiffs do. What I can tell you is that my usage definitely feels like it drains faster than Codex does, which is the closest comparison I run day to day. That is an impression from one operator, not a measurement, and it belongs in the same column as the complaint's numbers rather than above them.
Three numbers, one plan, and the only one with a published methodology is the one describing five-hour windows.
Which is precisely why the next section matters more than any of them. The multiplier is not what decides whether you should upgrade. Your habits are.
What it actually takes to exhaust a Max allowance, measured
Here is where I can offer something the rest of this SERP cannot: my own logs.
I mined nine weeks of my Claude Code usage on the 20x tier. The full measurement is its own post, but the number that matters for this decision is this one:
Limit and overload events came to roughly 119 hits across 220,861 log lines. About 0.05%, retries included.
And I am nearly never stopped.
The reason is that 97.4% of my tokens are prompt cache reads, not new work. Agentic work re-reads the same context constantly, and cache reads are far cheaper than fresh input. Heavy usage looks enormous and mostly is not.
So when you read a thread where someone says they blew through a Max allowance, the useful question is not "how many tokens." It is "how much of that was new."
Now the correction I owe you, because my own data could be read as a flex.
I hit the weekly wall during a sloppy stretch in August. On 20x.
That was me spending badly: long sessions riding up toward the maximum context window until they auto-compacted. Terrible habits. The wall is real even at the top step if you work that way, and pretending otherwise would be dishonest.The 60-second version, from my Shorts: One task ate my whole AI limit, weekly. Here's the fix..Start at 5x unless one of these is true
The honest default is the lower step. Move up if:You code heavily. Serious apps, games, websites, client work. That genuinely burns tokens in a way chat never will.
You are waiting instead of working. Not "I got a warning." Actually stopped, mid-task, on something with a deadline.
You run long agent sessions daily and have already fixed your context habits.And one warning for anyone new to Claude: expect to be inefficient at first. You will burn far more than you need to for a few weeks, and then you will learn the habits, a decent project instructions file, starting a fresh session around 500k of context, routing routine work to cheaper models. Do not buy a tier to paper over a learning curve you are about to climb anyway.
Why I went to 20x, and whether you should
I started on 5x and I kept running out.
Sometimes I would wait out the five-hour window. Then one day I was done with that. I was not going to sit there waiting when I wanted to edit a video or build a skill right then. Waiting is unproductive, and it is the most expensive thing on this page.
So I upgraded. It felt scary at first, genuinely expensive.
Then I did the math on video editing alone. If you hired an editor just to cut your videos, not even adding animation or visual aids, you would pay more than $200 a month and wait days for turnaround. The arbitrage is real, and I took it.My own billing screen. I am on Max 20x at $200 a month, which is the tier these measurements come from.
But notice that my reason was a specific, repeating, expensive job that I could price against an alternative. Not a feeling that the bigger number was safer.
Do not take my word for it. Do not even take your own word for it. Start at the $20 plan and work your way up. That is my rule for this ladder, and it is how I ended up here without regretting a step. Not one of my upgrades came from a comparison table. Every one came from a wall.
The two-5x-accounts question, answered honestly
This is the most-discussed idea in this space, and the top-ranked thread on r/ClaudeAI argues two accounts are the better buy for solo builders.
The math that makes it tempting is real: two Max 5x accounts cost $200, and so does one Max 20x.
Here is my honest read, and I am labeling it as my read rather than a calculation.
The account policy question is genuinely unresolved. Community threads contradict each other flatly. On r/ClaudeCode, one top answer warns that accounts get banned for this; another says Anthropic staff have publicly said it is fine. I could not verify Anthropic's current position, so I am not going to tell you it is safe or unsafe. Read the usage policy yourself before you try it.
Even setting policy aside, I think it is a bad trade. You would be logging out, losing context on in-flight chats, and logging back in throughout the day. Yes, you can point both at the same repo so nothing is technically lost. But the context switching alone eats whatever allowance you gained. My expectation is you net roughly nothing and spend the day managing logins.
And the deal itself favors one account. The multipliers stack on a single subscription, which is my argument for why one 20x beats two 5x at the same price rather than a precise quota comparison. Anthropic does not publish the numbers that would let anyone make that calculation exactly.
If you are genuinely maxed out and this is a real business, where you are running the top model hard on a problem that pays for itself, the aggressive move is a second $200 plan. Not two 5x accounts to reassemble the same total.
Two 5x accounts to reach the same ceiling is asking for inefficiency, and probably asking for trouble.
One billing trap before you switch either way
Users report a real gotcha worth knowing: downgrading through support rather than in Settings can cost you a legacy price you were grandfathered into.
Change plans in Settings > Billing where you can. Downgrades take effect at the end of your current billing period, and your chats, projects, and files stay with your account either way.
The one thing to measure this week, if you are on 5x
If you are on 5x and trying to decide whether to move up, here is the whole answer, and it is shorter than this page deserves.
Do not measure anything.
Seriously. You are not going to learn anything useful from a percentage, and staring at a meter is how people talk themselves into $100 a month. The multiplier argument above is genuinely interesting and it should not be your input.
Measure one thing instead: when you run out, can you wait?
Say you are building something simple and you hit the wall. If your reaction is "fine, I have other things to do, I will step away and come back in five hours," then it is not that serious. You have your answer, and it is no.
But if you need it right now, and you can feel that what you just did is the tip of the iceberg and there is a lot more you want to do, you might need to upgrade.
Let the pain tell you. That is the entire rule, and it is the one thing on this page that no amount of unpublished weekly-limit math changes.
If you are still decidingStill on the $20 plan and wondering if it holds? Is Claude Pro enough
Not sure Max is the right family at all? That is Claude Pro vs Claude Max
Not a developer and wondering if any of this applies? Is Claude Max worth it
Want the mechanics of the two meters? Claude usage limits explained
Comparing across vendors at the $100 line? Claude Max vs ChatGPT ProPublished August 14, 2026. Substantially updated September 10, 2026: added the filed class action, a real Max 20x meter, my own honest estimate of the gap, and a closing decision rule. On September 10 the feature-parity grid and the monthly-only "From $100" Max pricing were re-captured from claude.com/pricing and both are screenshotted above. The model-access and context-window rows in the table near the top were NOT re-verified on September 10 and still carry their August 14 check against Anthropic's Max plan support article; Fable 5.1 became Anthropic's most capable model on September 1, so treat those two rows as the oldest thing on this page. Multipliers, session and weekly mechanics and priority access were verified August 14 against Anthropic's Max plan support article, linked inline. Anthropic publishes no numeric quota for either step; every number attributed to users is labeled as a community report. The lawsuit figures are allegations in a filed complaint as reported by Engadget and Quartz, not findings, and no court has ruled. My 10x to 15x estimate is a subjective impression from a paying account, not a measurement. Anthropic's position on multiple simultaneous accounts was not verified and is deliberately not asserted here. I looked for additional named Max 20x subscribers posting about the multiplier with live links and could not verify new ones for this update, so that section was not added rather than filled with anonymous paraphrase.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: I pay $200 a month for Claude Max 20x, I use Claude Code most days, and I almost never hit a usage limit. Across nine weeks of my own logs, limit and overload events account for about 0.05% of 220,861 log lines.
The reason is not restraint. It is that 97.4% of my tokens are prompt cache reads, not new work.
Heavy usage looks enormous and is mostly cheap context re-reads. That single fact is missing from every "am I using too much" thread I have read, so I mined my own logs and published the numbers.
One thing up front: these are one operator's numbers on the top consumer tier, with specific habits. They are not a benchmark and they are not what your usage will look like.I walk through this on camera in Every Dollar I Spend on AI Automation (Real Numbers) (10 min).How I measured this, and what I cannot prove
Credibility here comes from the caveats, so they go first rather than in a footnote.
Raw verified window: June 11 to August 14, 2026. I streamed all 1,587 session log files and deduped token counts by request ID so retries and repeated writes do not double count.
Lifetime figures: these come from Claude Code's own stats cache covering December 31, 2025 to August 11, 2026. I label them "as reported by the app" everywhere they appear, because the raw logs from before July were rotated out and I cannot independently verify them.
That distinction matters. The headline number people would want me to lead with, 64.4 billion lifetime tokens, is the weakest number I have. The nine-week verified figure is the one I would defend.
One more honesty note, because it is funny and it tells you something about log data: my longest recorded "session" runs about nine days. That is not a nine-day session. That is a terminal window I left open.
If a stat below has no caveat attached, it came from the verified window.
The numbers
Verified, nine weeks (June 11 to August 14, 2026):Measure
ValueTotal tokens
11.29BPrompt cache reads
97.4% of tokensGenuinely new output
29.8M tokens (0.26%)Biggest single day
1.29B tokens (August 7, a Friday)Limit / overload events
119 hits across 220,861 log lines (**0.05%**)Median session length
3.2 minutesSessions under 1 minute
~30%90th percentile session
94 minutesMedian messages per session
53As reported by the app (Dec 31, 2025 to Aug 11, 2026): 2,742 sessions across 159 active days, roughly 17 sessions per working day, 64.4B tokens.
Model mix by request, nine weeks:Model
Share of requestsOpus
60%Fable
27.8%Opus 4.8
9.9%Sonnet
2.2%Haiku
0.04%Time of day: activity peaks at 8 AM and again from 3 to 4 PM ET. Dead between 4 and 6 AM.
The 97.4% is the whole story
If you take one thing from this page, take this.
Cache reads are not new work. When an agent works in a repo, it re-reads the same files, the same instructions, and the same conversation over and over. Anthropic caches that, and a cache read costs a fraction of fresh input.
The API price sheet makes the ratio concrete: Opus 5 input runs $5 per million tokens while a cache read runs $0.50. Ten to one.Anthropic's published API rates, captured August 14, 2026. The read row is why 11.29 billion tokens is a much smaller number than it looks.
So "64 billion tokens" is not 64 billion tokens of thinking. It is a small amount of new reasoning wrapped in an enormous amount of cheap re-reading.
This is why an extreme user rarely hits the wall. It is also why the token numbers people post in Reddit threads are almost meaningless without the cache split, and nobody ever posts the cache split.
Median session: 3.2 minutes
The other surprise in my own data was what the sessions looked like.
I expected marathons. I got quick draws.
Half my sessions are under 3.2 minutes and about 30% are under a single minute. The long tail is real, with a 90th percentile of 94 minutes, but the everyday pattern is: open it, ask for the thing, close it.
That maps to how I actually work, and it is worth saying because "power user" imagery usually shows someone locked in for six hours. Mine looks more like a couple of dozen short visits a day, going by the app-reported session count against active days.
The day the limit taught me routing
Here is the anecdote that changed my habits, and it is not flattering.
I was using the top model to QA a skill. Not to build anything. Just to check my work.
I burned more than 20% of my usage in one day. By day two or three I was at roughly 50%. By day four I was at 70 to 80%.
And I started getting stingy. Damn, I thought, I should not be using the top model for that.
The limit never actually blocked me. It taught me routing.
That is the reframe I would hand to anyone panicking about caps: the limit is not a punishment, it is a pricing signal. It tells you when you are spending a premium model on a commodity job.
I want to be honest about the current state of this too, because my own data could be read as bragging. I hit the weekly wall this week. On the 20x tier. That was me riding long sessions up to the maximum context window until they auto-compacted, which is exactly the habit I am about to tell you to fix.
The wall is real even at $200 if your habits slip.
Habit 1: a context meter and a fresh-session handoff
This is the one where most of my spending was hiding.
I run a context meter in my status line. It shows green, yellow, red. Around 500k it visually grabs my attention.
At that point a hook fires and auto-writes a NEXT-STEPS markdown file at the repo root. I close the window, open a fresh session, say "read the next-steps file," and carry on with zero context but the same knowledge.
Two reasons this matters, and the second one surprised me:At 800k context you burn tokens dramatically faster, because every single message re-reads all of it.
Ironically, the quality deteriorates. A model swimming in 800k of accumulated history is not sharper than one handed a clean brief.Before I built this, I rode sessions until they auto-compacted. That is where the 20%-plus days came from.
Habit 2: subagents, and the payload rule
The main agent stays the brain. Grunt work goes somewhere else.
Research, file sweeps, mechanical checks: I spawn those to subagents. Each one gets its own context window, so five or ten of them can run a large parallel sweep without touching my main session's budget at all.
The catch is real and most people hit it: give them a good payload. Clear instructions and the right tool access up front.
Send a subagent off with a vague one-liner and it will do basic research and hand you garbage back. The parallelism is free. The quality is not.
Habit 3: model and effort routing
Sonnet 5 is the floor for grunt work. Haiku only for genuinely mechanical jobs like renaming files.
Opus at medium effort handles routine, deterministic work. A LinkedIn post I build the same way every time does not need more than that.
Opus at high effort is for research and for leveling up skills. Opus hard is amazing, and honestly I underrated it for weeks by leaving it on medium.
The top model comes out only when the problem is genuinely hard.
The part people miss: the effort dial is part of the budget, not just the model name. Moving from medium to high changes both cost and quality, and switching that dial deliberately did more for my output than any model upgrade did.
What a heavy day actually looks like
The biggest day in the verified window was August 7, at 1.29 billion tokens. It was a Friday, and that is not a coincidence.
Friday is content day:Shorts scripts, including scraping competitor comments and X for topics
Blog work driven off Search Console opportunities, both new posts and enriching existing ones
The newsletterThe rest of the week is lighter: analytics reviews asking where we fell short, and tinkering to level up skills.
That tinkering compounds in a way that is easy to miss. My shorts skill started as "generate scripts." Then it found topics. Then it wrote in my voice. Now it edits the videos. I record 14 shorts in about 45 to 60 minutes, spend roughly 15 minutes editing, and then a polish skill runs for two to three hours autonomously.
Longform editing went from about two hours to 10 or 15 minutes through the same iteration loop. The current hook-animation work I would grade a C, and I expect to get it to an A the same way.
That is what the token count is actually buying. Not chat. Compounding tooling.The 60-second version, from my Shorts: My Real Claude Code Agentic OS Runs 6 Weekly Crons.What I would tell you to do with this
Do not use my numbers to justify the $200 plan. Read them the other way. A heavy daily user on the top tier sits at roughly 0.05% limit friction, which means I have headroom I am not using, and I still got walled once by bad habits. Habits move this number more than tiers do.
Do not read a big token count as a big workload. Ask for the cache split. Without it the number means nothing.
Fix context before you fix your plan. The fresh-session handoff was worth more to me than a tier upgrade would have been.
If you are choosing between the two Max steps, the measured version of that decision is Claude Max 5x vs 20x. If you are earlier in the ladder, Claude Pro vs Claude Max and Claude usage limits explained cover the mechanics I am measuring against here.
One last thing, and it is the reason I published my own logs instead of another opinion piece.
The only difference between AI slop and us is we actually have our experience. That is my rule, and it is why this page has numbers in it.
Published and last reviewed August 14, 2026. Verified figures come from 1,587 local Claude Code session logs streamed and deduped by request ID over June 11 to August 14, 2026. Lifetime figures are from Claude Code's own stats cache and are labeled as reported by the app throughout. Plan mechanics and API rates checked against claude.com/pricing that day. These are one account's numbers on Max 20x, not a benchmark.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: Claude's free plan gives you chat on every device, web search, memory across conversations, file creation with code execution, connectors, extended thinking, and the same 200k context window paid plans get. What it does not give you is a fixed number of messages, because that number does not exist.
Anthropic says so directly. Every site printing "about 40 messages a day" is guessing, and they all guess differently.
I am on Max at $200 a month, so I am not writing this from the free tier. I am writing it because the free tier is where I started, and because the advice around it is mostly invented.
What you actually get for $0
Straight from Claude's pricing page, checked today:Chat on web, iOS, Android, and desktop
Generate code and visualize data
Write, edit, and create content
Ability to search the web
Memory across conversations
Create files and execute code
Desktop extensions
Connect Slack and Google Workspace services
Connectors with remote MCP
Extended thinking for complex workAnthropic's own pricing page, captured August 14, 2026. Note what sits under Pro rather than Free: Claude Code, Cowork, Research, and "ability to use more Claude models."
That is a much thicker free tier than most articles describe. Memory, file creation, and connectors used to be paid features. They are not anymore.
What free does not include: Claude Code, Claude Cowork, Claude Design, Claude Science, Research, unlimited projects, and Claude for Microsoft 365. Anthropic lists Claude Code as included in all paid plans. If a tutorial tells you to run Claude Code on free, it is out of date.The 60-second version, from my Shorts: Three things people pay a pro for that Claude does free tonight.The number everyone prints is made up
Here is the actual state of page one. Four sites, four different answers, all stated as fact:Source
Claimed free limittruefoundry.com
"~40 short messages/day"gmelius.com
"roughly 30 to 100 messages per day"prompt.16x.engineer
"~40 short messages per day / ~45 short messages per 5 hours"nodemaven.com
"~50 to 100 messages per day"Now here is Anthropic, in its own pricing FAQ:"Every plan has usage limits that reset on a rolling five-hour session window, and paid plans add weekly limits on top. Your activity across Claude on web, desktop, mobile, and Claude Code all draws from the same pool. How much you can do depends on the length and complexity of your conversations, the model you choose, and the features you use, so there's no fixed message count. Free covers everyday questions."There it is: "there's no fixed message count."
That single sentence invalidates most of what ranks for this question. Anthropic does not sell you messages. It sells you a rolling five-hour budget, and how fast you spend it is up to you.
That is the Ship Lean read on Claude free, and it is the difference between feeling cheated and knowing what to change.
Anthropic also reserves room to move: "we may limit your usage in other ways, such as weekly and monthly caps or model and feature usage, at our discretion." So even the mechanic is not fixed.
You can see your own position anytime at Settings > Usage. It reports a percentage of a budget, not a message tally, which tells you everything about how this actually works.
What actually drains your budget
Four levers, straight from Anthropic's wording, in the order they usually bite:Length of the conversation. Every turn re-reads the whole thread. Message 40 in a long chat costs many times what message 2 cost.
Complexity of the request. A dense document analysis is not the same unit as "what's a good subject line."
The model you use. Heavier models cost more per turn.
The features you use. Extended thinking, file creation, and web search all add work.This is why the most common complaint on Reddit is not really a complaint about limits. It is a complaint about threads.
One user on r/ClaudeAI described it exactly: "i used to never exhaust my tokens, used to take me 1.5-2 hrs of conversation. but now, the tokens get exhausted within a single message???"
Nothing was nerfed. That single message was carrying an hour of history behind it.
The paid side shows the same physics. Another r/ClaudeAI user reported: "After my usage limits reset, I sent one prompt. Within 12 minutes, it ate 21% of my 5-hour limit." One prompt, a fifth of the window. Complexity, not count.
And the honest counterweight, from someone who actually ran a test on r/ClaudeAI: "I decided to test them limit. I got 100 conversations, each at least 3 messages long yesterday. I didn't get the usage warning."
One hundred conversations, no warning. That is the same free plan other people exhaust in an afternoon. The variable is not the plan.
Fix these three habits before you spend $20
This is what I would tell anyone who just got cut off mid-task. Do not open your wallet yet.
Be smart about your context. Do not dump everything into one chat window. Bring one specific use case per conversation. Context is your history of interaction on a topic, and it grows every turn. Ten messages deep, you are re-reading thousands of tokens on every single reply.
Yes, the context window is big. That is exactly why you burn through your usage faster.
Keep chats light. Start a new one when the task changes.
Choose down, not up. My rule wherever I have the choice: you can get away with Haiku for the small mechanical stuff and Sonnet for most real work. Do not reach for Opus on routine tasks. You will burn through your budget fast and often not get a better answer for it.
Fair warning on how far that goes on free: Anthropic lists "ability to use more Claude models" as a Pro upgrade, so your picker on the free plan is narrower than mine. The habit still matters the day you start paying, which is when it starts saving you real money.
Use free Claude for what it is best at. For me that is writing, plus code-related questions. Sign up for both free plans and split the job: Claude for writing, ChatGPT for general everyday research. Two free accounts cost nothing and cover more than one paid account at the same money.
What $20 actually adds
If you work through those habits and still hit walls on work that matters, here is what the money buys.Free
Pro ($20/mo, or $17 on annual)Usage
Everyday questions
At least 5x more per 5-hour sessionContext window
200k
200kClaude Code
No
YesClaude Cowork
No
YesProjects
Limited
UnlimitedResearch
No
YesModel choice
Narrower
"More Claude models"Fable 5
Not available
Usage creditsNote row two, because a lot of pages get it wrong: the context window is 200k on Free, Pro, Max 5x, and Max 20x alike. Free is not reading less of your document. It just cannot do as much before the meter stops you.
The four things that genuinely change your day are more usage, Claude Code, Claude Cowork, and access to more models. Everything else is convenience.
So is free enough for real work? No
I want to be straight about this rather than give you the comfortable answer.
If you have real workflows, depending on the free plan is a poor plan. Full stop.
For testing, absolutely yes. Test it on writing. A good email rewrite, yes. A few posts here and there, sure. That is a real and useful thing to have for $0, and I would tell anyone to start there.
But something serious? A batch of LinkedIn posts, a full document review, anything with real length? You are probably going to run out mid-task. And running out mid-task is worse than never starting, because now you are waiting on a reset with half a job done.
Here is how it actually went for me. I was on free, and I just found use cases. I would paste my problems in and get away with it for a few tries. After a while I felt the pain and thought, fine, I will pay the twenty bucks. Then I made that $20 last a long time.
Then Claude Code happened and everything changed. Way more use cases, building skills and agents, and the $20 plan stopped making sense. So I moved to $100. Eventually I hit that ceiling too, and moved to $200.Where that ladder ended for me. I am on Max 20x, which is why I can tell you the free tier is a starting line and not a destination.
Notice the pattern: every upgrade came after the wall, never before it. I never once paid because a comparison table told me to.
Start free. Let the wall tell you when to move.
What to read nextClaude usage limits explained if you want the full mechanics of the two meters
Is Claude Pro worth it? for the honest $20 verdict
Claude Pro vs ChatGPT Plus if you only have one $20 to spend
Claude Pro vs Claude Max if free already stopped being the questionPublished and last reviewed August 14, 2026. Free plan features, the 200k context parity, Pro and Max pricing, and the "no fixed message count" wording verified that day against claude.com/pricing, linked inline. Anthropic changes limits often and says so in its own terms, so check the source before you decide.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: yes, if you are on the Free or Go tier in one of nine countries. Plus, Pro, Business, Enterprise, and Education are ad-free. There are three documented ways out: pay for Plus or Pro, opt out on Free in exchange for fewer daily messages, or do nothing and know exactly what the ads can and cannot see.
There is no $2.99 ad-free plan. That number is Reddit folklore.
I pay for ChatGPT Pro and Claude Max, so I have never seen one of these ads myself. A lot of that spend is experimentation, because testing what actually stays in my workflow is the job. What follows is OpenAI's own documentation, checked today, plus my honest answer on whether ads should change what you pay for.
Who sees ads, by tier
Straight from OpenAI's announcement: "The test will be for logged-in adult users on the Free and Go subscription tiers. Plus, Pro, Business, Enterprise, and Education tiers will not have ads."Tier
Monthly
Ads?Free
$0
YesGo
$8
YesPlus
$20
NoPro
$100 / $200
NoBusiness
per seat
NoEnterprise
custom
NoEducation
custom
NoIf your employer gave you a ChatGPT account, it is a Business, Enterprise, or Education seat, and it does not show ads. That is the question most people at work are actually asking.
Two more carve-outs worth knowing. OpenAI does not show ads in accounts where the user is known or predicted to be under 18, and ads are "not eligible to appear near sensitive or regulated topics like health, mental health or politics."The 60-second version, from my Shorts: ChatGPT Just Started Showing You Ads and You Probably Did Not Notice.What changed on August 11: this is no longer a US-only test
Most pages ranking for this question describe a February pilot in the United States. That is out of date for a large share of readers.
The August 11, 2026 update reads: "ChatGPT Ads has now launched in the United Kingdom, Mexico, Brazil, Japan, and South Korea. We're continuing to expand to more markets this year."
The full rollout so far:Date
MarketsFebruary 9, 2026
United States (pilot begins)March 26, 2026
Canada, Australia, New ZealandAugust 11, 2026
United Kingdom, Mexico, Brazil, Japan, South KoreaSo if ads appeared for you this week and you are outside the US, nothing broke. Your country just got added.
The three ways out, and what each one actually costs
This is the table nobody prints. OpenAI states the exits in one sentence: "If you prefer not to see ads, you can upgrade to our Plus or Pro plans, or opt out of ads in the Free tier in exchange for fewer daily free messages."Your move
Cost
What you get
What you give upUpgrade to Plus
$20/mo
No ads, higher limits, better models
$20Upgrade to Pro
$100 or $200/mo
No ads, 5x or 20x Plus usage
A lot more moneyOpt out on Free
$0
No ads, still free
Fewer daily free messages, amount unpublishedDo nothing
$0
Everything you have now
Sponsored blocks under some answersNote the honest gap in row three: OpenAI does not publish how many fewer messages the opt-out costs you. Anyone quoting a specific number is guessing.
That gap matters more than it used to, and here is why.
The part nobody has connected yet: free got unlimited text five days before this
On August 6, 2026, OpenAI removed limits on text chats for all users, free accounts included, rolling out the week of August 10. Separate caps still apply to files, images, voice, and image generation.
The same announcement made GPT-5.6 Luna the default model for Free and Go, replacing GPT-5.5. Luna is the smallest model in the GPT-5.6 family.
Now read the opt-out sentence again. It trades ads for "fewer daily free messages" in a week where text messages on the free tier became unlimited.
OpenAI has not published how those two things interact. I am not going to pretend to know either. What I can tell you is that the free tier's constraint has quietly moved: it is no longer mainly about how many times you can type. It is about which model answers you, and about the caps that still bite on files, images, and voice.
The $2.99 myth, corrected
A widely-shared r/ChatGPT comment states: "Ads are displayed for users on the Free tier. Users can pay an additional $2.99 per month to remove these ads."
That price appears nowhere in OpenAI's announcement or help documentation.
The documented paid exit is Plus or Pro. If you are budgeting around a $2.99 ad-free upgrade, you are budgeting around something that does not exist.
The mechanism people do get right is the free opt-out. As one r/singularity commenter summarized it: "Or, you can reduce message limits to remove ads for free."
That one is real, and it is the exit almost no ranking page leads with.
What the ads can and cannot see
This is the second-biggest fear after cost, and OpenAI's wording is specific enough to quote rather than paraphrase.
What shapes which ad you see: "the topic of your conversation, your past chats, and past interactions with ads."
What advertisers get: "Advertisers do not have access to your chats, chat history, memories, or personal details. Advertisers only receive aggregate information about how their ads perform such as number of views or clicks."
So the targeting is real and it does use your chat history. The exposure is not: the advertiser sees a view count, not your conversation.
On the answers themselves, OpenAI states ads "do not influence the answers ChatGPT gives you" and are "always clearly labeled as sponsored and visually separated from the organic answer."
You control the rest in Settings. The ads controls panel carries ad history, interests, a one-tap delete for your ads data, and toggles for personalization. Per OpenAI's help documentation, deleting ads data "does not affect your chats," and may take up to 30 days to fully process.
Does this change my "buy ChatGPT first" advice? It makes it stronger
I am not going to force anyone to use AI. But this is the thing that is going to disrupt everything, so you might as well get started, and a great way to get started is ChatGPT free.
It is going to have ads. It is going to be a little annoying.
That is the point.
If you are using it enough that you get sick of seeing ads, or you run out of usage on the stuff that still has caps, that is your signal to spend the $20. You felt the pain first. Let the wall pick your plan, not the ads. That is my rule for free tiers, and it is the same rule I applied to myself on every upgrade I have made.
If you only have one $20 to spend, I still say buy ChatGPT before Claude, and I say that as someone who loves Claude and writes about it constantly.
Two reasons.
It is the generalist. ChatGPT plus the work features covers essentially everything a normal person needs, with a lot of usage and frequent resets. And the mobile app is better. I reach for ChatGPT on my phone far more than Claude.
Anthropic is currently the stingier one on usage. OpenAI resets limits often and is generous with them. Anthropic meters harder, and its free tier is thinner: no Claude Code, no Cowork, and the newest model is not available on it at all.
There is also a temporary boost on Anthropic's paid plans that is about to lapse on August 19, which only tightens things further. It never applied to the free tier anyway.
So no, ads do not change the call. They make it clearer.
Here is my one honest correction to my own instinct, though. I assumed the free tier was serving an older model and less usage. Free is now unlimited on text and runs Luna, which is a 2026-generation model. It is still the smallest one, and GPT-5.6 Sol is a huge step up from it, so the "you are not getting the best model" point holds. It just holds for a better reason than I thought.
What I would do this weekFree and annoyed by ads: open Settings, find the ads controls, and try the opt-out before you spend anything. It is free and it is documented.
Free and hitting caps on files, images, or voice: that is the real wall now. $20 fixes it.
On Go at $8: you are still in ad territory. The Go vs Plus math is the next thing to read.
On a work account: you will not see ads at all. Stop worrying about this one.
Deciding between vendors with one $20: ChatGPT first, and here is the full $20 comparison.If you want the bigger picture on how usage caps, not benchmarks, actually decide this whole category, that is Claude vs ChatGPT.
Published and last reviewed August 14, 2026. Tier scope, country rollout, targeting basis, and the three exits verified that day against OpenAI's own announcement page, linked inline. OpenAI does not publish a message count for the free opt-out, and this pilot changes often, so open the source before you decide.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using AI at your job without code.
Quick answer: one question decides it. Does this task need a file that lives on your laptop's hard drive? If no, fire it off from your phone and shut the lid. If yes, your desktop app has to stay open, and the laptop has to stay plugged in.
That second half is the part that bites. The task does not stop when your laptop sleeps or your battery dies, and it does not error out when it loses your files. It just quietly gets worse.
If you searched "claude cowork on phone," "how to use claude cowork mobile," or "claude cowork dispatch vs cloud," this page settles the decision first, then untangles the two different things both called "Cowork on your phone."
The one question that decides everything: does it touch a local file?
Everything else on this page is a footnote to this rule.
Anthropic's cross-device help article states it in one sentence: "A session in the cloud can read and write files in folders you've connected on your computer only while the desktop app is open on that computer. If the app is closed, the session keeps running but can't reach your local files."
Read that twice. The session keeps running. It does not stop, and it does not tell you.What the task needs
Laptop can be shut?
WhyEmail, calendar, Slack, Drive (connectors)
Yes
Cloud session reaches them over the internetWeb research and summarizing
Yes
Nothing local involvedWriting, drafting, restructuring text you paste in
Yes
The context is in the promptA folder on your hard drive (Downloads, a local vault, project files)
No
Needs the desktop app open on that machineAnything editing a video, a design file, or a local export
No
The file only exists on your machineLive artifacts and browser/computer use
No
Desktop-only per the help centerThe shorthand I already use for scheduled tasks holds here too: cloud in, cloud out; local files, local machine. That is my rule, and the surface you start from does not change it. I unpacked the scheduling half of it in Claude Cowork scheduled tasks; this page is the same rule applied to which device you are holding.The 60-second version, from my Shorts: 5 Claude Cowork connectors that do your boring work for you.Keep the laptop plugged in, or the run dies mid-task
This is my actual operational hack, and it is the one nobody writes down.
When I take work on the go, it is light tasks. The heavy stuff, editing my shorts and long-form videos, takes a long time and I am on it, so I let that run and step away. But I still take it with me, because Claude hits judgment calls.
It will stop and ask something like "Chris, you don't have the screenshot, do you want to use this," and then just sit there. Not running. Waiting.
That is why I go mobile at all. Not to do the work from my phone, but so I am not the bottleneck when it needs a decision. Anthropic frames the same thing in the Cowork web and mobile launch post: "When Claude reaches a call only you can make, it asks, and the question reaches your phone."
So here is the hack: have the computer on and plugged in.
If it fades away, if it goes to sleep, you are done. If the battery dies, you are done. Power is not a nice-to-have, it is the whole thing. My failure mode has never been getting back a confidently wrong answer. It has been the session simply not being alive when I came back to it.
And the mechanics back the hack up. Anthropic's Dispatch help article lists the requirements plainly: "Your computer must be awake and the app must be open for Claude to work on tasks." The setup flow even offers a toggle to keep your computer awake. If Anthropic ships a button for it, the problem is real.
One more trick worth stealing, because it is the thing that surprised me most: ask it for the file path. When it finishes a video render, I ask for the link and it hands me the path from the laptop. Click it, and I can watch the output on my phone. You do not need the file to be in the cloud to check the work.
The silent failure: your phone task looks finished and used none of your documents
This is the part the ranking pages skip, and it is the expensive one.
A normal software failure announces itself. This one does not. Your session keeps running, keeps its connectors, keeps its skills, and hands you a polished, confident, fully-formatted answer built on none of your actual files.
u/Secure_Sorbet_8671 named it in a PSA thread on r/ClaudeCowork: "The big one is local files. You connect folders on the desktop and a session only reaches them while the desktop app is open, so if your context lives in a local vault like mine, anything you fire off from the phone with the laptop shut has your connectors and skills but none of your files."
The interface does not save you either. u/Gh0stw0lf on the same subreddit put it bluntly: "They have projects on your computer that have local files be unable to do that and the UI doesn't clearly point that out."
So the defensive habit is simple and it costs you nothing.Before you shut the lid, ask yourself which folder this task will open. If you can name one, leave the desktop app running.
If you fire something off from the phone and the output feels generically correct, check whether it ever named one of your real files.
For anything pointed at a local folder, treat "laptop plugged in and app open" as part of the prompt.The trap is not that it breaks. The trap is that it succeeds at the wrong job.
Dispatch and cloud sessions are two different products, and only one survives the lid
Half the pages ranking for this query describe a world that ended on July 7, 2026. Here is the current split.Dispatch
Cowork cloud sessionWhere the work runs
Your own desktop
Anthropic's serversLaptop can be asleep?
No, must be awake with the app open
Yes, work continues in the backgroundPlans
Pro or Max
Max, Team, Enterprise; rolling out to ProThreads
One continuous thread, no way to start or manage a second
Normal sessions you start and resumeBest mental model
Remote control of your machine
A machine you do not ownDispatch is remote control. Per the Dispatch help article, it needs an awake computer, an open app, a Pro or Max plan, and an active internet connection on both devices. It also carries a limitation worth knowing before you build a habit on it: "There's no way to start a new thread or manage multiple threads." It is one long conversation you keep messaging.
Cloud sessions are the opposite. The Cowork launch post says it directly: "Work continues in the background. Close your laptop and Claude keeps going."
I had a hunch about this and asked to be fact-checked on it, so here is the honest scorecard: the hunch was right. With Dispatch you invoke a session on the machine and then walk away from it. Starting fresh from the phone gives you a new cloud session, which is genuinely a different thing, not a continuation of what is on your desk. That gap is exactly what bugs me about it, and it is why I have been leaning on Codex and GPT's remote work for on-the-go sessions lately. I do not have to invoke anything first. I can just say I want to start something new and start.
That is my preference, not a benchmark. If your work is knowledge work sitting in connectors, Cowork cloud sessions handle it fine and you never touch Dispatch. Anthropic's own line for the split, from the Dispatch doc: "Development tasks run in Claude Code; knowledge work runs in Cowork." If you are still deciding which cockpit you belong in, Claude Cowork vs Claude Code is the upstream decision.
If Cowork isn't on your phone, it's your plan, not you
There is a real cohort of people doing nothing wrong and seeing nothing.
u/tenaciousdweeb, on r/ClaudeCowork: "I have the Pro plan and I've tried everything and still no Cowork for me on mobile. I feel like I'm the only one who can't use it. I'm so frustrated."
That is the rollout, not a settings problem. Anthropic's help center, as of August 7, 2026: "Claude Cowork is in beta on web and mobile for Max, Team, and Enterprise plans, and will be rolling out to Pro plans over the next several weeks."
Three things to stop doing while you wait.Stop looking for a separate app. There isn't one. Cowork lives inside the regular Claude mobile app.
Stop expecting desktop sessions to appear on your phone. u/jaylan101 hit this: "Also my desktop cowork sessions don't show up on mobile so I don't see how pass off works." Resume behavior across surfaces is improving but is not a guarantee today.
Stop shopping for a second computer. The "buy a Mac Mini as your always-on host" advice still ranks and still gets repeated, and for cloud-eligible work it is now an expensive answer to a solved problem. You need an awake machine for local files. That is it.If you do have Dispatch on Pro but not mobile Cowork, that mismatch is documented, not a bug. Different features, different plan requirements.
The doubled limits ended August 5, so a phone habit costs more than it did last week
Worth knowing before you get comfortable firing tasks off all day.
Anthropic's launch post extended doubled Cowork usage limits "through August 5." That window closed two days before this was published. Whatever consumption felt like in late July is not what it feels like now.
This matters more on mobile than on desktop for a boring behavioral reason: phones make it trivially easy to fire off another task while waiting in line. Each one is agentic multi-step work, not a chat message.
The practical adjustment is not to use it less. It is to stop re-firing.
Ask for a status update on the session that is already running instead of starting a parallel one. If you want the full picture of what burns capacity and how the tiers actually behave, Claude usage limits explained covers it.
What Anthropic's own page promises, and what it quietly requires
Here is the pitch, straight from Anthropic's Cowork page, captured on July 9, 2026:"Steer from anywhere" is accurate. It is also doing a lot of work in that sentence.
Steering from anywhere is real: you can watch progress, answer the judgment calls, and redirect from your phone. u/strraand described exactly this on r/ClaudeCowork: "I can close the lid, go for a walk with my dog and still be able to follow the progress and adjust from my phone."
Working from anywhere is conditional. The condition is the local-file rule at the top of this page, and the small print for it lives in the help center, not on the marketing page.
What I would actually send from my phone first
Do not start with your most important task. Start with one where you would notice immediately if it went sideways.Best first phone task: something built entirely on connectors. Have it read your inbox and calendar and tell you what actually needs you today. Nothing local, so the lid can be shut the whole time.
The one I run most: a heavy job I started at the desk, that I then babysit from my phone so I can unblock it the second it hits a judgment call. Laptop open, plugged in, sleep off.
The trick worth copying: when it produces a file, ask for the path or link. You can open it on your phone and check the work without going back to the desk.
What I would not send from the phone: anything pointed at a local folder that I am not physically near. Not because it fails loudly, but because it doesn't.
Skip entirely for now: anything that depends on a live browser login or a desktop-only artifact. Those still want you at the machine.For the full menu of jobs worth handing over in the first place, 20 real Claude Cowork use cases has the list, and the connector-driven ones are your phone-safe candidates.
How this was checked
Product facts come from Anthropic's own pages, linked inline at each claim: the cross-device help article, the Dispatch help article, and the Cowork web and mobile launch post. User reports are public Reddit threads, linked where quoted.
The plugged-in rule, the judgment-call babysitting pattern, and the file-path trick are my own working habits, labeled as mine. I have not personally been handed a wrong answer by the local-file trap, so that section is built from Anthropic's documented behavior and named community reports, not from a story of mine. My failure mode has been the session dying with the machine, which is why the power hack is the part I actually enforce.
Things Anthropic does not document, so this page does not claim them: exact rollout dates by country or plan, what happens to an in-flight local-file task at the moment you close the app, and whether the doubled-limit promo affected anything beyond the Cowork limit.
Published and last reviewed August 7, 2026. Product facts checked against Anthropic's support center and launch post on that date. Cowork on web and mobile is in beta for Max, Team, and Enterprise with Pro rolling out; doubled Cowork usage limits ended August 5, 2026. These products change often, and the linked official pages are the source of truth.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: Anthropic does show you your Claude usage. Click your name, open Settings, then Usage, and you get two live meters. In Claude Code, type /usage. What Anthropic never publishes is a number, no messages per day, no tokens per plan, no table. So stop hunting for the figure that does not exist and read the two meters you already have.
Every site telling you "Pro gets 45 messages every 5 hours" made that up. I mean that literally. The number is not on Anthropic's pages, because Anthropic does not set your limit that way.
Here is the part that actually helps: once you stop looking for a number, the meters become useful, and four settings decide how fast those meters move.
The confusion, in one real thread
A guy on r/claude asked the most reasonable question in the world:"I know there is a concept called 'token,' but it's confusing for me. Again, I started just a few weeks ago; even the whole AI concept I recently started learning. Can someone please explain in simple words the concept of a token and also why my Claude is hitting its limits after just 2 messages?"That is u/ConfuseHead on r/claude, asking in plain words.
The top reply was a 400-word lecture on Byte Pair Encoding, embedding tables, and autoregressive sampling.
He asked why his account stopped working. He got a compression algorithm.
That gap is why this page exists. Everything ranking for this question is written for developers, in developer words, about Claude Code. If you are an HR manager who uploaded a policy PDF, asked two questions, and hit a wall, none of it is for you. This is.
Anthropic shows you your usage. It just never gives you a number
Both halves of that sentence matter, and most pages get one of them wrong.
It shows you. Two places, both free to check:In the Claude app: click your name in the bottom left, open Settings, then Usage.
In Claude Code: type /usage in the terminal.That is my own Settings screen on the Max plan, captured August 7, 2026. Usage is one row under Billing, right where you would look for it. It is not hidden, it is not a support ticket, it is two taps.
It never gives you a number. Anthropic's canonical page, how usage and length limits work, contains no message count, no token cap, and no per-plan table anywhere on it. What it says instead is this, verbatim: "Your usage is affected by several factors, including the length and complexity of your conversations, the features you use, which Claude model you're chatting with, and the effort level you've selected."
Read that list again, because it is the whole answer. Your limit is not a count of messages. It is a budget you spend at a variable rate.
That is also why two messages can eat your session. If each one re-reads a 60-page PDF and fires off web research, that is not two messages, that is two very expensive messages.
The Ship Lean rule: the meter is the number. Anthropic will not tell you how many messages you get, because there is no such thing as "a message" in their accounting. There is only how much you spent, and the meter shows you that live.
The two meters do two different jobs
Open Settings, then Usage, and you see two bars. Anthropic's own usage limit best practices page names them Current session and Weekly limits.Meter
What it measures
How it fails youCurrent session
How hard you are pushing right now
Snaps fast. One heavy afternoon and you are locked out until it recoversWeekly limits
Everything you have spent all week
Creeps. You never notice it until it is Wednesday and you are at 80 percentThe session meter is the one that interrupts you. The weekly meter is the one that ruins your week.
Nobody warns you about the second one. A user on r/ClaudeAI put it better than any documentation:"Omg I'm already at 20% weekly and my reset was yesterday, it creeps up on you while your working."That is u/InkedinSilver on r/ClaudeAI.
"It creeps up on you while you're working" is the weekly limit problem in eight words. The session bar is loud. The weekly bar is quiet, and it is the one that actually decides whether Friday works.
Two honest caveats so you do not get misled elsewhere. The "5-hour rolling window" everyone quotes is community convention, not something Anthropic states on its live pages, which describe a "Current session" bar without naming a duration. And the weekly reset schedule is contradicted by every third-party page that claims one, so I am not going to assert one either. Check your meter, do not trust the calendar math.The two meters on my account, September 5. Session bar on top, weekly bars below, and the reset times next to each. No message count anywhere, which is the whole point of this page.
One pool covers every Claude surface
This is the single most common surprise, and it is stated flatly in Anthropic's docs: "your usage of all different Claude product surfaces (claude.ai, Claude Code, Claude Desktop) counts towards the same usage limit."
One pool. Not one per app.
So the agent you left running in Claude Code is spending the same allowance you wanted for tonight's report in the browser. If you bought usage bundles, the same rule applies to those: that balance is a single pool across Claude, Desktop, Mobile, Claude Code, and Cowork.
If Claude "randomly" stopped working in the app, check what else you had running.
What burns the meters fastest, from someone watching them every week
I am on Max 20x, the $200 plan, and I watch both bars constantly. Here is what actually moves them, ranked by what I see:
1. Fable. By a distance. Fable is the thing I literally battle to not use every week. I have no trouble maxing out my Fable ceiling week over week, every week.
There is a documented reason for that, and it explains what most people experience as "Fable is broken." Per Anthropic's Fable 5 plan page, Fable 5 and Fable 5.1 “draw from your plan's regular weekly usage limits and use them faster than other Claude models.” They are capped at “up to 50% of your weekly usage limits” at no extra cost.
Anthropic even puts the obvious follow-up question in its own FAQ, “Will I get 50% more for my weekly limit for Fable models?”, and answers it: No.Anthropic's own plan page, captured August 7, 2026. This is the paragraph that explains why your week disappears faster than you expected.
So Fable is not a bonus allowance. It is a faster-draining straw into the same cup, with its own ceiling at half the cup. On Pro it is not included at all, it runs on usage credits, which is real money on top of your $20.
2. Opus, running everything. Opus is what I use 24/7 and I land in the 50 to 70 percent range most weeks. That covers everyday stuff, "hey check this," "help me optimize this listing," through to running my whole week of content, plus research.
3. Research and agents. When Opus spawns subagents to research something, that fans out. One button on your screen, many calls underneath. Research mode looks cheap and is not.
4. Long conversations. Anthropic notes that "longer conversations that trigger automatic context management consume more of your usage limit." Every message in a long thread carries the whole thread with it. Starting a new chat is free. Continuing a 90-message one is not.The 60-second version, from my Shorts: Your Claude usage isn't vanishing randomly. 4 quiet habits are eating it.Yes, it got stingier around August 5. Here is why
I noticed it maybe two weeks ago and I could not put my finger on it at first. Then I got the concrete moment: I started using Fable, glanced at the meter, and I was at 20 percent. In a single day. Not even straight through, just spread across a few hours.
Day one. Twenty percent. Six days left in the week.
That sucks, honestly. My reaction was immediate: okay, I gotta scale down a little bit.
There is at least one documented cause, and it is narrower than most people assume. The doubled Cowork usage limits ended August 5, 2026. Anthropic said so in its own Cowork launch post, which stated the doubled limits ran through August 5.
Read the scope carefully, because I am not going to overclaim it: what Anthropic documented is a Cowork promotion ending. If you leaned on Cowork, that bonus is gone and nothing about your plan changed. If you never touched Cowork, this is not your explanation, and I have not found a documented one.
I was not the only one. A thread on r/ClaudeCode titled "limits hitting fast for everyone" went up that same day and picked up 79 upvotes.
Now the part you need to separate out. That same week also had real outages. Anthropic's status page logged degraded performance for Claude Opus 5 on August 5, 13:51 to 14:34 UTC, plus a longer degraded-performance incident that morning, 07:05 to 14:14 UTC, covering Mythos 5, Fable 5, Opus 5, and Sonnet 5.
An outage is not a usage limit. They feel identical from your chair and they are completely different problems.
If Claude is slow, wrong, or erroring, check status.claude.com first. If Claude is refusing and pointing at your allowance, check your meter. Conflating the two is how people cancel a plan over a 40-minute incident.
One more thing worth knowing: errored responses still count against your usage. If a request fails and you retry, you paid for both.
The four things you can actually change today
Not vibes. These are the levers Anthropic documents on its own pages.
1. Lower the effort level. This is the highest-value setting nobody mentions, and Anthropic states the connection in plain English on its model and effort settings page: "Higher effort means more thorough responses, but they take longer and use more tokens, so you'll reach your usage limits faster."
The path, verbatim from that page: click the model name next to the send button, click "Effort," choose a level. The levels are Low, Medium, High, Extra high, and Max. Anthropic says Low and Medium "work well for routine tasks and stretch your usage further."
Two clicks. Reformatting a list does not need Max effort.The documented click path and the Opus 5 caveat, captured August 7, 2026, on the page almost nobody writing about Claude limits cites.
The extended-thinking caveat everyone else has wrong. Every page on this topic tells you to turn off extended thinking. As of today that is often impossible. Anthropic's own words: "Extended thinking cannot be turned off in Claude when using Claude Opus 5." If you are on Opus 5, that toggle is not available to you unless you switch models first. Advice written before late July is now actively wrong.
2. Pick the model for the job. My honest take by plan, and I am not going to pretend my setup is your setup:Your plan
What I would doMax 20x ($200)
Opus for everything. That is what I do. Watch Fable, it is the real drainMax 5x ($100)
Opus more carefully, Sonnet for everyday thingsPro ($20)
Sonnet way more than you think. Opus burns through a $20 allowance fast on most tasksFor the record: I run Opus 24/7 and I do not use Sonnet myself. My agents do, when Opus spawns subagents for research. Haiku, zero. That is a $200-plan luxury, not a recommendation for a $20 plan.
If you are still deciding which tier you need, is Claude Max worth it is the money version of this question, and Claude Pro vs Claude Max breaks down what changes between the tiers.
3. Use Projects instead of re-uploading the same file. This is the fix built for non-developers and almost nobody writes about it.
If you upload the same policy document, brand guide, or spreadsheet into a fresh chat every morning, you are paying for it every morning. Anthropic's docs: "When you upload documents to a project, they're cached for future use," and "Every time you reference that content, only new/uncached portions count against your limits."
Put the document in a Project once. Ask questions inside it forever.
While you are there: shorten your project instructions and delete project files you no longer use. Both are officially listed as ways to reduce usage.
4. Turn off tools and connectors you are not using. Anthropic recommends you "temporarily disable non-critical tools and connectors," including web search, Research, and MCP connectors, from the "Search and tools" settings. Its note on why, verbatim: "Tools and connectors are token-intensive."
If web search is on and you are just editing a paragraph, you are paying for a capability you are not using.
If you are blocked right now
Three options, in the order I would try them.Check status.claude.com. If it is an incident, no setting on your end fixes it and waiting is correct.
Start a fresh chat, drop the effort level, and switch to a lighter model. That is the cheapest recovery and it works on the session meter.
Buy a usage bundle, but only if the deadline is real. Pro, Max, and Team accounts can buy usage bundles at $50 for $45 of usage, $250 for $200, or $1,000 for $700. It is one pool across every Claude surface.My honest opinion on bundles: they are for a deadline, not for a habit. If you are buying them monthly, you have outgrown your plan and the upgrade is the cheaper answer.
And if you are on Free, know this going in: you do not get a meaningful meter. You find out you hit the limit when Claude stops mid-thought. That alone is a decent argument for the $20 plan.
What I would tell you if you asked me directly
Stop looking for your number. It does not exist, and the sites publishing one are inventing it.
Open Settings, then Usage. Look at the two bars. The session bar tells you how hard you are pushing this afternoon. The weekly bar tells you whether Thursday is going to work, and it is the one that will blindside you.
Then check the four levers: effort level, model, Projects, and connectors. In my experience the biggest single variable is which model you point at the task, and on any plan below $200 that means using Opus deliberately rather than by default.
The meter is the number. Everything else is guessing.
If Claude limits are what is making you compare plans in the first place, that is exactly the argument in Claude vs ChatGPT, where caps decide it rather than quality. For the full set of things Claude actually does in a normal work week, start at Claude at Work.
Published August 7, 2026. All Anthropic docs and status incidents cited here were checked on August 7, 2026. Usage policy changes often, so open the linked pages before you make a decision based on them.
Quick answer: Yes, if you make things for other people. No, if you don't. Max unlocks zero models that Claude Pro lacks. What $100 or $200 a month actually buys you is permission to stop rationing. I pay for Max 20x at $200 and it is worth it, but the reason is not "better Claude." The reason, by my own rough estimate, is about 14 hours a week I no longer spend editing.
Here is the problem with every page ranking for this question right now.
They are answering it for a developer burning tokens in Claude Code. If that isn't you, none of their math applies.I walk through this on camera in Every Dollar I Spend on AI Automation (Real Numbers) (10 min).The short answer, by who you areYou are
Do this
WhyNever paid for Claude at all
Pro, $20
Everyone should have the $20 plan. Scale from therePost online or run a business
Max is on the table
Content and client work generate enough volume to earn itAverage professional who doesn't post
Stay at $20
Resume work, job research, career planning fit inside ProHit the limit once last week
Wait it out
One wall is not a pattern. Let the pain talkHitting 50-60% of your weekly usage
Max 5x, $100
Now the meter agrees with youStopped mid-task daily on paid work
Max 20x, $200
Where I ended up, after monthsMax does not unlock a better Claude
This is the fact that should reframe the whole decision, and it comes straight off Anthropic's pricing page.
The model list is identical on every tier. Free, Pro, Max 5x, and Max 20x all list Fable, Opus, Sonnet, and Haiku.
There is no smarter Claude behind the $100 paywall.
What Max says it gives you is "5x or 20x more usage than Pro" and "higher output limits." That's capacity, not capability. Per Anthropic's Max plan page, Max 5x is $100 a month and Max 20x is $200, and the multiplier language there is worth reading carefully: it says more usage per session, not per month.
Max also carries two weekly ceilings, per that same page: one across all models, and a second one just for Sonnet.
So what does $100 actually buy? Permission to stop rationing
If the models are the same, the upgrade has to buy something else. It does, and it's this.
When Opus 5 shipped on July 24, 2026, Anthropic described it as "the new default model on Claude Max" and "the strongest model on Claude Pro."
Read that second half carefully, because most pages misread it. Pro users are not locked out of Opus 5. They get it. It just isn't their default, so they have to pick it on purpose, and it eats their allowance faster when they do.
That is the entire non-coder upgrade, in one sentence: on Pro you budget which model you deserve, on Max you stop thinking about it. That's the Ship Lean read on the Max question.
The clearest version of this I've seen from an actual non-developer came from r/ClaudeAI:"I'm currently trying out a max 5x plan and I don't code. I use Claude for analysis of meeting transcripts, writing documentation, and helping refine copywriting / marketing. ... The max plan lets me use Opus without needing to weigh the limit tradeoffs with sonnet."u/redbulb on r/ClaudeAI, May 2025Date that one honestly: it's from 2025 and refers to an older Opus, so treat it as evidence about the reason people upgrade, not as testimony about the current model. The reason has only gotten stronger since Opus 5 became Max's default.
The Fable 5 mechanic almost everyone gets wrong, including me
I want to correct something I believed until I checked the docs for this post.
I had assumed Fable 5 came with its own dedicated pool on Max, separate from everything else, because it certainly behaves like a different animal. It does not.
Per Anthropic's Fable 5 plan page, on Max, Fable 5 and Fable 5.1 “draw from your plan's regular weekly usage limits and use them faster than other Claude models.” What you actually get is a ceiling inside the same pool: you can use “up to 50% of your weekly usage limits on Fable models at no extra cost.”
And the FAQ on that page answers “Will I get 50% more for my weekly limit for Fable models?” with one word: No.
So it's a cap, not an allowance.
My instinct was half right, which is the useful half: Fable 5 is metered differently from every other model, because it's the only one with its own ceiling. And it does burn through your week faster than anything else. I max out my Fable ceiling almost every week. That's what a 50% cap inside one shared pool feels like from the inside, and it's why "I have Max so I never think about limits" is not quite true even at $200.
What actually justifies my Max plan: about 14 hours a week
Here is my honest answer, and it has nothing to do with code.
I edit a lot of video. Roughly 14 shorts a week.
It used to be just cuts. As someone who is not a professional editor, that alone was revolutionary to me, because cuts were all I knew how to do. Then you start watching your competitors and you notice they're doing something you're not: visuals.
Now Claude does the visuals too. Screenshots, real GUI captures, actual ChatGPT and Claude interfaces dropped into the edit.
Picture doing that by hand. Placing visuals one by one across a short is easily a couple of hours of work; let's round it down to one hour per short to be conservative. Fourteen shorts, one hour each, cutting and adding visuals. That's about 14 hours a week I no longer spend.
To be precise about that number: it's my own estimate of what the manual version would cost me, not a stopwatch measurement. I'm telling you how I did the math so you can redo it with your own numbers.
And that time doesn't turn into free time. It turns into more content, which is the actual point.
I could not do that without my Claude skills and the Max plan. So for me the answer is yes, a hundred percent.My billing screen. Max 20x, $200 a month. Captured on my own account.
What my usage actually looks like in a normal week
Since nobody on page one will show you a real number, here's mine.
On Opus and general work I run about 50 to 70 percent of my weekly allowance. Fable is the outlier: I'm maxing that ceiling out almost every week lately.
Am I always hitting my limit? Most of the time, yes. Not always. And I'm okay with that.
That's an important thing to say out loud, because "Max means never hitting limits" is a myth, and a $200-a-month user hitting the wall collected 1,522 upvotes for saying so. The upgrade doesn't remove the ceiling. It moves it far enough away that you stop planning your day around it.My meters on a normal Saturday: half the weekly allowance gone two days after the Thursday reset, and a third of the current session. That is the rationing I pay not to think about.
Do this before you spend $960 a year
Not one page ranking for this query tells a writing-first user to try the free fixes first. So here they are.Check Settings > Usage for two weeks. This is the whole decision, and it's free. If you never cross 40%, you don't have a volume problem. If you're pinned at 50-60% and getting stopped, you do.
Lower the effort level. Click the model name next to the send button, then click "Effort." Anthropic documents five levels: Low, Medium, High, Extra high, and Max. Its own wording: "Higher effort means more thorough responses, but they take longer and use more tokens, so you'll reach your usage limits faster." Low and Medium "stretch your usage further."
Know the Opus 5 caveat. Same page, verbatim: "Extended thinking cannot be turned off in Claude when using Claude Opus 5." So if you're on Opus 5 and trying to conserve, effort level and model choice are your levers. Turning thinking off is not one of them.
Put reference material in a Project so it's cached instead of re-pasted into a fresh chat every time.And know this before you trust anyone's numbers, including the ones in the SERP above this page: Anthropic publishes no concrete limits. No message counts, no hours, no token caps per plan. Its help page says usage depends on conversation length and complexity, which features you use, which model, and the effort level you picked.
Every "you get X messages per 5 hours on Max" figure on page one is invented.
The likelier mistake is buying too much plan
If you're a non-coder, the failure mode isn't underbuying. It's paying $100 for headroom you never touch.
The community evidence on this is unusually consistent. A daily Max user on r/ClaudeAI:"I have Claude Max right now, and I use it every single day for coding, for also working with text for my work: research, writing, rewriting emails, a reflection coach, and more. I get to about 25% of my weekly limit... It's crazy, but I might actually try going down to Pro because I literally never get above 30% of my weekly limit and I use it every day."u/RetroUnlocked on r/ClaudeAI, March 2026That's someone who codes and writes, using it daily, at a quarter of the plan.
The same thread surfaces the structural problem behind the overbuy: there's nothing between $20 and $100. u/OrbMan99 put it plainly: "The jump to the next tier is too high... Something like a $50 tier would probably bring a lot of those users back." That thread isn't an outlier either; a separate ask for a middle plan pulled 121 upvotes.
There is no $50 rung. So the honest options are Pro plus occasional usage credits, or the full jump.
One option nobody ranking for this query mentions: if you're using Claude for your actual job, a Team seat starts at $20-25 per seat per month on Anthropic's pricing page. Ask your employer before you put $960 a year on your own card.
Who should actually buy it
Here's my segmentation, and I'll be blunt about where the line falls.
Anybody who creates content online: yes, hands down. Multiple use cases, and the volume is what earns the plan back.
If you have a business: yes, hands down. Same reason.
If you're the average person who doesn't post anything online: stay at $20. And I mean that as a real recommendation, not a consolation prize. Pro handles resume optimization, job research, drafting your career goals. Cowork works on the $20 plan too. That's where it genuinely shines for a normal professional, and it's plenty.
Where it gets interesting is when you start building custom things: check my email every day, give me the news on these specific topics. Once you adopt it around your life it starts solving different problems, and that's usually when the usage curve bends.
I'd also gently push most people toward the content side. It's very low maintenance now. Two or three LinkedIn posts a week sharing what you know from the career you already have, or a Substack. You don't need to become a creator to benefit from creating.
Everyone should have the $20 plan. That's my honest opinion, and then you scale from there.
The decision rule: let the pain decide
If you hit your limit once last week and you're staring at the upgrade button, here's exactly what I'd tell you.
If you can wait, wait it out.
One wall is not a pattern. It's a Tuesday.
If you want to try it anyway, fine. Upgrade, then watch your usage. Are you actually hitting the limit? Are you maxing at least 50 to 60% of it? That's the number that answers this, and it's the only one that's yours.
Nobody can tell you whether $100 is worth it, because the answer depends entirely on what waiting costs you. So you have to know your usage. Let the pain decide. That's my rule for this upgrade, and it's the same rule I used on every tier I climbed.
The meter is in Settings > Usage. Go look before you spend.
Next forks worth readingClaude Pro vs Claude Max - the full tier-by-tier spec comparison if you want the mechanics
Is Claude Pro worth it? - the $20 version of this exact question
Claude Max vs ChatGPT Pro - if you're comparing the $100 tiers across vendors
Claude usage limits explained - why you got stopped when the meter said you had roomReady to put whichever plan you pick to work on real tasks? Start at Claude at Work.
Published August 7, 2026, and last reviewed that day. Pricing, tier multipliers, the Fable 5 ceiling, Opus 5's plan placement, and the effort-level settings were all checked against claude.com/pricing and Anthropic's own support documentation on the publish date, linked inline throughout. Anthropic changes these often, so open the source pages before you buy.
Quick answer: If your employer runs Google Workspace, you already have Gemini in your Gmail and Docs whether you chose it or not. So the real question is almost never "which one is smarter." It is "do I need to pay for a second AI on top of the one I already own?"
My rule: yes, a second subscription is worth it - but only once you can name the specific task that the tool you already have keeps failing at.
If you cannot name that task, you are shopping, not solving.
Where I am honest about what I have and have not used
I want to set expectations before you read a comparison from me.
Full disclosure on my own setup, because it is two different things and I do not want to blur them. I pay for the Gemini app through a family plan - and honestly, my wife uses it more than I do. Separately, my real heavy Gemini use was never the app at all: it was the API, running Nano Banana for image generation. Then I moved all of that over to OpenAI's ImageGen, and everything I ship now - Substack images, thumbnails, LinkedIn visuals, blog headers - runs through ImageGen inside my subscription instead of paid API calls.
What I have not done is use the Gemini app as a daily driver for real work recently. I have tried it - it is not that I could not. I just have so much abundance with Claude, ChatGPT, and Codex that I have no reason to reach for it.
So I am not going to tell you Gemini is worse at reasoning or better at research, because I have not earned that opinion. What I can give you is the decision framework, the actual current pricing on both sides, and what people who did switch report - which is more useful than another benchmark table anyway.
On images specifically, I will say this: Nano Banana is a sleeper. If you are working through the API, it is a genuinely top-tier option and it is roughly three times cheaper than ImageGen on average. Where it loses is fine detail - text rendering especially. ImageGen wins on quality in most angles I care about, which is why I migrated.
The tier ladders do not line up (and most pages get this wrong)Nearly every comparison article shows a tidy free/$20/$100 grid for both. That grid is wrong on both sides.Google AI
ChatGPTFree
$0, 15 GB storage
$0, may show adsCheap tier
AI Plus, $4.99/mo (2x free limits, 400 GB)
Go, $8/mo (US)Standard
AI Pro, $19.99/mo (4x free limits, 5 TB)
Plus, $20/moPower tier
AI Ultra, $99.99 or $199.99/mo
Pro, $100 or $200/moSources: Google's subscription page and OpenAI's Plus and Pro tier help pages, all checked July 31, 2026.
Two things worth pulling out.
There is a $4.99 Google tier sitting between free and $19.99. If your budget is genuinely tight, that is a real option and almost nobody mentions it.
And neither company offers annual billing on these plans. OpenAI states plainly that it does not support annual billing or paying for multiple months in advance.
One caveat on the ChatGPT side: ads are coming, not shipped. OpenAI's language is about testing ads on Free and Go in future. Do not buy or avoid anything today based on ads you cannot see yet.
You may already own Gemini without having chosen it
This is the part that changes the math for most people reading this.
If your company runs Google Workspace, Gemini is already in your Gmail, Docs, and Drive. If you carry an Android phone, it is there too. You did not evaluate it, you did not pick it, and you may be paying for a competitor to do things the bundled one already does.
That bundling is also, for a lot of people, the actual reason they keep using it:"I like Gemini more just because it's integrated into my Google Home, my Rivian, my Gmail, and apparently now in Chrome. The integration for me is what makes me keep using it."u/digitalcleavage on r/OpenAIIntegration is a real feature. It is not the same as being the better model, and it is worth being clear with yourself about which one is actually keeping you there.The 60-second version, from my Shorts: 6 FREE Google AI Tools Most People Have No Idea Exist.What switching actually costs
The loudest complaint from people who move is not about intelligence. It is about friction."I want to love Gemini mainly because I just got premium free for a year, but I've been super disappointed with it so far. For me, it's way slower than chatGPT even on 2.5 flash. And every time I ask it something I need to re-establish the context."u/Fit_Presence8008 on r/GoogleGeminiAIThat is the tax nobody prices in. Your saved chats, your custom instructions, your projects, and the accumulated context about how you write and what you are working on do not transfer between vendors.
If you have six months of history in one tool, switching does not start you at zero. It starts you below zero, because you now have to rebuild what you had while doing your actual job.
Which is a strong argument for adding rather than switching, if you can afford it.
The budget version: when you genuinely cannot pay for both
Plenty of people are in this position and the comparison pages ignore them entirely."I started a youtube channel for relaxation videos. Both ChatGPT and Gemini have help with creating those videos. ChatGPT is better for creating images and Gemini is best for generating my videos. Im stuck between the 2. The problem, I can't pay for both."u/Quirky-Background-80 on r/OpenAIIf that is you, work it in this order:Use the bundled one for everything it handles. If work gave you Gemini, make it your default for drafting, summarizing, and anything living inside Gmail or Docs. It costs you nothing.
Write down what it fails at for two weeks. Not vibes - actual tasks where you gave up or the output was unusable.
If that list has a repeating item, buy the other tool for that item. If the list is empty, you just saved $20 a month.
Consider the cheap tiers first. Google AI Plus at $4.99 or ChatGPT Go at $8 may clear your blocker without a $20 commitment.My rule for whether a second AI subscription earns its place
I pay for several of these, so let me be direct about how I justify it rather than pretending everyone should.
A second subscription is worth it when it removes a bottleneck you hit repeatedly, not when it is marginally better at something you do occasionally. The image-generation move I described earlier is exactly that shape: I was paying API costs per image, the work was constant, and moving it inside a subscription I already had made it effectively free at the point of use. That is a bottleneck being removed.
Compare that to "the other model scores better on benchmarks." That has never once saved me an hour.
So: if you cannot name the task, do not buy the tool. That is my rule for stacking AI subscriptions, and it applies just as well to Gemini, ChatGPT, or anything that launches next month.
Worth saying plainly since this is a fast-moving space: the frontier moves constantly. Do not treat any single comparison, including this one, as settled for longer than a few months.
How to decide this weekWork gave you Gemini, you have no specific complaint: do not pay for anything. Use what you have.
Work gave you Gemini and you keep hitting the same wall: buy the cheapest tier of the other tool that clears it. Go at $8 before Plus at $20.
You pay for ChatGPT and you are considering switching to Gemini: do not switch, and do not add either, until you have written down what is actually failing. Switching costs you your context.
You use AI seriously every day for work that pays: owning both is defensible. That is what I do, and the reason is coverage, not loyalty.
You are on the tightest budget: Google AI Plus at $4.99 is the cheapest real upgrade on this page, and almost nobody mentions it.Related decisionsClaude vs ChatGPT - the comparison I can speak to from daily use on both sides
Copilot vs ChatGPT - the same "my employer already gave me one" problem, Microsoft edition
ChatGPT Go vs Plus - what the $8 tier does and does not buy
Claude Pro vs ChatGPT Plus - the $20-versus-$20 forkIf you want to put whichever tool you land on to work on real tasks, start at Claude at Work.
Published July 31, 2026. Pricing for both vendors checked that day against gemini.google/subscriptions and OpenAI's official help pages, linked inline. Model versions on both sides move quickly; the pricing ladders and the decision rules are the durable parts of this page.
Quick answer: Yes, $20 Claude is worth it, hands down. But I am going to give you the answer most people reviewing this will not: if you are just the average person who wants one AI subscription for general work, start with ChatGPT instead. Claude is my daily driver and I still say that.
The reason is not quality. It is fit.
Two things almost every "is Claude Pro worth it" article on page one is currently wrong about: at $20 you no longer get Claude's newest model as part of the plan, and as of September 2026 the $20 ChatGPT tier does not get OpenAI's newest model in chat either. Both facts are below, from both companies' own pages.
Neither one changes my answer. Both change what you should expect for your money.
The short answer, by who you areYou are
Buy this
WhyGeneralist, want one AI for everyday work
ChatGPT Plus
More general, more use cases, easier first winCreative, writing, design, building things
Claude Pro
Better at design and at writing in your voiceTinkering, no strong use case yet
Neither, use both free tiers
There is no wrong answer at $0. Find the use case firstA teacher
Check Claude for Teachers
Verified educators get it free, not $20Doing real work daily on AI
Both, $40 total
Two separate allowances beat one bigger oneWhat changed on July 20, and why every ranking page is stale on itHere is the fact that reframes this whole question.
Until July 19, 2026, Anthropic included Fable 5, the newest model, in Pro's weekly usage limits as a promotion. That ended at 11:59:59 PM PT on July 19. Per Anthropic's Fable 5 plan page, the state today is:Free: “Fable 5 isn’t available on the Free plan.” Fable 5.1 is also limited to paid plans.
Pro ($20): “Fable 5 and Fable 5.1 aren’t included in your plan's usage limits. You can use them with usage credits.”
Max ($100+): included as standard, up to 50% of weekly limits at no extra cost.So $20 no longer buys you the newest Claude as part of the plan. You can still use it, you just pay per use on top.
Does that kill the value? No. But it means "Claude Pro gets you the best model" - a line still sitting in most ranking articles and in Google's own AI summaries - is now simply false, and you should not buy on that basis.
Re-verified against claude.com/pricing on September 4, 2026. Nothing moved:The September update: $20 does not buy the newest model on either side now
Claude putting its newest model behind credits sounds like a knock on Claude until you check what $20 of ChatGPT buys this month.
OpenAI shipped GPT-6 Astra on September 3, 2026. And the $20 ChatGPT Plus tier is the highest paid tier that does not get it in the chat model picker.
Per OpenAI's help center, in chat the model is branded GPT-6 Pro and it is listed for Pro $100, Pro $200, Business and Enterprise. Plus gets Astra in ChatGPT Work and Codex as it rolls out, not the chat box. Free and Go, the tiers below, get neither.
The mechanism is visible in the same article's plan table. The chat picker's reasoning levels run Instant, Medium, High, Extra High and Pro, and GPT-6 Pro sits under Pro. Plus does not have the Pro option:That table is about GPT-5.6 Sol reasoning levels, but it is the same door.
So both $20 tiers now gate their newest model. Different mechanics, same outcome:At $20
Newest model
How it is gatedClaude Pro
Fable 5.1
Available, but on separate usage credits rather than your included limitsChatGPT Plus
GPT-6 Astra
Available, but in Work and Codex rather than the chat pickerAt $20 you are no longer buying the newest model on either side. That is the honest 2026 read, and it is not a reason to skip either plan. It is a reason to stop choosing on which one has the shiniest model, because at this price neither one really hands it to you.
What $20 actually buys is a very good workhorse model plus the tools around it. Pick on the tools.
So is it worth $20? Yes, and here is the honest split
I will not hedge: yes, it is worth it. Claude is my baby.
What it is genuinely better at is design. Claude Design is included with the $20 subscription, and it is what I used to design my own website. That is the attestation I can give you - this site feels like a pro designer helped me build it, and Claude Design did it. It is also better at writing in your voice, and better at creative work generally.Sonnet 5 is powerful. Opus 5 is more powerful still. At the bare minimum you should be using the free tiers of both Claude and ChatGPT before you spend anything.
But here is the part that costs me nothing to admit.
For the more general user, I still push people to ChatGPT. It is more practical for general use, and on desktop the interface is a little less daunting. Even Codex, on the OpenAI side, feels less intimidating than Claude Code does to a newcomer.
So the honest question is not "which is better." It is: if you are creative, Claude will write and design better for you. If you are a generalist, ChatGPT will find more jobs to do.
That is my split on this question, and it is why I pay for both.
How I actually got here, because nobody answers this honestly
Every page on this question tries to give you one answer. There isn't one. Here is my real path, and I think it is the useful thing on this page.
I started on Claude free, about two years ago, using it while building n8n workflows because it was free and I figured why not.
Then I noticed something. I would ask ChatGPT, then ask Sonnet the same thing, and I could feel the difference. Smarter answer. Fewer errors to clean up.
So I kept doing that, and I kept hitting the free limit.
The day I paid was not a research day. It was a day I hit the wall, looked at waiting five hours, and thought: I am not waiting. I will pay the $20 and cancel it later.
I did cancel it, at one point. Then I paid again, because I kept reaching for it. Eventually I was paying for two subscriptions, and now I am at the top tier.
Nobody plans that. You use it as you see fit and you upgrade as you see fit. It is a gateway, one step at a time, and every step gets made by hitting a wall rather than reading a comparison.
Which is why my advice is boring and correct: do not buy Claude Pro because an article told you it is worth it. Use the free tier on real work until it stops you. If it never stops you, you just saved $200 to $240 a year depending on how you would have billed it. If it stops you constantly, you already have your answer and you will not need convincing.
What the free plan actually includes (the surprises)
Most articles get this backwards in both directions. Here is the verified matrix.Feature
Free
Pro ($20)Projects
Yes, capped at 5
Yes, unlimitedProject knowledge / RAG
No
YesMemory across conversations
Yes
YesSearch past chats
No
YesWeb search
Yes, counts against limits
YesClaude Code
No
YesCowork
No
YesClaude Design
No
YesFable 5 and Fable 5.1
No
Credits only, not includedTwo of those genuinely surprise people. Projects work on the free plan, capped at five - a lot of writers assume that is a paid feature and never touch it. And searching your own past chats is not free, which catches people the other way.
One honest note on the free tier: Anthropic's own docs contradict each other about whether free limits are five-hour sessions or daily. The get-started page says a session-based limit resetting every five hours with message counts that vary by demand. If you have found this confusing, it is not you.
How to make the free plan last longer
Before you pay, try these. They come from Anthropic's own usage guidance, not from me.Put reference material in a Project. Content in projects is cached and does not count against your limits when reused. Re-pasting the same background doc into fresh chats all week is the single most common way people burn a free plan.
Start a new chat for a new task. Very long conversations get more expensive as they grow, because the whole thread rides along.
Watch your attachments. Long documents and pasted URLs cost far more than short questions.If you do all three and still run out of road, that is real signal rather than guesswork.The 60-second version, from my Shorts: The moment Claude says 'you've reached your limit': do these 5 things.Where the free plan actually runs out
This is the specific answer to "when does paying earn its money back."
The free tier's limitations are: less usage, not the strongest model, and no access to the products that do actual work. You cannot use Projects properly with knowledge attached, you cannot use Cowork, and you cannot use Claude Code.
Free is good for free. I think it is smart that Anthropic keeps it that way - you get a peek. But it is a peek.
The moment paying earns its money back is when you stop asking Claude questions and start handing it work. That is the real line. Chatting fits inside free. Delegating does not.
And here is the honest warning that belongs in the same breath: Cowork is the best reason to pay and it burns your allowance fastest. Anthropic says plainly that working on tasks consumes more of your allocation than chatting. The feature that sells you the upgrade is also the one that runs you into the ceiling.
What real Pro users say
The most useful description of who this plan is for came from a lawyer on r/claude, not from a reviewer."I'm going to speak about my own experience using Pro. I am not a coder, nor is every Claude user, so bear that in mind... Not everyone needs a lot of usage; not everyone needs Fable all of the time. Now, if you need all that horsepower, then pay for more, but in my eyes, the Pro subscription is for people like me, not a heavyweight user that vibecodes with Fable 24/7."u/Lacoalfredo1 on r/claudeThat is the correct expectation to buy with.
The counterweight, and the reason to read the limits section above carefully, is what happens when you buy with the wrong one:"Yesterday I subscribed to Claude Pro ($20/month)... I worked on a WordPress plugin for 1 hour last night and 1 hour this morning... I just got the 'You've reached your limit' message. Two hours of actual work for 20 bucks?"u/kenaddams42 on r/ClaudeAIBoth experiences are real. The difference between them is not the plan. It is what each person expected to do with it.
Who should not payTeachers. Anthropic offers Claude for Teachers free to verified educators, with signup open through June 30, 2027. Check that before spending $20.
Anyone still tinkering. If you do not have a use case yet, both free tiers are enough. There is genuinely no wrong starting point.
Anyone buying it for the newest model. After July 20 that is not what $20 buys. Buy Pro for capacity, design, and Cowork.
Generalists who only want one subscription. Buy ChatGPT first. Come back to Claude when your work turns creative.The one-paragraph version
Claude Pro at $20 is worth it if you write, design, or build - it is the better creative tool and Claude Design alone can justify it. It is not worth it if you are buying it for Fable 5 or Fable 5.1, because neither is included in the plan's usage limits and both run on paid credits. As of September 2026 the $20 ChatGPT tier gates its newest model too, so stop shopping on model names at this price. And if you are a generalist buying your first AI subscription, get ChatGPT instead, then add Claude when your work asks for it.
Next forks worth reading:Claude Pro vs Claude Max - if you are already hitting limits at $20
Claude Pro vs ChatGPT Plus - the direct $20-versus-$20 comparison
Claude vs ChatGPT - the full picture, and why caps decide it
Claude Cowork use cases - what the paid-only feature actually doesReady to put it to work? Start at Claude at Work.
Published July 31, 2026. Last reviewed and updated September 4, 2026: re-verified the Fable credit rules against claude.com/pricing, added the September section showing that ChatGPT Plus now gates GPT-6 Astra out of the chat picker too, added my own free-to-paid path, and rewrote the title and description to lead with the verdict instead of a curiosity hook. Plan features and free-tier limits checked against claude.com/pricing, Anthropic's support documentation, and OpenAI's help center, linked inline. These change often.
Quick answer: Same models, different ceilings, and one real feature gate. Max 5x is $100 a month and Max 20x is $200. I am on the 20x tier. You do not buy Max because it is better. You buy it when waiting has started costing you more than the money does.
And you should know before you spend $200 that the "20x" on the tin is the subject of a proposed class action.
That is the part this page was missing, so let us start there.
The short answer, by who you areYou are
Get this
WhyCurious, just starting out
Pro, $20
Nobody should start at $200. You will not know what you need yetBrainstorming, drafting scripts, light Cowork
Pro, $20
This genuinely fits, if your process is dialled inHitting limits occasionally, can wait it out
Stay on Pro
If waiting is annoying but not expensive, you are not readyStopped mid-task on work that has a deadline
Max 5x, $100
The wait is now costing you real timeExperimenting constantly, building many skills
Max 20x, $200
Where I ended up, and only after a long climbWant Fable without credit top-ups
Max
This is the one true feature differenceAlready decided on Max and choosing between the two steps? That is Claude Max 5x vs 20x. Comparing across vendors instead? That is Claude Max vs ChatGPT Pro.
The 20x problem
Anthropic markets Max 20x as "20x more usage than Pro per 5-hour session." There is real, sustained pushback on whether that is what people actually receive, and it has now gone further than forum complaints.Filed June 14, 2026 and reported that week. Coverage by Quartz, which credits the original reporting to the Wall Street Journal, and by Engadget. These are allegations and nothing has been decided.
A proposed class action, Kahn v. Anthropic, was filed on June 14, 2026 in the Northern District of California on behalf of Washington D.C. resident Karl Kahn, targeting both Max tiers. The complaint states:"The actual usage provided by the Max 5x and Max 20x plans is far below the advertised amount of usage."The detail that made it concrete: after moving up to Max 20x, Kahn alleges a single five-hour work session consumed 15% of his weekly allotment. Anthropic declined the Journal's request for comment, and has not responded publicly since.
Now let me be careful, because this matters.
Nothing here is proven. A filed complaint is one side's account. Anthropic has not responded publicly and no court has ruled. I am not telling you the 20x claim is false, and I am not in a position to know.
What I can tell you is my own experience, and it is unremarkable: on the 20x tier I get more usage. It works. But my allowance feels like it drains faster than Codex does, and that is a feel, not a measurement, because Anthropic publishes no number I could measure against.
That is the actual grievance, and I think it is a fair one. The multiplier is a session-usage ratio, not a capability promise and not a queue position.
My own read, and it is only that: it was worded in a way that let people believe they were buying something more definite than they were. You are allowed to be annoyed about that as a consumer, whatever the court eventually decides.
So what does the complaint say the real numbers are?
Per reporting on the filing, it alleges Max 5x delivers roughly three and a half times Pro rather than five, and Max 20x lands somewhere between six and eight times rather than twenty. Kahn also points to Anthropic's own July 2025 communications about per-model weekly caps, which he says contradict the public marketing.
Again: alleged, not established.
One thing worth flagging against my own instinct. My rough recollection had been that the real figure was somewhere around 10x to 15x. The numbers in the complaint are lower than that, not higher, so if anything I had been giving Anthropic the benefit of the doubt.
You can see why this lands. The most Nico-shaped comment I found on the whole question came from u/needusbukunde on r/ClaudeAI, who had paid and still did not know what he had bought:"I just upgraded to Claude Pro for $200/year, upfront, and my service has actually gotten worse. My sessions get 'timed out' more frequently... What is 5 x account mean?"That, to me, is the question the marketing should have answered up front.
The two ceilings, and why you are not doing it wrong
Here is the single most useful mechanical thing to understand, and it is the reason so many people think they broke something.
Claude runs two separate limits at the same time. There is a rolling five-hour session cap, and there is a weekly cap.
You can be nowhere near your weekly limit and still get stopped by the session cap.
That is why lockouts feel random and unearned. They are not random. You hit whichever ceiling arrived first, and Anthropic does not publish a message count you could have budgeted against. The multipliers are all they state.
One more piece: everything shares one pool. Chat, Claude Code, and Cowork all draw from the same allowance, and Cowork specifically burns it faster because multi-step tasks are compute-intensive. Spend an afternoon on an agent task and you have also spent your chat capacity.The two ceilings on my own account, mid-Saturday: the 5-hour session bar and the two weekly bars. This is what "20x" looks like in practice: bars, not a message count.
The one real feature gate: Fable
This is the most under-reported fact in the comparison, and it is the only place the tiers stop being the same product at different sizes.
Until July 19, 2026, Fable was included in Pro's weekly usage limits as a promotion. That ended at 11:59:59 PM PT on July 19. Here is where the tiers landed, per Anthropic's Fable plan page:Plan
Fable accessFree
Not available at allPro ($20)
"Fable 5 and Fable 5.1 aren't included in your plan's usage limits. You can use them with usage credits."Max ($100 / $200)
Included as standard, up to 50% of weekly usage limits at no extra costAnthropic's own Fable plan documentation. This is the sentence that makes the tiers genuinely different rather than just bigger. Still current as of September 5, 2026.
Read that Pro row again, because it is the part people miss: you can still use the newest model on the $20 plan. You just pay for it separately, on top of your subscription.
That makes "core features are identical, only usage differs," the line nearly every comparison page repeats, no longer true.
I am on Max 20x, so Fable is included for me. That is exactly why I want to flag it clearly rather than let it sit in a footnote: the tier I happen to be on is the one where this change is invisible.
The "Max unlocks Opus" myth, corrected properly
This is the single most repeated claim about these tiers, and it is still wrong.
Here is the sentence the whole myth turns on, from Anthropic's Opus 5 announcement:"Opus 5 is the new default model on Claude Max, and the strongest model on Claude Pro."Read both halves.
Max gets Opus 5 as its default. Pro gets Opus 5 as its strongest available model. Same model, both tiers. What differs is which one you land on when you do not choose.
So "Max unlocks Opus" is false, and people repeating it are misreading a real sentence rather than inventing one. Anthropic does draw a Pro and Max distinction here. It is a distinction about defaults, not about access.
I previously wrote on this page that no official Anthropic page mentions Opus in the tier context. That was wrong: Anthropic published the sentence above on July 24, and this page went up on July 31. It was already wrong the day I published it.
The current lineup, verified on claude.com/pricing on September 5, 2026, is Haiku 4.5, Sonnet 5, Opus 5 and Fable 5.1. If you are unsure which to reach for, that is Claude Opus vs Sonnet, and picking a lighter model is often cheaper than buying a bigger plan.
The three rungs, and the trigger at each one
The query says "versus" but you have three choices, and the step between them is a different decision each time.
Pro, $20. Stay here if: you brainstorm, generate scripts, use Cowork occasionally, and you have your process dialled in. That combination genuinely fits inside $20 and a lot of people talk themselves out of it too early.
Trigger to leave Pro: you are trying to build something serious and you are hitting a stop. Not an inconvenience, a stop, on work with a deadline.
Max 5x, $100. Stay here if: the five-hour wall was your problem and $100 made it go away. For most people who outgrow Pro, this is the end of the ladder.
Trigger to leave Max 5x: you are experimenting constantly rather than executing a known process. Researching competitors, building ten skills, not entirely sure what you are doing yet. That mode burns capacity in a way that dialled-in work does not.
Max 20x, $200. This is where I am. And here is the honest caveat: you are still going to run out of usage. It depends entirely on how complex the thing you are building is. Nobody should buy the top tier expecting the meter to disappear.
That is the answer this page did not previously give, and it is the one I would want if I landed here from a search.
What is NOT the difference
"Claude Code and Cowork are Max-only." Both are included on Pro. Claude's pricing page lists "Includes Claude Code" and "Includes Claude Cowork" under the $20 tier.
"5x means five times my monthly total." Anthropic anchors that multiplier to per-session capacity, not a monthly pool.
Annual billing on Max. Anthropic's own page disagrees with itself here, which is worth knowing. Its billing FAQ says annual subscriptions are available for Pro and Team; its Max FAQ says both Max options are billed monthly; the comparison table lists Max's billing cycle as "Monthly and annual." Check what checkout actually offers you.
The real upgrade trigger, from someone who climbed the whole ladder
I did not decide to pay $200. I got walked there.
It started as a sparring partner. I had bugs or I wanted to build something, so I used Sonnet, and honestly I did not even see the power of Opus at the time. Sonnet alone was good enough that I thought, this is better than ChatGPT.
Then I kept hitting the limit. And I still did not upgrade. I just waited it out.
What changed everything was Claude Code. Not because I needed a coding tool, but because it was the first time the thing could take work off my load instead of just answering me.
The clearest version of that is what happened to my video editing, because it shows how the climb actually goes.
It started as "cut my shorts." Then it became "cut my long form," which is longer and more resource-hungry. Then it became "do not just cut, add visual aids": take screenshots of my actual laptop, different apps, so my raw talking-head footage has real material on screen while I talk.
That last version runs for hours to produce 14 shorts. And the same system now generates SEO content for me too.
The point is not the video. The point is that the skill keeps doing more and more while I spend about the same energy.
Every one of those steps raised my usage. That is what walked me up the tiers.
Here is the math that made $200 obvious. If I hired an editor for shorts, that is call it $50 a video on the low end, and realistically $100-plus if I want actual cuts with visual aids. Plus waiting days for turnaround.
So the ROI is not close.
And here is how the wall actually showed up in my day: I was waiting instead of working. That is the whole signal. Not a warning banner. Not a percentage. Just me sitting there unable to continue.
But notice the order of that story: the upgrade was the last step, not the first.My own Claude billing screen. I am on Max 20x at $200 a month, which is the tier where Fable is included in-plan.
Three cheaper things to try before $100
If you are getting stopped but not sure the upgrade is earned yet, work through these first.Use a lighter model. This is the cheapest fix on the list and most people skip it. Anthropic's model guide states that running Opus or Fable on a task Sonnet or Haiku could handle uses more of your limit for no gain. There is also an Effort control in the model picker that changes consumption on the model you already chose.
Buy usage credits instead. Pro and Max subscribers can purchase discounted usage bundles rather than jumping a whole tier. If your overflow is occasional, this is dramatically cheaper than $80 more a month, and it is also how you use Fable on Pro.
Put your reference material in a Project. Per Anthropic's usage-limit best practices, content in projects is cached and does not count against your limits when reused. People re-paste the same background documents into fresh chats all week and pay for it every time.There is a fourth that is really a habit: very long single conversations cost far more than short ones, because every turn re-reads everything before it. When I measured nine weeks of my own usage, starting a fresh session instead of riding one conversation was the change that moved the number most. Before I fixed it I was burning 20 to 30% of a week's top-model allowance in a single day.
If you work through all four and you are still getting stopped on work that matters, that is your answer.
So should a non-coder pay $100?
Here is the answer I would actually give you, and it is probably not the one you expect from someone paying $200.
If you have not hit the wall, congratulations. You are good.
If you do not create video, and you write LinkedIn posts, tweets, and Substack drafts with a few skills running on the $20 plan, that is all you need. Do not overcomplicate it.
I have a different use case. I tinker constantly and I create content about AI, so I am naturally tinkering and documenting all day. That is what outgrew the $20 plan, not some general truth about the tiers.
The limits are not a coding thing, though. They bind on any compute-heavy work: long document analysis, research, Cowork sessions. Plenty of non-coders hit them.
Here is what I would actually tell you: when it becomes a no-brainer, you will not question it. If you are overthinking the upgrade, that usually just means you have not found a strong enough use case yet. That is my rule for this decision, and it has held for every tier I have moved through.
Start at $20. Let the pain tell you when to move.
One billing warning before you change anything
Users report a real gotcha: changing your subscription through support rather than in Settings can cost you a legacy price you were grandfathered into.
Use Settings then Billing where you can. Upgrades take effect immediately with unused time credited; downgrades take effect at the end of your billing period, and your chats, projects, and files stay with your account either way.The 60-second version, from my Shorts: I made a $20 AI plan feel like the $200 one (routing, not a hack).If you are still decidingClaude Max 5x vs 20x - if Max is settled and the step is not
Claude Opus vs Sonnet - if a lighter model would fix this for free
Is Claude Pro worth it? - if you have not paid for anything yet
What Claude's free plan gets you - if you are not sure $20 is earned yet
Claude Max vs ChatGPT Pro - if you are comparing $100 tiers across vendors
Claude Cowork vs Claude Code - the one that burns your allowance fastestIf you want to put whichever plan you pick to work on real tasks, start at Claude at Work.
Published July 31, 2026. Last reviewed and updated September 5, 2026: added the Max 20x usage lawsuit and the consumer grievance behind it, restructured the page around the Pro to Max 5x to Max 20x ladder with the trigger at each step, added the honest caveat that you can still run out on 20x, replaced a half-remembered "10x to 15x" recollection with the figures the complaint actually alleges, added the lighter-model fix as the cheapest thing to try first, flagged Anthropic's self-contradiction on Max annual billing, and refreshed model names to Haiku 4.5, Sonnet 5, Opus 5 and Fable 5.1. Prices, plan limits and the Fable tier rules checked that day against claude.com/pricing and Anthropic's own support documentation, linked inline. Anthropic changes these often, so open the source pages before you buy.
Quick answer: Yes, most people should still pay for their own AI subscription - but not for the reason every other article gives.
It is not that Copilot is the dumb one - it runs current-generation models from the same labs. It is that there are two different products called Copilot, and the free one your company handed you probably cannot see your work data at all. And even when it can, it is your employer's tool, bought for your employer's benefit. That is the part nobody says out loud.
I use Copilot at my day job and ChatGPT on my own dime, so this is a comparison I actually live rather than one I assembled from spec sheets.
First: find out which Copilot you actually have
This is the most useful thing on this page, and almost no comparison article covers it. From Microsoft's own licensing docs, there are two tiers with nearly the same name:
Copilot Chat (free) - automatically included with an eligible Microsoft 365 subscription at no extra cost. Web-based chat: it sees the public internet and whatever you paste or upload. It has enterprise data protection, but it cannot read your organization's content.
Microsoft 365 Copilot (paid add-on) - a per-user license your company buys on top. This is the one from the advertising: work-based chat that grounds answers in what your work account can access - your email, files, meetings, and chats - plus in-app help inside Word, Excel, PowerPoint, and Teams.
The 10-second test: ask it something only your inbox or your files would know. "What did my manager email me about the Q3 budget?" If it cannot answer, you have the free one. Everything the ads promised you - the meeting summaries, the "draft this from my documents" magic - lives behind the paid license.
If nobody at your company bought the add-on, you are comparing ChatGPT against a web chatbot in a Microsoft wrapper. That is a very different comparison from the one you thought you were making.The paid product. Microsoft 365 Copilot's own page, captured July 27, 2026. If your company didn't buy this add-on, the Copilot you have is the free chat without access to your email, files, or meetings.The 60-second version, from my Shorts: The AI you already pay for is hidden inside Docs, Word, and Notion.What using it at work is actually like
My honest read, from using it: the interface is clunky. It feels like 2023 or 2024, like ChatGPT when it first came out. In the rollout I have lived with, it is just a chat, and the "agents" I could make were really prompt templates with a fancier name.
To be fair to Microsoft: at licensed tiers they do ship real prebuilt agents now, with names like Researcher and Analyst. Whether you ever see them is, again, an IT-department question, not a Microsoft question.
I am not the only one living this gap. One sysadmin evaluating both put it as: "We like the appeal of CoPilot being integrated with Outlook and Teams already... but the things it can do is honestly subpar at best compared to ChatGPT" (u/Apprehensive-Heat994 on r/sysadmin). And from someone locked into a restricted tenant: "I feel like a snail in a horse race" (u/Willing-Button-6452 on r/csMajors).
But the bigger thing is how much it varies. What Copilot can do depends entirely on your company and which features they enabled. Some organizations do not even have internet access switched on. There are a lot of constraints, and they are not the same constraints from one employer to the next.
Which leads to a caveat I want to be fair about: it is not really fair to compare Copilot against ChatGPT unless your company has both. Your Copilot experience is partly a review of your IT department, not of Microsoft.
Having used both, ChatGPT is more feature-rich - voice, custom GPTs that are genuinely easy to make, skills, Codex. For the work I do, it beats Copilot across the board. My honest opinion is that Copilot is well behind.
The plot twist: it is not a model problem
Here is what saves this from being a lazy dunk. Copilot is not running some inferior discount model - Microsoft ships current-generation models inside it, from the same labs whose products you are comparing it against.
So when Copilot gives you a worse answer than ChatGPT on the same prompt, the model usually is not the variable. The wrapper is. The context it is allowed to see is. The permissions are. A capable model that cannot read your files, cannot browse, and sits behind a clumsy interface will lose to a slightly different model that can do all three.
That reframe matters, because it tells you the fix is never "wait for Microsoft to get a better model."
The real reason the free one is not enough
This is the argument I care most about, and it has nothing to do with benchmarks.
You want to own your own data and your own processes.
You should not mix your personal life into your company's AI. When you are on the go, you are not going to use your employer's Copilot to work through your schedule, your goals, your finances, or your health. Nor should you - it is their tenant, their logs, their tool.
The reverse is just as true, and more likely to get you in trouble: do not leak company data into your personal AI. Check your policy. In a lot of organizations, pasting a client contract into personal ChatGPT is a fireable violation, and the affiliate-driven comparison pages will never tell you that.
So the split is not about which is smarter. It is:Work data stays in the work tool. That is what it is for, and it is the one your security team approved.
Your life stays in your tool. Schedule, brainstorming, news, goals, side projects, learning.Your company's Copilot is a tool the company gives you for their benefit, to do their work. That is not a criticism - it is just what it is. If you want to use AI to improve your own life, that is a different tool, and it is worth $20.
You can still use the work one well within those lines: brainstorming, abstract sessions, coding problems, help tracking and managing your schedule. Just nothing proprietary going out, and nothing personal going in.
Where Copilot legitimately wins
Compare, do not co-sign. Copilot has real advantages that have nothing to do with model quality:It is the one you are allowed to paste a contract into. That is not a small thing - it is often the entire decision at work.
It runs in the tenant your company already pays for, under the compliance and data-protection terms your legal team signed.
It is inside the apps - Teams meeting summaries, Excel work in place, Word drafting - if your company bought the add-on.
It costs you nothing personally.If your AI use is summarizing meetings, cleaning up emails, and light Excel work, all on data you legally cannot paste anywhere else, the free Copilot is genuinely enough. Some readers should not spend a dollar on a second tool. That is a real answer.
Here is my theory about why Copilot wins deployments anyway: it is the enterprise default. Low friction, already part of the family, bundled with the suite your company pays for. Nobody gets fired for switching it on.
That is a real kind of winning. It is just not the same thing as being the tool I would pick. With my own money, it is the last option I would choose, behind ChatGPT, Claude, Grok, and even Gemini.
The verdict
Free Copilot is enough if: your AI use lives inside Microsoft apps, on work data, doing work tasks - meeting notes, email cleanup, formulas.
Pay for your own if: you want memory across sessions, real research, image generation, voice, custom assistants, or you are building anything on the side that is not your employer's business. Also if you simply want a place to think that your employer does not administer.
The giveaway: if you keep pasting things out of your work tools to get a better answer, you already need the second tool. You have been running the experiment without noticing.
How to buy it: start with the free tier. If you constantly get logged out or maxed out on usage limits, that is a good sign it is time to upgrade. And when you do - go look at what else you are subscribed to. You would be surprised how much money you are already wasting on things you do not use.
If you want the cheaper rung, ChatGPT Go is $8/mo and covers ordinary writing work. Which tool to buy at $20 is its own question - I cover it in Claude vs ChatGPT and, tier for tier, in Claude Pro vs ChatGPT Plus.
For what it is worth, the tools I think are actually worth your attention right now are OpenAI, Anthropic, Grok, Gemini, and Perplexity. Most of the rest is noise. There are strong local models too - GLM, Gemma, Kimi - and you can run them on your own machine, but that is a more complex path than this post is about.
More on using AI inside a normal job: Claude at Work.
Quick answer: Go is $8 a month, Plus is $20, and the $12 gap buys model access, not capacity. Plus gets GPT-5.6 Sol, legacy models, full deep research and Sites, reaches GPT-6 Astra through Work and Codex, and will never show you ads. Go gets none of those. What the gap does not buy is context: both tiers sit at the same 54K instant and 256K reasoning window.
I have never paid for Go. I saw the $20 plan, started there, and have been upgrading ever since, so trying Go now would make no sense for how I work.
That means everything factual below is read off OpenAI's own pages rather than lived. Including one thing this page previously got wrong.
First, the correction I owe you
The earlier version of this page told you that Go's reasoning tops out at GPT-5 Thinking Mini and framed the gap as "Go has no real reasoning."
That is wrong, and OpenAI says so directly on the Go help page:"Are reasoning models included with ChatGPT Go? Yes. ChatGPT Go users can access Think on the web and in the ChatGPT mobile app... Think uses GPT-5.6 Luna. ChatGPT Go does not include GPT-5.6 Sol."So Go reasons. It just reasons with Luna instead of Sol.
The gate is which reasoning model, not whether there is one. That is a smaller and more honest distinction than the one I published, and I would rather correct it in public than quietly swap a sentence.
I also removed the comparison graphic that used to sit on this page, because it carried the same stale claim and two model names OpenAI no longer lists.
The short answer, by who you areYou are
Pick
WhyNot hitting any limits on Free
Stay free
You are shopping for a problem you do not haveDrafting, rewriting, summarizing, everyday questions
Go, $8
Genuinely enough, and the context window is the sameBothered by ads on principle
Plus, $20
Go is ad-eligible and you cannot buy your way outAsking it to think through work you will act on
Plus, $20
This is what Astra and Sol are forAttached to a legacy model workflow
Plus, $20
Go drops legacy models entirelyWanting full deep research and Sites
Plus, $20
Go gets a limited version; Sites is Plus onlyGo is $8, and our own tracker had it wrong
OpenAI renders its plan prices in the browser rather than in the page source, which is why articles quote wildly different numbers and why our own plan tracker carried a wrong one until today.
So here it is, rendered and captured.ChatGPT's pricing page as it actually renders, captured September 5, 2026. Go is $8. Plus is $20.
Go is $8 a month. Plus is $20. The gap is $12.
Our AI plan tracker previously listed Go at $12, confusing the gap with the price. That is corrected as of today, with the correction logged on the tracker itself rather than silently edited out.
What the $12 gates: Astra, Sol, legacy models and ads
From OpenAI's pricing comparison as it rendered on September 5, 2026.Free
Go, $8
Plus, $20GPT-6 Astra
No
No
Yes, in Work and CodexGPT-5.6 Sol
No
No
YesGPT-5.6 Luna
Yes
Yes
YesGPT-5 Thinking Mini
Yes
Yes
ExpandedReasoning (Think)
limited
Yes, on Luna
Yes, expandedLegacy models
No
No
YesAds
eligible
eligible
neverInstant context window
27K
54K
54KReasoning context window
Varies
256K
256KInstant input maximum
~12 pages
~40 pages
~40 pagesDeep research
limited
limited
YesSites
No
No
YesResponse times
limited on bandwidth
limited on bandwidth
FastTwo rows in that table deserve to be pulled out.
One clarification on that Astra row, because it trips people up. Plus reaches Astra through ChatGPT Work and Codex, not through the chat box, where the model is branded GPT-6 Pro and gated to Pro and above. Go does not reach it on any surface. The full breakdown is on which plans get Astra.
The context windows are identical. 54K instant, 256K reasoning, on both tiers. This is the fact I would tattoo on the page. You are not paying $12 for more room to work; you are paying for a better thinker inside the same room.
Response times differ, and nobody mentions it. Free and Go are listed as "Limited on bandwidth & availability." Plus and Pro are listed as "Fast." At peak times, that is a real difference in a way a feature checklist does not convey.
The ads question, which settles it for a lot of people
This alone decides it for some readers, and almost no comparison page leads with it.
From OpenAI's ads FAQ, which was stamped "Updated: 4 days ago" when I read it on September 5, 2026:"Ads may appear for users on the Free and Go plans. Plus, Pro, Business, Enterprise, and Edu accounts will not have ads."The US test began on February 9, 2026, with ads appearing below responses, clearly labeled and visually separated. OpenAI says they do not influence answers and that advertisers never receive your chats, chat history or memories.
Here is the part that matters for this decision: on Go you cannot pay your way out of ads.
The "Ads-Free" option is a Free-plan setting, and it trades ads for lower message limits and reduced feature access. OpenAI's own FAQ spells it out: if you are on Go and you want Ads-Free, you would have to switch back down to Free, or up to a plan that is already ad-free.
Turning off ad personalization is not the same thing. It changes which ads you see, not whether you see them.
One inconsistency worth knowing about, because it is OpenAI's and not mine: the ads FAQ describes a live test covering Go, while the Go help page still says "We may start testing ads in ChatGPT Go in the future." Two OpenAI pages, two tenses. Check your own account rather than trusting either page, or this one.
Go is no longer region-locked, and both tiers chat without limits
Go is no longer region-limited. OpenAI's help center now states that "ChatGPT Go is now available in all ChatGPT supported countries." A lot of content still describes Go as a select-regions product, including some of OpenAI's own marketing.
Free and Go both get unlimited everyday text chats. OpenAI states it plainly: "Free and Go users both have unlimited everyday text chats." What Go raises is the limits on the expensive things: image generation, file uploads, data analysis.
That reframes the $8 a bit. You are not buying the ability to chat more. You are buying more of the tools.
Where Go genuinely wins
Go is a real product for a real person and I want to say so clearly, because the internet's instinct is to treat cheap tiers as traps.
If your work is volume of ordinary language, drafting emails, rewriting paragraphs, summarizing documents, everyday questions, Go covers it. You get the same context window as Plus, you get Think on Luna, you get projects, custom GPTs and scheduled tasks.
Users who moved down and stayed happy describe exactly this. u/varik21 put it simply:"For just conversations, there is absolutely no difference... I am going back to Go."The other side of that same thread is just as consistent. u/Stade_gangster:"answers feel much more geared towards being fast than being thoughtful... if you mainly use ChatGPT for in-depth discussions, I can definitely understand why people choose to stick with Plus."Read those two together and you have the whole decision. Conversations, Go is fine. Thinking you will act on, pay the $12.
Downgrading is not refunded, and three other traps
Read these before you click downgrade.Downgrades are not refunded. OpenAI states your current plan runs to the end of the billing cycle, then you switch to Go and pay the Go rate. u/Icy-Imagination5185, who did it by accident, put it plainly: "Now I can't go back to Plus for a FULL MONTH."
No annual billing. OpenAI does not offer annual or multi-month prepay on Go, Plus or Pro.
Legacy models disappear on Go. If you have a workflow pinned to an older model, it goes away.
Ads-Free is a Free-plan setting, not something you can add to Go.
No Sora on Go. The video model, along with Sites, record mode and full Codex, sits above the $8 tier. ChatGPT Work is on both, but Free and Go get it as "Limited (desktop app)" while Plus gets desktop, web and mobile.My rule for cheap tiers
Here is the rule I actually use, and it is the only reason I have ever moved up a tier: pain is the upgrade trigger.
Start free. If you only occasionally bump a limit, do not upgrade. You are not really using it yet.
Once you find real use cases and the tool becomes part of your life, managing email, planning your day, the stuff you connect through apps and connectors, you will want the state-of-the-art model and more usage. That is when $20 becomes obviously worth it.
So where does that leave Go? Go is what you pay when you are past free, you keep hitting limits, and you are fine without the newest models. That is what the $8 buys. Be honest about which person that is.
The cheap tier is a trap in exactly one situation: when your work needs judgment and you bought volume instead. A wrong answer you act on at work costs far more than the $12 you saved.
One option the comparison pages skip: $8 on Go plus a second free tier from another provider often gets you further than $20 on one. More friction, more model.
How I checked this
Reasoning access, the ads policy, availability, the unlimited-chats language and the refund rules come from OpenAI's Go help page and its ads FAQ, both read September 5, 2026. Prices, model rows and context windows come from chatgpt.com/pricing rendered in a browser the same day, which is the only way to read those numbers reliably.
I have never subscribed to ChatGPT Go, so there is no first-hand usage report here and I am not going to invent one. OpenAI publishes no message counts for these tiers either, so any "X messages per day" figure you see elsewhere is folklore.
Related readingChatGPT Plus vs Pro - the next rung up, if $20 is already settled
Is ChatGPT Plus worth it - the $20 decision on its own
Claude Pro vs ChatGPT Plus - if you are choosing between $20 plans
Claude vs ChatGPT - the full comparison, and why caps decide itThis post is part of Claude at Work, the hub for using AI at your job without code.
Published July 27, 2026. Last reviewed and updated September 5, 2026: corrected this page's central claim that Go lacks real reasoning (it has Think on GPT-5.6 Luna), re-cut the gate around GPT-6 Astra, Sol, legacy models, deep research and Sites, settled the price at $8 against OpenAI's rendered pricing page and fixed the wrong $12 figure on our own plan tracker, added the ads section and the response-time row, noted that Go is now available in all supported countries, and removed a comparison graphic that carried the stale reasoning claim. All facts checked that day against OpenAI's own pages, linked inline.
Quick answer: I pay for both. Not $20 each anymore - $100/mo for ChatGPT Pro and $200/mo for Claude Max, and I got there the same way most people will: I hit the caps, got tired of waiting, and upgraded.
That is the part the comparison articles skip. Everyone argues about which model writes better. Almost nobody tells you that the decision you will actually make is not "which one is smarter" but "which one stops me mid-task, and how often." Quality gets you interested. Caps are what make you move your money.
So here is the comparison from someone who pays real money on both sides, including the honest starter advice that costs me nothing to give: if you have $20 and you are just beginning, get ChatGPT.
What changed since I wrote this (updated August 7, 2026)
Both sides moved in two weeks. None of it breaks the thesis above, but it changes what you get for your money.July 24: Claude Opus 5 shipped. It is now the default model on Claude Max and the strongest model on Claude Pro, per Anthropic's announcement. My honest take on whether I felt it is further down, and it is not the take you expect.
July 30: OpenAI shipped GPT-5.6 and renamed the family: Sol (frontier), Terra (balanced), Luna (fast and cheap). Source.
August 6: ChatGPT's free tier stopped being a demo. OpenAI is making GPT-5.6 Luna the free default and, verbatim, "expanding access with unlimited text chats," plus a new Think button for harder questions. Source.Two things to hold onto before you get excited about that last one.
It is a staged rollout. OpenAI's own wording: Luna becomes the default for Free and Go "this week," and "starting next week, they'll also have unlimited text chats and access to a new Think button." So do not assume it is live in your account today. Mine was still the old behavior when I checked on August 7.
And the caveat that keeps this whole page true, also verbatim: "Limits will still apply for file uploads, images and other tools." Unlimited text chat is not unlimited work. The moment you start dropping in PDFs and spreadsheets, you are back in cap country.
OpenAI also published an accuracy number worth knowing and worth discounting appropriately. In its own internal evaluation of financial, medical, and legal prompts, "responses containing at least one factual error were about 62% less common with GPT-5.6 Luna and 68% less common with GPT-5.6 Sol than with GPT-5.5 Instant." That is a vendor testing its own product, not an independent benchmark. Treat it as a direction, not a fact.
One more scope note, because people misread this release: the version of GPT-5.6 Sol that powers Work and Codex is not changing in this update.
The short answer, by who you areYou are
Get this
WhyStarting out, one subscription, $20
ChatGPT
More general, more use cases, easier first winWriting, designing, or building things you ship
Claude
Better creative work, better instruction-followingUsing AI seriously every day
Both at $20
Two separate allowances beat one bigger oneConstantly locked out mid-task
Upgrade the one you live in
The cap is the signal, not the benchmarkA developer or heavy agent user
Claude, and expect to outgrow $20
Claude Code shares your chat allowanceWhat each one actually is right now
Both ladders look almost identical, which is why price alone tells you nothing.Claude
ChatGPTFree
$0
$0Cheap tier
-
Go, $8/moStandard
Pro, $17/mo annual ($200 up front) or $20 monthly
Plus, $20/moPower tier
Max, from $100/mo (5x or 20x Pro usage)
Pro, from $100/moModels
Opus, Sonnet, Haiku on every tier. Fable 5 and Fable 5.1: no on Free, credits-only on Pro, included on Max
GPT-5.5 Instant, GPT-5.6 Sol / Terra / Luna, GPT-5 Thinking MiniContext window
200K default; Fable 5.1, Opus 5, and Sonnet 5 support 1M on paid plans
54K instant / 256K reasoning on PlusCoding agent
Claude Code, included in paid plans
Codex, included on PlusCorrection, and it matters more than anything else in this table. When this page first published on July 27, that models row said "Fable, Opus, Sonnet, Haiku" for Claude with no tier split. That was wrong, and it had already been wrong for a week. Here is the actual state per Anthropic's Fable 5 plan page: the promotion that included Fable 5 in Pro's weekly limits ended July 19, 2026 at 11:59:59 PM PT. Free gets no Fable 5 or Fable 5.1. On Pro, “Fable 5 and Fable 5.1 aren’t included in your plan's usage limits” - you use them with usage credits, which is real money on top of your $20. On Max both are included as standard, up to 50% of your weekly limits at no extra cost.
So the newest model is now a $100-tier feature in practice.
Sources: claude.com/pricing, Anthropic's Fable 5 plan page, and chatgpt.com/pricing, rechecked 31 July 2026. These change often - open them before you buy.
Note on my own plan: the table shows Max starting at $100, which is the 5x tier. I am on the 20x tier at $200, which is why Fable 5 and Fable 5.1 are included for me; I use Fable 5 in-plan.
Does that change the verdict at the top of this page? It sharpens it.
The whole argument here is that caps decide this, not quality. Now the newest model itself sits behind a cap on the $20 tier. If you want the most frontier model, compare the benchmarks - and if ChatGPT's Sol at extra-hard effort beats Opus 5, that is a real difference: you get a better model for the same price, and ChatGPT is straight up more generous about resetting limits. I did not experience that until I was paying for both. Now I can attest to it.
That is not a knock on Claude. It is the same point this page opened with, just with a bigger price tag attached: the thing that moves your money is what you get stopped from doing.
One structural difference is worth more than the whole table: ChatGPT has an $8 tier and Claude does not. If your budget is genuinely under $20, that is not a tie, it is a one-horse race. I break down what that $8 does and does not buy in ChatGPT Go vs Plus.Claude's real pricing page, captured July 27, 2026. The ladder the whole decision runs on.
Which model is your plan actually giving you
This is the question nobody answers, and it is the one that decides whether you feel ripped off.
Both companies sell you a plan, not a model. The plan then decides which brain you get by default, which ones you can switch to, and which ones quietly cost extra. Here is the plain-English ladder for each, weakest to strongest.
Claude: Haiku, then Sonnet, then Opus, then Fable, then Mythos.Your plan
Default you get
Can you reach the top?Free
Sonnet
No Fable 5 or Fable 5.1 at allPro ($20)
Opus 5 is the strongest included model
Fable 5 and Fable 5.1 are not in your plan limits. They bill as separate usage credits, real money on top of your $20Max (from $100)
Opus 5 is the default
Fable 5 and Fable 5.1 are included, drawing from your regular weekly pool at up to 50% of limits, and they burn that pool faster than other modelsThe trap on Max: you do not get 50% extra capacity for Fable. It comes out of the same pool, just quicker. Anthropic's own Fable 5 plan page is explicit about it.
ChatGPT: Luna, then Terra, then Sol, then Sol Pro.Your plan
Default you get
Can you reach the top?Free
GPT-5.6 Luna, plus a Think button for harder questions
No SolGo ($8)
Same models as Free, more uploads, voice, and memory
No SolPlus ($20)
Sol, Terra, and Luna, with a new effort slider
Sol, yes. Sol Pro, noPro ($100+)
Everything above
Adds GPT-5.6 Sol ProThe thing to notice: the model a free ChatGPT user talks to and the model a Plus subscriber talks to are not the same model. Luna is the cheap fast one. Sol is the flagship. When someone tells you "ChatGPT got dumber" or "ChatGPT is amazing at research," ask which one they were on, because they are probably not describing the same product you are paying for.
Same on the Claude side. A person raving about Fable and a person on Pro who has never touched it are having two different experiences of one $20 subscription.
Steal this rule: before you judge either tool, open the model picker and read what it actually says. Most "this AI is bad" complaints are really "my plan put me on the cheap tier and did not tell me."
The real difference nobody leads with: caps, not quality
Here is the mechanism, because it explains almost every angry post you will read about either tool.
Claude runs a rolling five-hour session limit plus a weekly cap on paid plans. Critically, everything shares one pool: your chats, Claude Code, and Cowork all draw from the same allowance. Spend an afternoon on an agent task and you have also spent your chat capacity. When you hit the wall, you stop.
ChatGPT leans on rolling limits and, when you exhaust the good model, tends to move you down a tier rather than shutting the door.
That difference in style is the whole ballgame. A hard stop mid-workday feels completely different from quietly getting a smaller model. It is why you see posts like this one from r/ClaudeAI, where a paying customer put it better than any reviewer:"I cannot pay for a product, use it normally for two hours, and then be locked out. I especially cannot accept a weekly lockout." ... "The jump from Pro to Max is $80/month. That's not a tier, that's a cliff."u/mcburgs, r/ClaudeAII am not going to pretend that is wrong, because I am the guy who climbed that cliff. I started on both at $20. I hit the caps regularly, I did not want to wait hours, and the math worked, so I upgraded. That is not a defense of the pricing. It is just what happens once you depend on the tool.
Does the August 6 free-tier change soften this? For chatting, yes. For work, no. Unlimited text chat does not cover file uploads, images, or tools, and Claude's shared pool did not change at all. If you want the actual numbers before you pick a side, I broke the whole system down in Claude usage limits explained, and is Claude Max worth it is the honest math on the $100 jump.The 60-second version, from my Shorts: Claude or ChatGPT? Wrong question. Here's when to use each..How I actually run it
I pay $200/mo for Claude Max and $100/mo for ChatGPT Pro. What I get for it, concretely: Claude does video editing work, visual aids for videos and shorts, script ideas for long-form, and turns my YouTube long-form into a newsletter. Claude also writes my Substack newsletter and LinkedIn posts; my ChatGPT subscription generates the visuals that go with them.
The starting point of that ladder is still two $20 subscriptions, and that is what I would tell almost anyone reading this. It is also, independently, what people land on in the threads - one reply in that same r/ClaudeAI discussion just said: "Get 2 $20 subscriptions! That's in the middle!" Two separate $20 allowances genuinely beat one $40-equivalent allowance on a single account, because the caps do not pool.
Upgrade when the pain shows up - not before. If you are not hitting limits, a bigger plan buys you nothing.
Did Opus 5 change anything for me? Not really.
Opus 5 has been the default on my Max plan since July 24. I use Claude for hours most days.
I have not noticed a difference.
That is not a criticism, and it is not me saying the model is not better on paper. It is me telling you the truth about what a frontier upgrade feels like from inside the daily work: mostly like nothing. If you were waiting for a new model to fix your workflow, the workflow was the problem.
What I did notice was Claude acting strange starting around August 5, which is a different thing from a model being worse. I looked it up rather than trust my own vibes, and there was a real incident. Anthropic's status page logged "Degraded performance for Claude Opus 5" on August 5, 2026 from 13:51 to 14:34 UTC, on top of a longer "Degraded performance of multiple models" window that morning, 07:05 to 14:14 UTC, hitting Mythos 5, Fable 5, Opus 5, and Sonnet 5.
Useful habit: when a tool suddenly feels dumber, check the vendor's status page before you rewrite your prompts or cancel your plan. Sometimes it is genuinely them.
About the affiliate question, straight
There is a fair suspicion going around that people recommend Claude because the affiliate money is better.
So, flatly: I have no affiliate relationship with Anthropic. I do not earn anything when you subscribe to Claude. I pay $200/mo for Max and $100/mo for ChatGPT Pro out of my own pocket, which is the opposite of a business model.
Here is the version of my opinion that should be hardest to fake, since it cuts against the thing I supposedly shill: GPT-5.6 Sol at the top effort setting is genuinely phenomenal for research and for hard questions. It is not a consolation prize. If your work is mostly "help me think through something complicated," that is a great place to be.
I still think Fable is a little bit better. That is my preference from daily use, not a benchmark result, and I would not argue hard with someone who lands the other way.
If I had to pick one for a year
Claude. Easily.
Content, design, creative writing, website building - all of it is just better, and the code is genuinely top tier. I want to be fair about the other side though: Codex comes very close, and in some cases it is better - better research, and good as a sparring partner. They are both good.
But here is the honest cost of that choice: I would lose image generation, so I would pay for OpenAI image gen through the API instead, probably $10-30/mo. At which point I might as well have just kept the $20 ChatGPT subscription. Or I would use Nano Banana and sacrifice a little quality.
That is the real shape of "just pick one." It rarely stays one.
The starter advice that surprises people
If you have $20 and you need to start: just get ChatGPT.
I know how that reads coming from someone who pays $200/mo for Claude. But it depends on your goal, and for a general "I have twenty bucks and I want to begin," ChatGPT is the better first purchase. You can get past the writing quality - train it on your voice and it will get you 80% there. That is not a deal breaker for most people.
Does the August 6 free tier kill this advice? It makes it stronger.
My actual rule has always been: use the free tier, and let the pain force you to subscribe. Let that be the guide. The problem before was that the free tiers were too thin to generate honest pain. You would hit a wall in ten minutes, learn nothing about whether the tool fits your work, and either quit or pay blind.
Now there is a real floor to stand on. Unlimited text chat on ChatGPT Free means you can actually run your work through it for a week before spending a dollar.
So: start free. Use it for real tasks, not test questions. When you get stopped by something that matters - a file you cannot upload, an image you cannot make, a tool you cannot reach - that is the signal. Pay then, and you will know exactly what you are paying for.
Pick based on your goal, not on which brand a YouTuber likes.
Which one wins each job
The head-to-head is settled per task, not in general. These are the task-level breakdowns:Claude vs ChatGPT for writing
Claude vs ChatGPT for PDFs
Claude vs ChatGPT for Excel
Claude vs ChatGPT for PowerPoint
Claude vs ChatGPT for email - only one can actually hit send
Claude vs ChatGPT for meeting notesAnd the two tier-level decisions:Claude Pro vs ChatGPT Plus - the $20 question
Claude Max vs ChatGPT Pro - who actually needs $100+
Claude Pro vs Claude Max - the within-Claude fork, and where the Fable 5 change actually lands
Is Claude Pro worth it? - if you are still deciding whether to pay at allAnd if Claude wins your head-to-head, the next fork is which agent door to open: Claude Cowork vs Claude Code - the no-terminal answer.
If your work is inside a company that hands you Microsoft tools, the comparison you actually want is Copilot vs ChatGPT - or, against Claude directly, Claude vs Copilot, written from using Copilot at my day job.
The rest of the field, judged the same way: Claude vs Gemini for the Google-natives, Grok vs Claude and Grok vs ChatGPT for the fast-moving newcomer, Perplexity vs ChatGPT for the search-first crowd, and Claude vs ChatGPT for research for the job where my verdict flips. Every plan price on all of these, verified and dated: the AI plan tracker.
The one-paragraph version
Claude is the better tool for work I ship. ChatGPT is the better first subscription and the better generalist. The caps, not the benchmarks, are what will eventually move your money - and when they do, the honest middle step is two $20 plans before either $100 tier. Start with the goal, not the hype.
Building automation instead of just chatting? Start at Claude at Work.
Published July 27, 2026. Last reviewed August 7, 2026. What was rechecked and added today: OpenAI's August 6 ChatGPT update, confirmed against openai.com and quoted verbatim, including the staged rollout wording and the "limits will still apply for file uploads, images and other tools" caveat; the GPT-5.6 rename and Claude Opus 5 release dates; the Claude Fable 5 plan mechanics, re-confirmed against Anthropic's Fable 5 plan page, which still show credits-only on Pro and included-to-50%-of-limits on Max with no extra capacity; and the August 5 Opus 5 degradation incident on Anthropic's status page. A new plain-English model ladder for both vendors was added. Dollar prices in the pricing table were not re-verified this pass, because chatgpt.com/pricing renders its amounts client-side and I could not confirm them from the source, so those rows still date to the July 31 check. Open both pricing pages before you buy.
Quick answer: pick by billing model, not by features. Claude Cowork draws from the Claude plan you already pay for, $17/month annual or $20 monthly on Pro, with nothing extra to buy. ChatGPT Work bills as metered credits from a pool it shares with three other agents.
Feature-wise they do the same job. The differences that will actually bite you are the ones nobody on this search result page has written down yet.
My rule for the four surfaces: Chat answers. Cowork and ChatGPT Work do your office work. Claude Code and Codex build software.
That extends the split from Claude Cowork vs Claude Code, where the rule was "Chat answers. Cowork does your office work. Claude Code builds software." OpenAI has now shipped its own version of the same three-lane structure, which is why this comparison exists at all.
If you searched "claude cowork vs chatgpt work," "chatgpt work vs claude cowork," "chatgpt cowork equivalent," "claude cowork vs chatgpt agent mode," or "claude cowork vs chatgpt codex," this page is the corrected version. Four things the current results get wrong, then the honest verdict.
The comparison in one tableClaude Cowork
ChatGPT WorkWhat it is
Agent for non-coding knowledge work
"An agent designed for longer, multi-step work and finished deliverables"Announced
Early 2026, GA on paid plans
July 9, 2026, rolling out over "the coming days"Where it executes
Remotely, on Anthropic's servers
Cloud, same as CodexLocal file access
Through the Claude Desktop app only
Desktop app only. Web and mobile cannotHow you pay
Included in your plan allocation. No separate SKU
Metered credits from a shared agentic poolShares its budget with
Nothing. It is your plan
Codex, ChatGPT for Excel, Workspace AgentsEntry price
Pro $17/mo annual, $20 monthly. Max from $100/mo
Included with eligible paid accounts, per OpenAI's rollout note. Credits meteredSibling surfaces
Chat, Cowork, Claude Code
Chat, Work, CodexCompliance API coverage
Not captured, per Anthropic's docs
Follows standard ChatGPT workspace controlsTwo products, same job, four corrections needed before the table above means anything.
Correction 1: ChatGPT agent mode is retired, so most comparisons are describing a dead product
Search this topic today and you will get a stack of articles comparing Claude Cowork to "ChatGPT agent mode."
That product is gone.
OpenAI's help center says it in one line: "ChatGPT agent is no longer available. Use ChatGPT Work for longer, multi-step tasks and finished deliverables." The same page notes that Operator, the earlier browser-driving agent, was folded into agent mode before that, and "the Operator website is no longer accessible."
So the lineage runs Operator, then agent mode, then ChatGPT Work. Two renames in about a year.
The product you are actually comparing against Cowork is ChatGPT Work, announced July 9, 2026. ChatGPT now has three surfaces: Chat for conversation, Work for multi-step deliverables, Codex for software. If that structure sounds familiar, it is the same three-lane split Anthropic shipped with Chat, Cowork, and Claude Code.
One practical note before you go looking for it: OpenAI says ChatGPT Work "is gradually rolling out to eligible accounts over the coming days." If it is not in your account yet, that is why.
Correction 2: the local vs cloud story is backwards in both directions
Here is the claim you will see repeated everywhere, including in AI-generated summaries at the top of this search: Claude Cowork is the local-first one, ChatGPT is the cloud one.
It is backwards.
Anthropic's own getting-started documentation states that "Cowork runs your tasks remotely (in beta). Claude's work runs on Anthropic's servers, in an isolated environment." The local part is a bridge, not the execution: "When a task needs something on your computer, like a local file or your browser, Claude reaches it through the Claude Desktop app on that computer."
Meanwhile OpenAI's documentation is equally blunt about its own limit: ChatGPT "Work on web and mobile cannot directly access files on your computer." Local file access requires the desktop app, and on the free and Go tiers it is listed as "Limited (desktop app)" on OpenAI's pricing page.
Read those two side by side and the real picture is boring: both products execute in the cloud and reach your local files through a desktop app. Neither one is running on your laptop.
A dated correction to my own earlier post
This is where I have to correct my own page, because the same mistake is on it.
Claude Cowork vs Claude Code, published July 9, 2026, closes with the line that Cowork's "home-field advantage is working directly on your local files and folders." That line needs an update as of July 2026. Anthropic's current documentation describes remote execution on their servers with the desktop app acting as the bridge to local files, which is a meaningfully different architecture than "works directly on your local files."
The practical consequence: tasks built on connectors run in the cloud without your machine, and tasks pointed at a local folder still need an awake desktop. Same for ChatGPT Work.
If a comparison tells you one of these is the "local" option, it has not read either company's docs this month.Correction 3: the billing models are structurally different, and that is the actual decision
Everything above is trivia compared to this. The two products meter you in genuinely different ways, and nobody explains it.
Claude Cowork draws down the plan you already have. Anthropic's Cowork docs list it as "available for paid plans (Pro, Max, Team, Enterprise)" with no separate purchase. There is no Cowork SKU. What there is, printed in the same docs: "Working on tasks with Cowork consumes more of your usage allocation than chatting with Claude," and specifically that "auto mode consumes more of your usage limit than the other modes."
That matches the verdict from Claude Cowork vs Claude Code: "Claude Cowork consumes limits faster than Chat," and Claude Pro caps you two ways, a 5-hour session limit plus a separate weekly cap. The plan-sizing conclusion there still holds. The $20 tier is sized for chat-first use with some agent work on the side. If Cowork becomes your daily workhorse, that is what the Max tiers from $100/month are for, which is the same upgrade trigger covered in Claude Max vs ChatGPT Pro.
ChatGPT Work meters credits from a shared pool. OpenAI's Codex plan-usage article states that "usage from Codex, ChatGPT Work, ChatGPT for Excel, and Workspace Agents draws from the same agentic usage and credit pool when those features are available on your plan." Work follows the same usage structure as Codex.
Sit with that for a second, because it is the sharpest practical difference on this page.
If you use Codex for a build, ChatGPT for Excel on a spreadsheet, and ChatGPT Work on a report, all three are eating the same budget. Heavy use of any one starves the other two. Cowork has no equivalent cross-product contention; it competes only with your own Claude chat usage.
So the honest question is not "which is cheaper." It is: do you want one metered pool shared across four agents, or one plan allocation that agent work drains faster than chat?
Nobody has published a real usage number, including me
Somebody on r/ClaudeCowork asked the exact right question and got nothing back. u/Proper_Club8511:As per title, does anyone have any information on the effective difference in usage between these two? I would like to know which is more efficient in general tasks. i.e. given the same task, and the same tier of plan (Pro), which would use more/less and by how much.That thread produced no useful answer.
It is the unmet question this page exists to address, and the honest response is that no public number exists. Neither company publishes numeric quotas. The metering units are not even comparable: Cowork burns a plan allocation, ChatGPT Work burns credits from a shared pool. There is no exchange rate.
Anyone handing you a "Cowork uses 2x more" figure invented it.
What the burn actually feels like on the Claude side
I cannot give you a Cowork-vs-Work efficiency number (nobody honestly can yet; the Reddit thread asking got zero useful answers). What I can tell you from daily use: Cowork burns more than chat because it does more. It is interactive, it reads reference files, it reads my wiki before it acts. That is the work you are paying for, not waste.
The place I actually feel a dent is not Cowork chores; it is heavy coding sessions in Claude Code, like editing my video-editing pipeline. If you are on a $20 plan, the efficiency rule that matters: plan with the big model, execute with the cheaper one (Opus to plan, Sonnet to execute). I skip that dance because I am on the $200 Max plan and just run Opus. And watch the open-ended autonomous runs: telling an agent "here's a goal, go" on something vague is the fastest way to torch a week's allocation on any plan.
Where ChatGPT Work genuinely wins
This page leans Claude, so take this section as the one it had to argue itself into. ChatGPT Work has real advantages, and two of them are structural.
One: the agent budget is portable across the whole stack. That shared pool is a downside when four agents fight over it and an upside when you want flexibility. A week where you do no coding means all of it goes to Work. Cowork cannot borrow budget from anywhere.
Two: it lands inside an account most workplaces already have. ChatGPT's enterprise footprint is bigger. If your company already pays for ChatGPT, ChatGPT Work arrives on your existing account with your existing admin controls and your existing Workspace Agents, no new vendor review. That is not a small thing when procurement is the actual bottleneck.
Three: three surfaces, one login. Chat, Work, and Codex live in one product with one billing relationship. Running Claude means Chat, Cowork, and Claude Code plus, if you also pay for ChatGPT, a second subscription.
There is also a real argument that Cowork's agent behavior is less predictable. u/jmillionair3 put it well on r/AI_Agents:Seems like GPT is going for more of a build and agent and deploy it places, and cowork deploys agents at its own discretion to carry out tasks but more or less similar in nature. What am I missing?He is not missing much, and that reading is fair.
ChatGPT's model leans toward agents you configure and place. Cowork's leans toward describing an outcome and letting Claude decide the steps. Neither is better in the abstract. If you want to see exactly what the second style produces in practice, 20 real Claude Cowork use cases has the community receipts.The 60-second version, from my Shorts: What to hand ChatGPT Work first (OpenAI's own advice).Where Claude Cowork genuinely wins
No separate meter to think about. Cowork is in the plan. You are not watching a credit balance while deciding whether a task is worth running, and you are not rationing between four agents. For someone who just wants the Monday report to happen, that mental overhead matters more than it sounds.
A more mature scheduled-task story. Cowork's recurring tasks run remotely on their schedule, hourly through weekly, even with the laptop closed, with the local-files exception noted above.
It is generally available, not rolling out. Cowork is live on Pro, Max, Team, and Enterprise today. ChatGPT Work is still arriving on eligible accounts, with Cowork's own web and mobile beta rolling out starting with Max.
And a caution against over-committing to either: portability is worth something. u/backhand_snipe on r/ClaudeCowork, writing about switching to ChatGPT Work:Because my workflows and historical decisions lived in plaintext on my machine rather than locked deep inside Claude's proprietary backend, I was completely platform-agnosticNew paragraph, because that is the transferable lesson on this whole page: keep your prompts, context files, and standard operating procedures in plain text you own. Then switching costs you an afternoon instead of a year.
His stated reason for leaving was not features either:The real catalyst for me jumping ship today was growing frustration with Anthropic. The constant teasing of Fable 5's availability, the moving target dates (July 9th, then the 12th, now the 19th), and the resetting limits just felt like being dragged along as a customer.Roadmap trust is a real evaluation criterion. Worth naming honestly on a page that otherwise leans Claude.This is the kind of thing I mean by Cowork winning on the desk. I recorded myself doing the task once, and it drafted the repeatable version, told me which access path it preferred, and asked before saving.
Correction 4: Cowork activity is not captured in the Compliance API
If you are evaluating this for a team rather than yourself, stop here first.
Anthropic's own Cowork documentation states it plainly: "Cowork activity is not captured in the Compliance API at this time."
For an individual professional, that is a footnote. For anyone in a regulated industry, or anyone whose security team requires an audit trail of what AI touched which documents, it is potentially disqualifying, and it is the first question a reviewer will ask.
It does not mean Cowork is unsafe. It means the audit surface your organization may require does not exist for it yet. Check before you roll it out to a team, not after.
None of the pages currently ranking for this comparison mention it. It is the highest-stakes fact on the page.
How this comparison was built
Straight answer on methodology, because it changes how much weight to give the verdict.
No head-to-head first-party test was run, and I will not fake one. I live in Claude Cowork and Claude Code daily; I have not run ChatGPT Work on anything real. So read this page for what it is: a documentation-and-real-users comparison from someone with deep hours on exactly one side. ChatGPT Work also only started rolling out on July 9, 2026, and a real test needs both products in one account doing identical work.
What this page is built from:Primary documentation, linked inline at every claim that matters: Anthropic's Cowork help articles and pricing page, OpenAI's help center for ChatGPT Work, the retired agent mode notice, and the agentic billing article.
Real user threads, named and linked, from r/ClaudeCowork and r/AI_Agents.
My own published Cowork verdicts, linked to their source posts, including the correction to one of them above.What it is not: a benchmark. When a usage number gets published or I run both against the same task, this page gets updated and the date at the bottom changes.
Pick your lineYou already pay for Claude Pro or Max and mostly do document, email, and report work: Cowork. It is included, it is generally available, and the scheduled tasks are the feature that makes it stick.
Your company already runs on ChatGPT and procurement is your real bottleneck: ChatGPT Work. Same job, no new vendor review.
You code and do office work in roughly equal measure: ChatGPT's shared pool across Work and Codex is genuinely more flexible than paying for Claude twice over.
You are on a $20 plan and agent work is becoming daily: the plan is your constraint, not the product. Read Claude Max vs ChatGPT Pro before switching tools, because switching will not fix a sizing problem.
You are evaluating for a regulated team: start with the Compliance API gap above. Feature comparisons are premature until that clears.
You have never used either: start with whichever subscription you already have. They do the same job, and the cost of picking wrong is one afternoon, provided you keep your prompts and context in plain text.For what it is worth, here is how the split actually shakes out in my week, across the two ecosystems: ChatGPT is my well-rounded everyday generalist (general questions, quick lookups, and QA-ing Claude's output), and Claude is where the real work happens: creative work, scripts, content, building and refining the systems. Cowork does the office work, Claude Code builds the software. Pick the ecosystem that matches the bulk of YOUR week and you will not be far wrong.
Common questions
Is ChatGPT Work the same as ChatGPT agent mode?
No. Agent mode is retired. OpenAI's help center states that ChatGPT agent is no longer available and points users to ChatGPT Work for longer, multi-step tasks. Operator was absorbed into agent mode before it was retired, so both older names now resolve to ChatGPT Work.
Does Claude Cowork run on my computer or in the cloud?
In the cloud. Anthropic states that Cowork runs tasks remotely on their servers in an isolated environment. When a task needs a local file or your browser, Claude reaches it through the Claude Desktop app on that machine, which makes the desktop a bridge rather than the place the work happens.
Do I have to pay extra for Claude Cowork or ChatGPT Work?
Neither has a separate subscription. Cowork is included on all paid Claude plans, starting at $17/month annual or $20 monthly for Pro, and draws from that plan's allocation. ChatGPT Work is included with eligible paid ChatGPT accounts but meters credits from an agentic pool shared with Codex, ChatGPT for Excel, and Workspace Agents.
Which uses fewer credits for the same task?
No public number exists. Neither company publishes numeric quotas, and the two products meter in non-comparable units, a plan allocation versus a shared credit pool. Anthropic does say that Cowork's auto mode consumes more of your limit than its other modes, which is the one concrete lever you control.
Should I switch from Claude Cowork to ChatGPT Work?
Only for a structural reason: your company already runs on ChatGPT, you want one agent budget shared with Codex, or you need an audit trail Cowork's Compliance API gap cannot provide. Feature parity is close enough that switching for features alone is not worth the setup cost.Published July 18, 2026. Last reviewed July 2026. Product facts checked against Anthropic's Cowork support documentation, Claude pricing, OpenAI's ChatGPT Work help article, the retired agent mode notice, and OpenAI's agentic billing article on that date. No head-to-head first-party test was run; see the methodology section above. ChatGPT Work was announced July 9, 2026 and was still rolling out at publish time.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Every guide ranking for this query teaches you to build a skill in the first five minutes. Folder, frontmatter, ZIP, done. That order is backwards, and it is the reason most non-coders try Skills once and quietly stop.
The build step is the last step.
My rule for this one: you are going to change your mind, so run the task manually until you stop changing it, and only then write it down.
That is the rule I landed on in the Claude vs ChatGPT for Excel breakdown, and the useful surprise is that Anthropic's own documentation agrees with it, not with the tutorials. More on that in a second, because it is the whole post.I walk through this on camera in The 5 Levels of Claude Code Skills (Most People Are Stuck on Level 2) (17 min).A Skill Is a Folder of Instructions Claude Picks Up Only When It Needs Them
Anthropic's definition is one sentence: Skills are folders of instructions, scripts, and resources that Claude loads dynamically to improve performance on specialized tasks.
The word doing the work is dynamically. Per the same page, when you ask Claude to complete a task, it reviews available skills, loads relevant ones, and applies their instructions.
So a skill is not always-on. It sits there until a task matches, then fires.
That is also why the description field matters more than anything else you write. A skill needs two things in its frontmatter: a name capped at 64 characters and a description capped at 200. Anthropic's skill-creation docs call the description critical, because Claude uses it to decide when to invoke your skill. A perfect procedure with a vague description never runs.
Skills, Projects, and Custom Instructions Do Three Different Jobs
Most of the confusion on this query is people building a skill when they wanted a Project. Anthropic draws the line themselves in the Skills doc: Projects provide static background knowledge that is always loaded, while Skills provide specialized procedures that activate dynamically.What it is
When it loads
Use it forCustom instructions
Standing preferences
Every conversation, always
Tone, format defaults, how you want to be addressedProjects
Static background knowledge
Always, inside that Project
Client docs, brand guide, product spec, past reportsSkills
A procedure with steps
Only when a task matches the description
A repeatable process you run the same way every timeThe quick test: if the answer is stuff Claude should know, that is a Project. If it is how Claude should do a thing, that is a skill. If it is how you like to be talked to, that is custom instructions.
Getting this wrong costs you weeks, because a badly-scoped skill fails silently. It just never triggers.
Anthropic's Own Trigger for Building a Skill Is Retrospective
This is the line that should end the "build your first skill in 5 minutes" genre. From Anthropic's skills documentation: create a skill when you keep pasting the same instructions, checklist, or multi-step procedure into chat, or when a section of your standing instructions file has grown into a procedure rather than a fact.
Read the tense. When you keep pasting. Past evidence of repetition you have already lived through.
The tutorials teach the trigger prospectively: have an idea, encode it, ship it. Anthropic's trigger is the opposite. It requires that you already did the thing enough times to be annoyed by doing it again.
That difference is not academic. Encode a procedure you have not tested and you have not saved time, you have installed a wrong answer that fires automatically and looks confident every time. You will not catch it, because the output will be plausible.
The vendor's own doc is the receipt here. Build second.
Why You Can't Find Skills in Your Settings (the Code Execution Gate)
If Skills is missing from your Claude settings, you almost certainly have not turned on code execution, and no tutorial mentions this.
Skills are available for users on Free, Pro, Max, Team, and Enterprise plans, per Anthropic, so the plan is rarely the problem. The gate is a capability toggle.Free, Pro, and Max: open Settings, go to Capabilities, turn on "Code execution and file creation." Then Skills appears under Customize.
Team and Enterprise: an Owner has to enable it first in Organization settings under Skills. If you are not an Owner, you cannot unlock this yourself. Ask.Yes, a non-technical person has to enable something called code execution. That is the gate, that is what it is called, and flipping it is the whole job.
Where skills work once it is on: claude.ai on web, desktop, and mobile; Claude Code; Claude Cowork; the Microsoft 365 add-ins for Excel, PowerPoint, Word, and Outlook; and the API.
One trap worth knowing before you plan around Cowork. Cowork is a paid surface only, on Pro, Max, Team, and Enterprise per the Cowork docs, which is a narrower list than Skills itself. And per Anthropic's skills docs, Cowork sessions and cloud sessions, including routines, do not read the skills folder on your machine. Both interactive and scheduled Cowork sessions load the skills enabled for your claude.ai account, synced at session start.
Translation: a skill saved locally will not show up in your scheduled Cowork run. It has to live on your account.
Start With the Skills Anthropic Already Built for You
You can get real value out of Skills without ever authoring one, which reframes this entire query for a non-coder.
Anthropic ships built-in claude.ai skills for Excel spreadsheet creation and manipulation, Word document creation, PowerPoint presentation generation, and PDF creation and processing. Partner skills from Notion, Figma, and Atlassian sit in the Skills Directory.
That covers a large share of what actually lands in a normal work inbox. Deck, doc, spreadsheet, PDF.
So the honest first week is: turn on code execution, browse the directory, use the bundled skills on real work, and author nothing. If you never get past this step, Skills still paid for itself.The 60-second version, from my Shorts: Stop installing random Claude skills. These 5 actually hand you the file.The Graduation Ladder: Manual Chat, Then Reps, Then Encode, Then Schedule
The right order is a ladder with four rungs, and every rung has to earn the next one.Run it manually in chat. Paste the input, describe what you want, read the output.
Do it again, several times, and change your mind out loud. This is the rung people skip.
Stop changing your mind. When the format holds across runs, the procedure is finally knowable.
Write it down as a skill, then automate it.My inbox automation is the worked example, and I published the sequence in the Claude Cowork scheduled tasks breakdown. I ran the inbox summary manually first. I started by asking for what I wanted without fully understanding what I actually needed, then iterated on the format run by run: I need this, I do not need this, I like this format. Only once it repeatedly did the job right did I move it onto a schedule.
The manual reps were not wasted time before the automation. They were the spec.
His companion rule from the Excel breakdown is the same shape: start with one simple skill and keep doing everything else manually until you trust it. One skill, fully earned, beats five you are quietly babysitting.
If you are still deciding where any of this fits in your week, the Claude at Work pillar lays out which lane each task belongs in before you automate anything.
The Signal You're Done Iterating
The real tell is not a duration. It is having a clear goal and knowing what good looks like, and the only way you know what good looks like is that you have done the work manually yourself.
Here is how I actually judge it with my own skills: I use the product, so I can read an output and place it. This is good enough. This is on par with what I would have done. This is better than I would have done it. That scale only exists in your head because of the manual reps. Skip them and you literally cannot know when the skill is done, because you have no standard to judge against.
The practical version: after each manual run, note whether you edited the output. When runs stop surprising you and you catch yourself saying "on par with mine," you have the green light.
And be pragmatic about production. You are not going to spend five weeks perfecting a skill while shipping nothing. Get it to good enough for production, ship with it, and keep refining. Realistically that is a week or two of iterating when you have time, and either way the work keeps moving: run it manually while the skill catches up, or refine until it clears your bar. Once it is in production it earns its keep on real output you are QA-ing anyway.
What Actually Goes in Your First Skill
Your first skill should be the most boring repeatable thing you do, not the most impressive.
Good first candidates share a shape: same input type every time, same output format every time, and a judgment call you have already made and settled.The weekly status update you write from the same three sources
The way you reformat a client's messy export before you look at it
The specific structure your boss wants meeting follow-ups in
The checklist you run before sending anything externalWrite it as instructions, not prose. Steps, the format you want, and one line about what to do when input is missing. Then spend real effort on the description field, because that string is what decides whether the skill ever fires.
Three No-Code Ways to Actually Create One
There is a terminal-free upload path, a skill whose job is writing skills, and since July 21, 2026, a way to record yourself doing the task and let Claude write the skill from the recording.
The upload path. Create the skill folder, package it as a ZIP, then go to Customize, Skills, the plus button, Create skill, Upload a skill. One structural requirement that trips people up: the skill folder must be the ZIP root, and per Anthropic's skill guide, files should not sit directly in the ZIP root. Folder in the ZIP, not loose files.
skill-creator. Anthropic maintains a skill that authors skills. Its own description reads: create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill. If writing structured instructions is the part you dread, this is the part you delegate.
Record a skill. This one landed on July 21, 2026, after this page first went up, and it is the most non-coder path of the three. In Cowork on Claude for Mac, click the plus button in the composer and pick "Record a skill" (or go to Customize, Skills, Add, "Record your screen"). Do the task the way you normally do it, talk through it if you want, click Done. Per Anthropic's skill guide, Claude then "starts a Cowork task and reviews the recording, then proposes a skill" for you to approve before it is saved. It is Pro, Max, and Team only, Mac only, and a recording runs about 10 minutes.A test run on my Mac mini: I recorded myself turning a Monday brief into a three-line summary in TextEdit, and Claude watched the 15 steps and proposed a skill that does the same thing with file reads and writes instead. That is the part that matters for this page: it built the skill from a task I had already done by hand, which is exactly the order I am arguing for.
Two cautions from the same doc. Everything on your screen is captured while you record, so no passwords, no private tabs. And the video is not kept; what is saved afterward is a set of screenshots from the session, which you can open under the "Recorded demonstration" step in the task.
One correction while you are here. The tip circulating on LinkedIn, where you open an old Cowork chat and prompt Claude to turn the conversation into a Skill, is not documented by Anthropic anywhere in its support material. It may work. It is not a supported path, so do not build your plan on it.
After the Skill Exists, Verify It Like You Do Not Trust It
A skill that runs automatically is a skill nobody is reading closely, which is exactly when wrong output survives longest.
The habit I keep from the Excel breakdown is trust but verify: run the automated version in parallel with your manual process until it matches your answers, and when a number matters, paste the output into a second tool and ask it to QA the first. Spot-check numbers against the source. Read the whole thing on the first few automated runs even when you are sure it is fine.
Skills do not remove the review step. They remove the typing.
If you want to see what the finished automated versions look like in practice, the 20 real Claude Cowork use cases page is the catalog of what people actually run on a schedule.
When a Skill Is the Wrong Tool
Skills are the wrong call more often than the tutorials admit, and knowing when saves you the most time.One-off tasks. If you will do it twice this quarter, just prompt it. Encoding costs more than the task.
Anything still changing. New process, new client, new format the team is still arguing about. Encode it and you will be maintaining a stale procedure inside a folder you forget exists.
Judgment calls. Deciding what to include is not a procedure. Formatting what you already decided is. Skills automate the second one.
Stuff that is really reference material. If you want Claude to know things, put them in a Project. Skills hold procedures, not facts.
Work where being wrong is expensive and unreviewed. Anything financial, legal, or client-facing that would go out without a human reading it. Keep a person on that.The failure mode is always the same: encoding too early, then trusting output that quietly drifted from what you meant.
Frequently Asked Questions
Do Claude Skills require coding?
No. Anthropic ships built-in skills on claude.ai for Excel, Word, PowerPoint, and PDF work, plus partner skills from Notion, Figma, and Atlassian in the Skills Directory, and using them requires no authoring at all. Creating a custom skill also has a no-terminal path, and since July 2026 a record-your-screen path in Cowork on Mac: package the skill folder as a ZIP and upload it through Customize, Skills, Create skill.
Why is Skills missing from my Claude settings?
Skills sit behind the code execution capability. On Free, Pro, and Max, open Settings, go to Capabilities, and enable "Code execution and file creation," after which Skills appears under Customize. On Team and Enterprise, an Owner has to enable Skills in Organization settings before any individual user can see it.
When should I create a Claude Skill instead of just prompting?
Anthropic's stated trigger is repetition you have already experienced: create a skill when you keep pasting the same instructions, checklist, or multi-step procedure into chat. If you have not run the task manually enough times that two consecutive runs need no edits, prompting is still the right tool and a skill would just lock in a guess.
Are Claude Skills available on the free plan?
Yes. Skills are available for users on Free, Pro, Max, Team, and Enterprise plans, though code execution must be enabled first. Claude Cowork is stricter and paid-only, on Pro, Max, Team, and Enterprise, so a free user can use skills in chat but not run them on a Cowork schedule.
Do my local skills work in Claude Cowork?
No. Anthropic states that Cowork sessions and cloud sessions, including routines, do not read the skills folder stored on your machine. Both interactive and scheduled Cowork sessions load the skills enabled for your claude.ai account, synced at session start, so a skill has to live on your account to run there.
Do This in This OrderThis week: enable code execution, then use the bundled Excel, Word, PowerPoint, and PDF skills on real work. Author nothing.
Weeks two through four: pick your single most repetitive task and run it manually in chat every time it comes up, changing the format whenever it annoys you.
When two runs in a row need no edits: write it down as one skill, with real effort spent on the description field.
Only then: put it on a schedule, and read the first few automated outputs like you do not trust them.
If the process is still changing, or the call is a judgment call: do not build the skill. Keep prompting.The tutorials sell you the last step first because it is the only step that looks like progress. The reps are the work.
Published and last reviewed July 18, 2026. Product capabilities, plan tiers, and setup paths checked that day against Anthropic's official support and developer documentation, all linked inline. The claim that a Cowork conversation can be converted into a skill in one click is explicitly not documented by Anthropic and is flagged as unverified above. These products change often; the official pages are the source of truth.
Until late August 2026 this page opened with "Claude cannot send your email." That is no longer true. Anthropic's Gmail connector now sends, replies and forwards, and asks for your approval before each one by default. ChatGPT has worked that way for a while. So both can hit send, both ask first, and every other page ranking for this query is still arguing about tone.
Every page ranking for this query is arguing about tone. That is the wrong argument.
My rule for this one: tone is the cheapest thing to fix and the loudest thing on the SERP. What matters is which tool can touch your inbox, and how it treats the words you put in it.
Whether you searched "claude vs chatgpt for email," "chatgpt vs claude for gmail," or "can claude send emails," the answer starts with capability and privacy, not warmth. Here is the split, with the official documentation behind each line.
The Split in One Table (updated September 2026)Claude
ChatGPTReads your inbox
Yes, via Google Workspace connector
Yes, via connectorsCreates a draft
Yes, in Gmail
YesSends the email
Yes, since August 2026. Asks for approval first by default
Yes, gated as an "important action"Reads attachments
Metadata only, not content
See OpenAI's connector docs per appTrains on connector-retrieved email
No, stated explicitly
Not by default on Business, Enterprise, EduTrains on email you paste in
Possible on Free, Pro, Max if chat training is on
Possible on Free, Plus, Go, Pro if the setting is onAvailable on the free tier
Yes, connectors are on all tiers
Connector behavior varies by tierIn-browser help
Claude in Chrome, beta, paid plans, labeled risky
Not covered hereSaved email voice
Projects with project instructions
Projects and custom GPTsThe third row is the one most pages still get wrong.
Claude Can Send Now (It Could Not Until August 2026)
Until late August 2026 the Google Workspace connectors documentation said "Claude creates drafts in your Gmail account, but cannot send emails on your behalf," with an FAQ line that read "Can Claude send emails on my behalf? No." That is the version most comparison pages still quote. It changed in late August 2026. The same FAQ line now reads:Yes. Claude can send, reply to, and forward emails from Gmail, and asks for your approval by default before each of these actions.The approval step is the part to hold onto. Ask Claude to reply to a thread and it drafts the reply, shows it to you, and waits. On Team and Enterprise plans an owner decides whether members can let those actions run without asking each time. I covered the switch the week it landed:The 60-second version, from my Shorts: Claude just got a send button for your Gmail.The setup screen makes more sense now too. Google's OAuth screen mentions email sending permissions, and Claude actually uses them. The setting to leave on is the approval prompt.My own connector list in Cowork. Gmail is on, which is all the setup the send path needs; the approval prompt does the rest.
One more limit worth knowing before you build a habit on this: attachment content is not directly accessible through Gmail, metadata only. So "read the contract Sarah sent and reply" does not work end to end. You get the filename, not the file.
Connectors are available on all Claude tiers including free. On Team and Enterprise, an Owner has to enable them for the organization first.
ChatGPT Can Send, But It Has to Ask First
ChatGPT has sent email for a while, and its gate works the same way Claude's does now. Sending is not silent, and OpenAI's own category for it is the useful detail. From the connectors documentation:An important action is an app action that could have a meaningful effect outside ChatGPT... Examples may include: Sending or editing an email, message, comment, post...The default behavior is spelled out on the same page: ChatGPT "uses Important actions, which allows reading from apps automatically but asks before actions that may have a meaningful effect outside ChatGPT."
Read that as a two-speed system. Reading your inbox is automatic. Sending stops and waits for you.
One caveat no competitor on this query mentions. OpenAI's Outlook email and calendar app documentation lists mail sending scopes while describing retrieval as read-only. OpenAI's own doc is inconsistent here, so on Outlook, verify the behavior on a low-stakes message before you trust it with a client thread.
The Privacy Question Is Not "Which Tool," It Is "How Did the Email Get There"
Here is the finding that should change what you do tomorrow morning: on both platforms, the copy-paste habit everyone defaults to has the worst data posture available.
Anthropic's connector page states the protection clearly: "We do not train our models on your Gmail, Drive, or Calendar connector data." Then it draws the line that matters.If you are using our consumer products (e.g. Claude Free, Pro, and Max)... and you have chosen to allow us to use your chats... then any content you copy/paste from your Gmail, Drive, Calendar... may be used to improve our models.That is corroborated on Anthropic's model training policy page. Connector-retrieved email is never trained on. The same email, pasted into a chat box on a consumer plan with training allowed, can be.
OpenAI draws a parallel line on tier instead of method. From the connectors doc: for ChatGPT Free, Plus, Go, and Pro users, OpenAI may use information accessed from apps to train models if your "Improve the model for everyone" setting is on. For Business, Enterprise, and Edu customers, OpenAI does not use information accessed from connectors to train models by default.
So the checkable rule for both:Connector, on a work plan: the strongest posture on either platform.
Connector, on a consumer plan: protected on Claude for retrieved data; on ChatGPT, check your training setting.
Copy-paste into a chat window on a consumer plan with training on: the weakest posture, and it is the one most people are using right now.If your inbox contains client names, salary figures, legal threads, or anything under an NDA, that last bullet is the whole reason to read this section twice. Turn the setting off, or connect the account instead of pasting.
The Verdict Spine: Email Is Generalist Work, Not Creative Work
This is where the decision actually gets made, and it is the same fork the Claude at Work pillar exists to settle: which tool owns which job in your week.
My split across my own week, published in Claude Cowork vs ChatGPT Work, runs like this. ChatGPT is the well-rounded everyday generalist, and it doubles as the QA layer for Claude's output. Claude is where the real work happens: creative work, scripts, content, building and refining.
Email is correspondence. Most of it is generalist work.
And one reframe from my own week, because almost every comparison on this page's SERP assumes "AI for email" means writing emails. My heaviest AI email win is not drafting at all. It is management: my inbox system organizes, archives the junk, unsubscribes, and pings me on the stuff that matters, like a person reaching out or a bill that needs to be paid. Drafting responses is the smaller half. If your inbox is cluttered, the highest-value email automation you can build has nothing to do with prose (the full setup is in Cowork scheduled tasks).
Confirming a meeting, chasing an invoice, answering a vendor, summarizing a thread for your manager: that is ChatGPT's lane by that split, and either tool can now carry it through to send.
The exception is real, though. Some email is writing. The resignation note, the pitch to a decision maker, the apology to a client who is already annoyed, the message where the wrong word costs you something.
Those are Claude tasks by the same split. And the approval prompt before send is a feature there, not friction: nothing irreversible goes out until you have read it.
Run the Cross-Check Before Anything Important Leaves Your Outbox
The rule I landed on in the Excel breakdown applies here without modification: run the output past the other model before it goes out. If you are using ChatGPT, paste it into Claude and ask it to QA it.
For email, the trigger is easy to state.
Anything going to a client, your boss, or anyone who can fire you gets a second-model read before send.
It takes about thirty seconds and it catches the real failure mode of AI email, which is not bad grammar. It is confident tone drift, or a draft that quietly agrees to a deadline you never agreed to. Ask the second model one question: does this commit me to anything I did not intend to, and does the tone match a message to this person.
One catch nobody else will tell you. Pasting the draft into the second tool is exactly the copy-paste path from the privacy section, so on a consumer plan check your training setting first, or run the cross-check on a work account.
Save Your Email Voice Once Instead of Re-Explaining It Every Week
Both tools have the same fix for "it does not sound like me," and it is a container you set up once.
Claude has Projects, where you can define project instructions for each project, including instructing Claude to use a more formal tone. The free tier is capped at 5 projects, so an "Email" project is a reasonable use of one of them.
ChatGPT has Projects too, with a precedence rule worth knowing, per OpenAI's Projects doc: project instructions only apply inside the respective project and will override your global custom instructions. Your email project can be more formal than your default without you editing your global settings back and forth.
Custom GPTs go one step further, and OpenAI's guide to creating a GPT draws the split more clearly than most people apply it.Use knowledge for reference material, not rules or behavior. Put rules, tone, and workflow guidance in instructions.That gives you a concrete build. Your voice rules go in instructions: sentence length, greeting and sign-off, whether you use exclamation points, what you never say. Then 5 to 10 of your actual sent emails go in knowledge as reference material.
Building GPTs is paid-only and web-only, so plan around that.
If you want the longer comparison of those two containers, it is in Claude Projects vs custom GPTs.
One timing warning before you spend an evening on this. My manual-reps rule, also from the Excel breakdown, says do not encode a workflow before you have run it manually enough times to stop changing your mind.
So do not write your email voice instructions on day one. Draft emails the slow way for two or three weeks, keep the corrections you make by hand, and then turn those corrections into the instruction block.
The corrections are the instructions.
Where Claude Genuinely Wins on Email
Now that both tools send and both ask first, the capability table is close to a tie. The win moves to the words.
Claude reads a hard email better. It holds your tone across a long reply, it does not slide into the hedge-everything register that gives AI email away, and when you tell it "shorter, and drop the apology" it actually does that instead of rephrasing the apology.
Leave the approval prompt on. That is your pause: nothing leaves your account until you have read it, no confirmation dialog you clicked past at 4:50 on a Friday. For anyone whose email carries legal, financial, or relationship weight, that prompt is worth more than the two seconds it costs.
Add the editing and refining advantage from the split above, and the cohort is clear: if your hard emails are the ones that matter, Claude is the better email tool for you.
How This Comparison Was Built
No first-party head-to-head test was run for this page, and no screenshots of private inboxes appear on it.
What it is built on instead:Official vendor documentation, linked at every capability claim: Anthropic's Google Workspace connectors page, model training policy page, Projects page, and Claude in Chrome page; OpenAI's connectors page, Outlook app page, Projects page, and GPT creation guide.
Verbatim quotes rather than paraphrase wherever the exact wording is the point, especially on sending and on training data.
My published verdicts on adjacent questions, linked at each claim: the daily-driver split from the Cowork vs Work comparison, the cross-check rule and the manual-reps rule from the Excel breakdown.One flagged inconsistency: OpenAI's Outlook app documentation lists mail sending scopes while describing retrieval as read-only. Outlook users should verify behavior on a test message. Both companies ship changes constantly, and their pages win over this one.
Updated September 5, 2026: Anthropic's connector page flipped its FAQ from "Can Claude send emails on my behalf? No." to "Yes," with approval required by default. The title, table, and the two sections on sending were rewritten to match. I covered the change in a Short on August 25.
Frequently Asked Questions
Can Claude send emails from Gmail?
Yes, since late August 2026. Anthropic's Google Workspace connector documentation now answers it directly: "Yes. Claude can send, reply to, and forward emails from Gmail, and asks for your approval by default before each of these actions." Before that it was draft-only, and most pages on this topic still say so.
Can ChatGPT send an email without asking me?
Not by default. OpenAI classes sending an email as an "important action," which it defines as an app action that could have a meaningful effect outside ChatGPT. The default configuration reads from connected apps automatically but asks before actions like sending, so the message is shown to you first.
Is it safer to paste an email or connect my account?
Connecting is safer on both platforms. Anthropic states it does not train on Gmail, Drive, or Calendar connector data, but says pasted content from those sources on Free, Pro, or Max may be used to improve models if you allowed chat training. OpenAI may use information accessed from apps for training on Free, Plus, Go, and Pro when "Improve the model for everyone" is on, and does not by default on Business, Enterprise, and Edu.
Can Claude read my email attachments?
No. Anthropic's connector documentation states that attachment content is not directly accessible through Gmail, metadata only. Claude can see that a file was attached and what it is called, but it cannot open the contents, so "read the attached contract and draft a reply" does not work end to end.
Do I need a paid plan to connect Claude to Gmail?
No. Google Workspace connectors are available on all Claude tiers including free. On Team and Enterprise plans, an Owner has to enable connectors for the organization before individual users can set them up.
Why does the Google consent screen mention email sending permissions?
Because Claude now uses them. Anthropic's documentation says the OAuth screen mentions sending permissions and that Claude can send, reply to, and forward emails, but only with your explicit approval by default. On Team and Enterprise plans, owners decide whether members can let those actions run without asking each time.
Can Claude write directly inside my Gmail compose window?
Claude in Chrome, in beta on paid plans (Pro, Max, Team, and Enterprise), has built-in knowledge of how to navigate Gmail and other popular platforms, and Anthropic still labels it risky even with safety classifiers enabled. Treat in-browser assistance as an experiment rather than the reliable path. The dependable Claude email surface is the Google Workspace connector, which drafts, and with your approval, sends.
Pick By What You Need the Tool to Actually Do
There is no universal winner here, only a fork based on your inbox.You want the whole loop handled, drafting through to send: either one now. Both send, both ask first by default, so pick by the rest of this list.
Most of your email is routine correspondence: ChatGPT, per the generalist half of the daily-driver split.
Your hard emails are the ones that matter: Claude. Better at voice and refining, and the approval prompt is your pause before anything irreversible.
You handle confidential threads: connect the account instead of pasting, and check your training setting. The method matters more than the logo.
You are on a free plan: Claude connectors work on free tiers, which is the cheapest real path into an AI-assisted inbox.
You want it to sound like you every time: a project with instructions on either tool, built after two or three weeks of manual reps, per the Excel breakdown.Decide how the email gets into the tool, and who reads it before send. Then argue about tone.
Published July 18, 2026. Updated September 5, 2026, when Anthropic's Gmail connector gained send, reply and forward with approval required by default. Capabilities, permissions, and training-data policies checked against Anthropic's and OpenAI's official support documentation, all linked inline. No first-party head-to-head test was run for this page. These products change often; the official pages are the source of truth.
Claude will not take your recording. Not "handles it worse." It does not accept audio or video files at all. ChatGPT will, through record mode, but only on the macOS desktop app, capped at 4 hours, auto-joining Google Meet and nothing else.
So the real question is not which one writes prettier summaries. It is whether you can record at all.
My rule for this one: capture is a hardware problem, minutes are a writing problem, and only one of these two tools can do the first.
Whether you searched "claude vs chatgpt for meeting notes," "chatgpt vs claude for note taking," or "can claude take meeting notes," that is the answer the entire first page of Google forgot to tell you. Here is the evidence, the official limits, and a verdict split by whether a recorder is allowed in your room.
The Split in One Table (July 2026)Claude
ChatGPTAccepts an audio file
No. 10 supported file types, none audio
Yes, via record modeNative transcription
None
Yes, with speaker labelsWhere recording works
Nowhere
macOS desktop app onlyRecording length cap
n/a
4 hours per recordingAuto-joins your calls
No
Google Meet only, not Zoom/Teams/WebexWhich plans
n/a
Plus, Pro, Business, Enterprise, Edu (not Free or Go)Turning messy notes into minutes
Stronger
CompetentScheduled, recurring processing
Claude Cowork
Scheduled tasksEntry price
Pro $17/mo annual, $20 monthly
See OpenAI's pricing pageRead that top row again, because it is the row nobody writes down.
Claude Does Not Accept Audio Files, Full Stop
Anthropic publishes the supported upload list, and it is short: PDF, DOCX, CSV, TXT, HTML, ODT, RTF, EPUB, JSON, XLSX, plus images. Ten document types and pictures.
No mp3. No m4a. No wav. No mp4.
There is no asterisk, no "coming soon," no premium tier that unlocks it. If you drag a recording of your Tuesday standup into Claude, there is nowhere for it to go.
The near-miss that confuses people is voice mode. Claude has one, and it is genuinely useful, but it is conversational input and output: you talk to Claude, Claude talks back. It is not a recorder pointed at a conference table.
To be fair to both apps: talking TO them on mobile is in good shape now. Claude's mobile voice input was a little buggy a few weeks back, and in my use it has caught up; both transcribe what you say to them just fine. That is dictation. It is a different capability from handing the app a recording of a meeting, and the upload list above is still the wall for that.
This is the wall. Somebody pays $17 a month for Claude Pro after reading a listicle that called it "more structured for meeting summaries," opens the app with an hour of audio, and discovers the product will not take the file.
ChatGPT Is the Only One of the Two That Can Capture a Meeting
Credit where it is due, and this is the section where ChatGPT wins outright.
OpenAI's own Record mode doc says it plainly: "ChatGPT can transcribe and summarize audio recordings like meetings." Record mode transcribes, labels who said what, and hands you a summary. On the capture question, Claude is not in second place. It is not on the field.
Now the limits, because every ranking page skipped these too.macOS desktop app only. Not the web app, not Windows, not iOS or Android.
4-hour cap per recording. Your all-day offsite does not fit in one file.
Google Meet auto-join only. Zoom, Microsoft Teams, and Webex are not supported for auto-join.
Plus, Pro, Business, Enterprise, and Edu. Not on Free, not on Go.
Off by default on Enterprise and Edu, so your admin has to turn it on before you can use it.
Audio is deleted after transcription, and it works best in English.Read that list as a filter, not a feature spec. A Windows-based project manager running Teams calls gets nothing from record mode, and no comparison post on this query says so.
Every Ranking Page Is Answering the Second Question First
Search this topic and you get summary-quality opinions. Claude is more structured. ChatGPT is more conversational. Pick your favorite prose style.
That is the second question.
The first question is whether the tool will accept your input, and one of these two answers no. Writing quality only becomes the deciding factor after capture is solved, which is exactly why so many people buy the wrong subscription for this job.
I have written the writing-quality comparison, and the verdict there holds: in Claude vs ChatGPT for writing the pattern from months of real user reports is that Claude holds voice and follows constraints like "one sentence only" more reliably by default, while ChatGPT wins on volume, versatility, and tasks that mix files and data. That is a real advantage for turning a transcript into minutes.
It is just not the advantage you should shop on first.
The Pain Is in the Write-Up, Not the Recording
The people asking this question are not confused about what a summary is. They are drowning in the after-part.
A project manager, u/folkarlow93 on r/projectmanagement, laid out the math: "a 1 hour meeting can take me anywhere from 2-4 hours (sometimes more) to type up minutes for... my manager seem to be able to take minutes and get them sent out 30 minutes after the meeting which I find is insane, almost like I'm doing something wrong."
Four hours of write-up for one hour of meeting, against a manager who ships in 30 minutes. That gap is the whole product category.
An executive assistant, u/AmMF808 on r/ExecutiveAssistants, named the other half of it: "I can take notes during the meeting, but then I feel like I'm only half listening... Some meetings are online, but a lot are in person meetings or quick follow ups where setting up a full meeting bot feels weird."
Two different problems hiding under one search query. One is transcription. The other is attention, etiquette, and what happens to the scribbles afterward.The 60-second version, from my Shorts: Never Write Up a Meeting Again: The 2-Minute AI Workflow After Every Call.The Handoff Between Tools Is Where the Time Actually Goes
Even people who solved capture are still doing manual labor, and they know it.
From r/ObsidianMD, u/Interesting-Post4178 put it better than any vendor page: "The issue is not really capturing the meeting anymore. There are already good tools for transcription and summaries. The part that still feels manual is what happens after... Right now this still feels like a lot of copy/paste between a meeting tool, Claude/ChatGPT, notes, and email."
That is the honest 2026 state of things.
Heads up: the Fireflies links below are affiliate links — a small kickback to me at no cost to you. I only recommend tools I've actually run.
Transcription is a commodity. Otter, Fireflies, Tactiq, and Granola all do it well, and several will sit in a Zoom or Teams call that ChatGPT record mode will not auto-join. The unsolved part is the chain from transcript to decisions to owners to the email that actually goes out.
What I actually do after a call
My path is deliberately boring, and it sidesteps the whole wall: I record and transcribe on the phone itself, with Voice Memos or Apple Notes transcription on the iPhone, and then paste the TEXT into the AI. Easiest thing you can do. Text uploads work everywhere, the transcript is on-device in one step, and I never have to care which chatbot accepts which file type.
I also do not lean on in-app recording for any of these tools. The recording windows are short in my experience (ChatGPT's felt like roughly a ten-minute window when I tried it) and the apps can freeze mid-capture, which is exactly the failure you cannot afford in a meeting. Capture on the device, process in the AI.
For the processing itself, my split: everyday summarize-and-draft work goes to ChatGPT, which is my well-rounded daily generalist, and anything that has to become real content (a script, a post, a doc with my voice in it) goes to Claude. When a summary really matters, I paste one tool's output into the other and ask it to QA it.
The Cohort Nobody Writes For: You Are Not Allowed to Record
Every article on this query assumes a bot is in the room. For a lot of professional life, it cannot be.
In-person meetings. Board meetings where recording is prohibited by policy. Legal, healthcare, and security-restricted workplaces. Two-party consent states. And the case u/AmMF808 named, where spinning up a meeting bot for a five-minute hallway follow-up just feels weird.
If that is you, the Claude no-audio limitation is completely irrelevant.
You are never uploading audio anyway. You have a page of hurried notes, a few half-sentences, three names, and a decision you are pretty sure was made. Your problem is processing, not capture, and processing is exactly where Claude is strongest.
Here is the workflow, and it takes about five minutes.Take rough notes during the meeting. Bullets, fragments, initials for people. Do not try to write minutes live. That is the trap that makes you half-listen.
Paste the mess into Claude within the hour, while you can still fill gaps from memory.
Ask for four things in one prompt: a structured summary by topic, decisions made, action items with an owner and a due date, and open questions that were not resolved.
Make it flag its own gaps. Add "list anything ambiguous or missing that I should confirm" so you get a checklist instead of confident invented detail.
Ask for the follow-up email in a second message, addressed to the attendees, short, decisions and owners only.Step 4 is the one people skip, and it is the one that keeps the output honest.
For meetings you have every week, do not run this by hand forever. That processing step is Cowork's job.
Route It With the Chat, Cowork, Code Rule
The rule I give everyone, from Claude Cowork vs Claude Code: Chat answers. Cowork does your office work. Claude Code builds software.
Pasting one set of notes into a chat window is a Chat task. Fine, do it, it works.
Turning every Monday's notes into a brief without you asking is a Cowork task. Anthropic's own getting-started documentation names this exact use: Cowork can "extract themes, key points, and action items from meeting notes, interviews, or lecture recordings."
This is not theoretical. In 20 real Claude Cowork use cases, use case 9 is people dumping Teams chats, meeting notes, emails, and scattered docs into Cowork on a schedule and getting one readable brief back. Use case 15 is podcast host Aakash Gupta running Cowork across years of his own transcripts to find where guests contradicted each other, then building the ten most quotable insights into a deck.
That is the after-capture lane, already documented, already working.
Two caveats before you go build it. Cowork is on paid plans only (Pro, Max, Team, Enterprise) with no separate SKU, and it burns your usage allowance faster than chat does. Claude also has a Google Calendar connector that can view and create events, which is how you get "schedule the follow-up we agreed to" without leaving the app.
The Third Answer: Do Not Make These Two Your Only Options
The way this comparison is set up is a little rigged, so let me break it.
A dedicated transcription tool plus Claude is a legitimate, often better answer than either chatbot alone. Otter, Fireflies, Tactiq, and Granola exist because capture is a real product problem with real requirements: joining Zoom, Teams, and Webex, running on Windows, handling recurring calendar invites, storing transcripts you can search a year later.
None of that is what a chatbot is for.
Let the recorder record. Then hand the transcript to the tool that writes better, and store the output wherever your work lives, whether that is Notion, Obsidian, or a Claude Project you keep for a given client. That chain beats a purist "one AI does everything" setup almost every time.
How This Comparison Was Built
Straight up: I did not run a head-to-head first-party test for this page. No side-by-side of the two tools summarizing the same meeting, and no screenshots of my own runs.
Fabricating those would be the actual failure, so here is what this page is built on instead.Official vendor documentation, linked inline at every capability claim: Anthropic's file upload list, voice mode page, Cowork getting-started page, and Google Workspace connector page; OpenAI's record mode help article.
Real user threads, named and linked: r/projectmanagement, r/ExecutiveAssistants, and r/ObsidianMD, quoted verbatim above.
My own published verdicts on adjacent questions, linked at each claim so you can check the reasoning.One honest gap: exact dollar figures for ChatGPT's Plus and Pro tiers could not be verified against OpenAI's pricing page at publish time, so this page names tiers instead of prices and links the pricing page for the current numbers. Claude's are on claude.com/pricing: Pro at $17/month billed annually or $20 monthly, Max from $100.
Both companies ship changes constantly. Their pages win over this one.
Frequently Asked Questions
Can Claude take meeting notes?
Claude cannot record or transcribe a meeting. Its official upload list covers 10 file types plus images, with no audio or video format anywhere on it. Claude can turn notes or a transcript you already have into clean minutes and action items, but something else has to capture the meeting first.
Is there a Claude AI note taker?
Not in the sense most people mean when they search that. There is no Anthropic-built bot that joins your call and takes notes for you. Claude is the processing layer that runs after capture, either in chat or on a schedule through Claude Cowork.
Which is better for note taking, Claude or ChatGPT?
For turning raw notes into structured minutes that follow your format rules, Claude. For getting the raw material in the first place, ChatGPT, because record mode is the only native capture path between the two. Many people end up using ChatGPT or a dedicated recorder for capture and Claude for the write-up.
Can ChatGPT join my Zoom or Teams meeting automatically?
No. Record mode auto-joins Google Meet only. For Zoom, Microsoft Teams, or Webex you either record locally with the macOS desktop app in the room, or use a dedicated meeting tool like Otter, Fireflies, or Granola that supports those platforms.
Does record mode work on Windows?
No. ChatGPT record mode is macOS desktop app only as of July 2026. Windows users have no native capture path in either Claude or ChatGPT and should pair a dedicated transcription tool with whichever chatbot they prefer for the write-up.
Pick By Whether You Can Record
There is no universal winner here, only a fork based on your room.You are on macOS, meetings are on Google Meet, recording is allowed: ChatGPT. Record mode is the only native path between the two, and it is included on your Plus or Business plan already.
You are on Windows, or your calls are Zoom, Teams, or Webex: neither chatbot captures for you. Pair Otter, Fireflies, Tactiq, or Granola with Claude for the write-up.
You cannot record at all (in person, board, regulated, or it just feels weird): Claude. Capture stays manual, and the five-step notes-to-minutes workflow above is the whole win.
You have the same meeting every week: Claude Cowork on a schedule. Stop doing the processing step by hand.
Your write-up has to match a strict template or house format: Claude, per the instruction-following pattern in Claude vs ChatGPT for writing.
You are only paying for one and this is a real part of your job: decide capture first, then read the $20 tier breakdown before you enter a card, because the plan you pick changes how much of this work you can actually get done.Solve capture. Then argue about prose.
Published and last reviewed July 18, 2026. Product capabilities and limits checked that day against Anthropic's official support documentation and OpenAI's record mode help article, all linked inline. User quotes are drawn from public Reddit threads, linked at each quote. No first-party head-to-head test was run for this page. These products change often; the official pages are the source of truth.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
You have a 90-page contract, report, or policy PDF and one question you actually need answered. Every ranking page on this query hands you a context-window spec sheet and calls it a decision.
That is a hardware answer to a workflow question.
Three things decide this, and none of them is the size of the context window: where you put the file, whether the PDF is a scan, and whether you are summarizing or judging.
My rule for this one: the bigger window does not read your document, and when being wrong is expensive, run both models against each other before you trust either one.
One note on the window itself, since every other page leads with it. Anthropic's context window doc states that Claude Sonnet 5 supports a 1M token context window on all paid plans, that Opus 4.8, Opus 4.7, Opus 4.6, and Sonnet 4.6 support 500K, and that outside of these models the window is 200K, with no free-tier number published anywhere. OpenAI does not clearly publish context window sizes for its consumer tiers at all.
So the "bigger window" argument compares a published number to an unpublished one. That is one vendor being more transparent than the other, which is a real finding, just not the one those pages think they are making.
What will actually break your afternoon is file size caps, image handling, and retrieval behavior. All three are documented. None appear on that SERP.
Where You Put the File Changes the Rules More Than Which Brand You Picked
The most useful table for this job is not Claude versus ChatGPT. It is chat versus project knowledge, on both platforms.Claude chat
Claude project files
ChatGPT chat
ChatGPT project knowledgeFile size cap
500MB per file
30MB per file
512MB per file, 2M tokens
512MB per file, 2M tokensFile count
Up to 20 per chat
Unlimited, must fit context window
80 files per 3 hours (Free: 3/day)
10 uploaded at a timeWhat gets read from a PDF
Text and visuals under 100 pages
Text extraction only, except multimodal PDFs
Text and visuals (Enterprise: visual retrieval)
Text only, images discardedWhen it flips to search
Automatically at context limit (RAG)
Automatically at context limit (RAG)
Not documented for consumer tiers
Not documented for consumer tiersStorage
Not published per-user
Not published per-user
25GB per user, 100GB per org
25GB per user, 100GB per orgSources: Anthropic's upload documentation, OpenAI's file uploads FAQ, and OpenAI's Projects documentation.
Read the first row again. That is the row this whole post exists for.
The Popular Advice About Big Files Is Backwards on File Size
The standard advice, including the top-voted Reddit answers on this, is that large files belong in the knowledge base rather than in chat. On total corpus volume, that is correct. On a single big file, Anthropic's own numbers say it is exactly backwards.
Anthropic's upload doc puts chat uploads at 500MB per file, up to 20 files per chat. The same page documents project files at 30MB per file.
That is a 16x gap running the opposite direction from the advice.One enormous PDF, used once: drop it straight into a chat. The project will reject it long before the chat does.
Forty documents you will query for months: project knowledge. Unlimited file count there, and the chat's 20-file ceiling becomes the binding constraint instead."Knowledge base for big stuff" sounds obviously right. It is right about how many, and wrong about how big.
There is a second cost the size cap hides. Anthropic specifies project files as "Text extraction only (except for multimodal PDFs)." Your chart-heavy quarterly report can lose its charts by being filed in the place you assumed was more thorough.
The Scanned PDF Trap Is the Most Expensive Surprise in This Comparison
If your PDF is a scan or a photo of a document, this section decides your answer by itself.
OpenAI's file uploads FAQ states it directly: "ChatGPT Enterprise supports Visual Retrieval for PDF files... All other plans and document files only support text-based retrieval. This means that ChatGPT will extract digital text from the file and discard any images."
Their dedicated visual retrieval page repeats it: "This capability is available only to ChatGPT Enterprise customers. It is not supported for ChatGPT Free, Pro, Team, or Edu accounts."
A scanned document has no digital text. Extract the digital text, discard the images, and you are left with nothing.
Claude's documented behavior on the same input: "Claude models can analyze both text and visual elements (like images, charts, and graphics) in PDFs that are under 100 pages." No tier condition attached to that sentence.
Neither vendor uses the word OCR in consumer documentation, so nobody should claim either product does or does not perform it. What is documented is what gets read and what gets discarded, and that is enough to make the call.
But Claude's advantage has an undocumented middle. Anthropic's other threshold is that PDFs over 1000 pages get text only, and between 100 and 1000 pages the documentation says nothing. That band covers most real contracts, annual reports, and policy manuals, which is to say it covers the exact document that brought you here. Any page that tells you what happens in it is making it up.
So test your own file: ask Claude to describe a specific chart or signature block on a specific page. If it can, visuals are being read. If it cannot, you are getting text extraction and should plan accordingly.
The nested version of this trap that will genuinely catch you
Same PDF, same ChatGPT account, two different behaviors depending on how it got there.
OpenAI documents it: "PDFs uploaded as GPT Knowledge or Project Files are processed using text-only retrieval. PDFs uploaded by users during interactions with a published GPT or within a Project conversation are processed using visual retrieval."
So filing a PDF into your project's knowledge strips its images. Dragging the identical file into that project's chat window does not. The tidy habit is the one that costs you the images, which is worth knowing before you commit a workflow to it. Same argument as Claude Projects vs Custom GPTs.
Claude has the same problem, in the project-files line quoted above. On both platforms, the knowledge base is the lossier destination for anything visual.
Why Claude "Forgot" Something That Was Definitely in Your Project File
The most common complaint about long-document work has a documented mechanical answer.
Anthropic explains that when retrieval augmented generation is enabled, "Claude uses a project knowledge search tool to retrieve relevant information... Instead of loading all project content into memory at once, Claude intelligently searches and retrieves only the most relevant information needed to answer your questions."
The trigger is the part that matters:RAG automatically activates when your project approaches or exceeds the context window limits. You'll see a visual indicator.It reverses if the project drops back below the threshold, and Anthropic says it can expand capacity by up to 10x.
At your desk: below the threshold, everything you filed is loaded. Above it, Claude searches instead of reading, and only finds what its search judged relevant to your phrasing.
Same project. Two behaviors. The switch happens on its own.
That is why the clause you know is in there comes back as "I don't see that in the documents." Not a hallucination, a retrieval miss. Ask again using the document's specific language rather than a paraphrase.
One caveat. Anthropic's RAG page says "RAG for projects is available for all Claude plans (free, Pro, Max, Team, and Enterprise)," while their Projects article describes enhanced project knowledge with RAG as paid-plans-only. Two Anthropic pages, same period, opposite claims. Check your own account for the indicator rather than trusting either page.
Summarize This and Find the Clause Are Two Different Jobs With Two Different Risk Levels
Every page on this query treats "work with a PDF" as one task. It is two, and only one can hurt you.
Summarize and orient is low stakes. What is this report about, what are the main findings, what do I need before the meeting. If the model flattens a detail, you find out in the meeting and correct it. Both tools are fine here.
Find the clause and judge it is not. What is the termination notice period, does this indemnity cap cover third-party claims, does this policy apply to me. Being wrong has a dollar figure attached.
That second job is where my trust-but-verify rule from the Claude vs ChatGPT for Excel breakdown applies without exception: the fully autonomous agent does not exist yet. Run the AI in parallel with your manual process until its answers match yours, and keep QA-ing after they do. His cross-check is the specific move for documents: if you are using ChatGPT, paste the output into Claude and ask it to QA it. Then reverse it. AI is already smarter than us at plenty of things and it can still be wrong.
For a contract, that means a concrete loop:Ask model A for the clause, requiring it to quote the exact passage and give the page number.
Open the PDF and read that page yourself. Everyone skips this and it takes ninety seconds.
Paste model A's answer into model B and ask it to QA the interpretation against the same document.
Where the two disagree, that is your list for a human who is paid to be right.Step one does most of the work. A model forced to quote and cite cannot hand you a smooth paraphrase of a clause that is not there.
Two setup habits from that same breakdown make either job go better:One job per session. Do not dump 20 files covering 20 topics and fire off 10 questions. The exception that works is same-domain consolidation. Four contracts from one vendor relationship is one job. Four unrelated policies from four departments is four jobs, and merging them is how you get answers that quietly blend documents.
Decide the output format before you upload. That preference is the actual skill. It separates "summarize this contract," which returns prose you have to re-read, from "give me a table with clause, page number, plain-English meaning, and risk to me."Capture and processing are separate problems here the same way they are in meeting notes, and whether this work belongs in a chat, a project, or something scheduled is the routing question covered in the Claude at Work pillar.
Where Each One Genuinely Wins
Neither is the universal answer, and pages that pick a winner outright are not being straight with you.
ChatGPT wins on volume and throughput. 80 files every 3 hours is a generous documented allowance, and the 2M token per-file cap does not apply to spreadsheets at all. If your PDF work sits next to CSVs, that exemption is real. Storage is published too, at 25GB per user and 100GB per org, which is more than Anthropic publishes on that dimension.
ChatGPT's free tier is honest about being a trial. 3 file uploads per day tells you immediately whether it fits your week.
Claude wins on scanned and image-heavy documents on any tier. Anthropic documents visual analysis of PDFs under 100 pages with no tier condition, while OpenAI restricts visual retrieval to Enterprise. That is the largest documented capability gap between the two for this job.
Claude wins on published transparency. Context windows by model and plan, page thresholds, per-destination file caps. You can plan against numbers instead of guessing.
Both share one weak spot: non-PDF documents. Anthropic states that for DOCX, TXT, and the rest, "Claude extracts text only from these files. If they contain embedded images, Claude won't be able to read or interpret them." Same for Google Drive: "Claude extracts text content only from Google Drive files. Images embedded in documents are not processed." A Word doc full of screenshots is a bad input on either platform.
Frequently Asked Questions
Is Claude or ChatGPT better for PDFs in 2026?
For one large PDF, especially a scanned or image-heavy one, Claude. Anthropic documents that Claude analyzes both text and visual elements in PDFs under 100 pages on any tier, while OpenAI documents that visual retrieval for PDFs is Enterprise-only and all other plans extract digital text and discard images. For a large library of documents you query repeatedly, both handle it, and the deciding factor is where you put the files rather than which brand you picked.
What is the maximum PDF size Claude can handle?
Anthropic documents 500MB per file for chat uploads, up to 20 files per chat. Project files are capped much lower at 30MB per file, with unlimited file count as long as total content fits the context window. Separately, Claude analyzes text and visuals in PDFs under 100 pages and processes text only for PDFs over 1000 pages.
Can ChatGPT read a scanned PDF?
Not on consumer plans. OpenAI documents that visual retrieval for PDFs is available only to ChatGPT Enterprise customers and is not supported for Free, Pro, Team, or Edu accounts. On those plans ChatGPT extracts digital text from the file and discards any images, so a scanned or photographed document has no extractable text and yields nothing usable.
Why did Claude miss something that was in my project file?
Anthropic documents that retrieval augmented generation activates automatically when a project approaches or exceeds context window limits. Below that threshold everything is loaded at once. Above it, Claude switches to searching and retrieving only what it judges relevant, which means the same project can behave two different ways depending on how full it is. A visual indicator appears when the switch happens.
How many PDFs can I upload to ChatGPT?
OpenAI documents up to 80 files every 3 hours, with free users limited to 3 file uploads per day. Individual files are capped at 512MB and 2M tokens, spreadsheets at roughly 50MB. Inside Projects, only 10 files can be uploaded at the same time, and the project count per plan runs from 5 on Free to 40 on Edu, Pro, Business, and Enterprise.
The Verdict, By What Is Actually on Your Desk
No single winner. A fork, decided by the document rather than the brand.One huge PDF you need answers from today: either tool, dropped into a chat window, not filed into a project. Chat takes 500MB on Claude and 512MB on ChatGPT. The project route caps at 30MB on Claude and strips images on both.
A scan, a photo, or anything image-heavy: Claude, unless you are on ChatGPT Enterprise. Consumer ChatGPT discards the images, and images are the entire document.
A library you will query for months: project knowledge on either platform, with the images caveat understood, plus the awareness that Claude switches to search once the project outgrows the window.
A contract you are going to sign: neither one alone. Make the model quote and cite the page, read that page yourself, then cross-check with the other model per the Excel breakdown's verify rule.
You are only paying for one: ask whether your documents are born digital or scanned. That question decides it faster than any context window number.Put the file in the right place. Test whether the visuals are being read. Then argue about which one writes prettier summaries.
Published and last reviewed July 18, 2026. Limits verified that day against vendor documentation linked inline: Anthropic's upload page (itself dated April 22, 2026), Anthropic's context window and RAG for projects pages, OpenAI's file uploads FAQ, visual retrieval FAQ, and Projects documentation. No first-party head-to-head test was run for this page. These numbers drift, and both vendors ship changes constantly, so check the official pages before relying on any figure here.
Both make real, editable .pptx files. That is the part the first page of Google gets wrong.
Several of the slide-tool vendor pages ranking for this query still say the chatbots do not produce slides at all. A MindStudio page dated May 4, 2026 puts it plainly: Claude "is in the same boat as ChatGPT for visual design. It doesn't produce slides."
That was roughly true when it was written. It is not true now, and Anthropic's and OpenAI's own documentation is where you can check.
This is just what happens when a fast-moving product category meets evergreen comparison content. MindStudio itself published a newer piece on June 3, 2026 comparing ChatGPT PowerPoint, Claude, and Gamma that references the PowerPoint add-in and editable decks. The claim got updated. The older page is the one still ranking.
So before you trust any AI comparison, including this one, check the date on it.
The real decision, once you know both make files, is the tier gates, and they invert.
Which lands on the Ship Lean split for slides: the tool makes the file, you make the argument.
Whether you searched "claude vs chatgpt for powerpoint," "chatgpt vs claude for slides," or "can claude make a powerpoint," the premise of the question is out of date. Here is what each one actually produces, which plan unlocks which mechanism, and what to do when you have a deck due Thursday and a company template you have to match.
Both Tools Generate Editable .pptx Files as of July 2026
This is the claim that resets the whole comparison, so here are the primary sources.
Anthropic's file-creation doc: "Claude can create Excel spreadsheets (.xlsx), PowerPoint presentations (.pptx), Word documents (.docx), and PDF files." It does this by executing code to create files directly in your conversations, inside an isolated, sandboxed container.
OpenAI describes ChatGPT for PowerPoint as "a PowerPoint-native AI experience that lives in a sidebar inside Microsoft PowerPoint. It can help you create, edit, understand, and polish presentations directly, while preserving editable slide structure."
Editable slide structure. Not an image of a slide, not a wall of HTML. A file you open, click into, and fix.
Both of those pages are dated. That is the standard you should hold this one to as well.
There Are Three Mechanisms Here, Not Two Tools
The mistake is comparing "Claude" to "ChatGPT" as if each has one slide feature. Claude has two separate ones with different plan requirements.
Claude, in-chat file creation. Claude writes code in a sandbox and hands you a .pptx. Anthropic states that "code execution and file creation is available to all Claude users (Free, Pro, Max, Team, and Enterprise)" on web, desktop, and mobile. Files can save to Google Drive. The cap is 30MB per file.
Claude for PowerPoint, the add-in. An add-in that integrates Claude into your PowerPoint workflow, running in PowerPoint on web, Windows, and Mac. Anthropic states it "is available to Pro, Max, Team, and Enterprise plans." Free is excluded.
ChatGPT for PowerPoint, the add-in. The sidebar described above. OpenAI states it "is available globally to Free, Go, Plus, Pro, Business, Enterprise, ChatGPT Edu, and K-12 users. Free and Go include limited usage access."
Read those three tier lists next to each other and the inversion falls out.
The Tier Gates Invert, and That Is the Actual Decision
A free Claude user can get a .pptx file but cannot get the sidebar. A free ChatGPT user can get the sidebar but is metered.Claude, in chat
Claude for PowerPoint
ChatGPT for PowerPointFree plan included
Yes, all users
No. Pro, Max, Team, Enterprise
Yes, limited usageWhere it runs
Web, desktop, mobile chat
PowerPoint on web, Windows, Mac
Sidebar inside PowerPointOutput
A .pptx file you download
Slides in your open deck
Slides in your open deckEdits an existing deck
Not documented
Yes, pinpoint slide edits
Yes, add or revise slidesReads your slide master
Not documented
Yes, stated explicitly
Hedged, "may not always match"Metered by credits
No
No
Yes, roughly 20 to 110 per taskOther file types
.xlsx, .docx, .pdf too
Slides only
Slides onlyOne caution on that credit row. OpenAI conditions its credit figures on flexible pricing being effective, which makes it rollout-dependent. Check the current help page before you plan a week of deck work around a number from a blog post, including this one.
Only the Add-Ins Have a Verified Claim to Edit an Existing Deck
If your Thursday deck already exists and needs revisions, this is the row that decides it.
Claude for PowerPoint is documented to "make pinpoint edits to specific slides without regenerating entire decks." You select a slide and tell Claude what to change. ChatGPT for PowerPoint is documented to "add or revise slides in an existing deck."
Claude's in-chat file creation is documented for creating files. There is no equivalent published claim that it will take your uploaded .pptx and edit it in place.
That distinction matters more than it sounds. Generating a fresh deck and revising slide 7 of an approved deck are different jobs, and only one of them is what most people are actually doing at 4pm on a Wednesday.
If you need the second one, you need an add-in, which means a paid Claude plan or a metered ChatGPT session.
Every Vendor Says the Same Thing About Company Templates: Start From the File
This is the one point where Anthropic, OpenAI, and Microsoft all agree, and it is the actual technique.
Microsoft's Copilot documentation is the bluntest: "Start with your organization's template. If your organization uses a standard presentation template, start with this file before creating a presentation with Copilot. Starting with a template will let Copilot know that you'd like to retain the presentation's theme and design."
OpenAI's Work documentation says to ask ChatGPT to match an existing file, master deck, or reusable template, and draws a useful line: a reference file applies to one request, a template can be used again. Worth noting for planning purposes, though, that OpenAI's Work documents article states PowerPoint is not included in that Work desktop flow at launch.
So the technique, regardless of which tool you pay for:Open the approved company template file first. Do not start from a blank prompt and hope.
Work inside that file with the add-in, so the tool can see the master, layouts, and theme colors already in it.
Give it the argument, not the design. It has the design.
Check the first two generated slides against a known-good deck before you let it produce twenty.Nobody is going to praise your gradient. They will notice immediately if the logo is in the wrong corner.
Claude Makes the Confident Template Claim, ChatGPT Makes the Honest One
Give each tool its real win here, because they are different kinds of win.
Anthropic's is the strongest sentence either company publishes on this topic, from the add-in doc: "Claude reads the slide master, layouts, fonts, and color scheme in your deck and uses them when generating or editing slides." That is an affirmative, specific claim about the exact thing a corporate template is made of. There is also a nice compounding detail: skills you have enabled in your Claude settings are available inside the PowerPoint add-in too, so any house style rules you have already written travel with you.
OpenAI hedges, in writing, on the same question: "Template adherence: The product is designed to work within existing presentation templates where possible, but generated or edited slides may not always match a preferred style perfectly." It goes further and flags that "some advanced PowerPoint editing, chart, shape, formatting, and slide-management capabilities may be limited or still in development."
Read that hedge as information, not weakness.
A vendor telling you in advance that formatting might drift is more useful than a confident claim you discover is approximate at 11pm. Plan a formatting check either way. ChatGPT just told you to.
ChatGPT's real win is reach. It is in the sidebar on Free, in a product your employer already licenses, on a plan you may already have. Claude's sidebar requires Pro or above.
What Neither One Knows Is What Your Audience Needs to Believe
This is the section that decides whether the deck lands, and no documentation page covers it.
In Claude vs ChatGPT for Excel, my argument is that being technical stopped mattering. Beautiful output and data interpretation are commodities now. The moat is taste, plus your existing procedure for how things should be done.
Slides are the purest version of that. Both tools will produce a clean, on-brand, well-spaced deck. Neither knows that your VP kills any recommendation that appears before the cost slide, or that the number on slide 4 is the one the whole room came to argue about.
There is a transferable pattern in how I handle this on the content side. In my write-up of Claude Cowork scheduled tasks, I describe trying to fully automate long-form content and watching it fail: you sound robotic reading it, you lose track of what you're talking about. What works instead is that the system generates the outline and I supply the substance.
Same shape for a deck.
Let the tool build the structure and the file. You own the argument, the order, and the cut. The failure mode is identical in both cases: a fluent artifact that says nothing you would defend in a room.
One more rule worth stealing from that Excel piece: one job per session. Do not dump twenty files on twenty topics and ask ten questions at once. One deck, one audience, one decision you want made.
If you are still deciding how much of your work belongs in these tools at all, the Claude at Work pillar lays out the lanes.
How to Actually Build the Thursday Deck
Not a feature list. The order of operations.Write the argument first, in plain text. Five to eight bullets, in the order you want the room to accept them. This is the part you cannot delegate.
Open the approved template file. Not a blank deck, not a theme you picked. The file your company sends around.
Open the add-in you have access to. Claude for PowerPoint if you are on Pro or above, ChatGPT for PowerPoint otherwise.
Ask for structure from your bullets, one section at a time, not the whole deck in one shot.
Check slides one and two against a known-good deck before generating the rest. Fonts, logo placement, color, title case.
Revise slide by slide. Both add-ins support targeted edits. Use them instead of regenerating.
Read it out loud as an argument. If you cannot say why each slide exists, cut it.Step 1 is the one people skip, and it is the one the tools cannot do.
If you want the free-tier path instead: ask Claude in chat for a .pptx built from your bullets, download it, then paste your content into the company template manually. Slower, but it costs nothing and you keep the template intact.
Frequently Asked Questions
Can Claude make PowerPoint presentations?
Yes. Anthropic's documentation states Claude can create PowerPoint presentations (.pptx), along with Excel, Word, and PDF files, by executing code in a sandboxed container inside the conversation. This file creation is available to all Claude users including Free, on web, desktop, and mobile, with a 30MB per-file limit.
Does ChatGPT work inside PowerPoint?
Yes. ChatGPT for PowerPoint is a sidebar that lives inside Microsoft PowerPoint and can create, edit, understand, and polish presentations while preserving editable slide structure. It is available globally to Free, Go, Plus, Pro, Business, Enterprise, Edu, and K-12 users, with Free and Go getting limited usage access.
Which one matches my company template better?
Claude makes the stronger documented claim. Anthropic says Claude for PowerPoint reads the slide master, layouts, fonts, and color scheme in your deck and uses them when generating or editing slides. OpenAI hedges, stating that generated or edited slides may not always match a preferred style perfectly. Either way, start from the template file rather than a blank prompt.
Can AI edit an existing PowerPoint deck?
Yes, through the in-PowerPoint add-ins. Claude for PowerPoint can make pinpoint edits to specific slides without regenerating entire decks. ChatGPT for PowerPoint can add or revise slides in an existing deck. Claude's in-chat file creation is documented for creating files, not for editing an uploaded .pptx, so use the add-in when you need to edit.
Do I need a paid plan for AI slides?
It depends which mechanism you want, and the gates invert. Claude's in-chat .pptx creation includes Free users, but Claude for PowerPoint requires Pro, Max, Team, or Enterprise. ChatGPT for PowerPoint includes Free and Go with limited usage, but it is credit-metered, with a typical task consuming roughly 20 to 110 credits.
How This Comparison Was Built
Full disclosure up front: I do not make decks. Nothing on this page pretends otherwise; it is built from the official documentation and named real users, linked at every claim.
No first-party head-to-head test was run for this page. No screenshots of a deck built in either tool.
Every capability claim above is linked inline to the vendor's own documentation: Anthropic's file creation article and Claude for PowerPoint article, OpenAI's ChatGPT for PowerPoint help page and its Work documents page, and Microsoft's Copilot presentation guidance. Where a vendor hedges, the hedge is quoted rather than smoothed over.
The MindStudio quote is verbatim from that page, dated May 4, 2026, and is cited as an example of dated content rather than an error. The same publisher's June 3, 2026 comparison already reflects the newer capabilities, which is worth saying plainly.
Pricing and credit figures move; OpenAI's own page conditions its credit numbers on rollout, so treat the linked pages as the source of truth over this one. Including on the question of whether this page is still accurate when you read it.
Pick By Which Mechanism You Need
There is no universal winner here, only a fork based on your plan and your deck.You are on a free plan and just need a deck to exist: Claude, in chat. It is the only free path to an actual .pptx file, and it also does .xlsx, .docx, and .pdf.
You are on a free plan and need to work inside PowerPoint: ChatGPT for PowerPoint. It is the only sidebar that includes Free, though usage is limited and metered.
You already pay for Claude Pro or above: Claude for PowerPoint. The slide master claim is the most specific one either vendor publishes, and your enabled skills come with you.
You must match a strict corporate template: open the template file first, then use an add-in inside it. That advice comes from Anthropic, OpenAI, and Microsoft alike.
You need to revise an existing approved deck: an add-in, not in-chat generation. Only the add-ins have a documented edit-existing-deck capability.
You are deciding which subscription to hold for work generally: read Claude vs ChatGPT for writing and Claude vs ChatGPT for Excel before you pick, because slides are rarely the only thing on the list.The file is the easy part now. The argument never was.
Published and last reviewed July 18, 2026. Product capabilities, plan tiers, and limits checked that day against Anthropic's, OpenAI's, and Microsoft's official documentation, all linked inline. No first-party head-to-head test was run for this page. These products change often; the official pages are the source of truth.
I pay for both and use both, so here is my honest starting point: this is not really a "which one is better with numbers" debate anymore. A year ago ChatGPT genuinely choked on Excel files. Today both handle a real spreadsheet without drama. They are both smart. What actually decides it is preference and fit: how you want the data extracted, how you want it visualized, and which workflow you are already living in.
The task-level split still exists and it is worth knowing: Claude gets the calculations clean and builds a formatted spreadsheet from scratch; ChatGPT is faster at editing a live sheet in place, handles bigger data workloads, and does not hit a usage wall as fast. Serious spreadsheet users who can justify it end up keeping both and routing by the job.
Whether you typed "claude vs chatgpt for excel" or "chatgpt vs claude for spreadsheets," that is the short answer. The rest is the evidence: the mechanical reason Claude's math comes out cleaner when it matters, where ChatGPT genuinely wins, a decision table, and the working rules I actually use so either tool earns your trust.
First, clear up the Copilot confusion
Before any comparison, one thing has to be nailed down, because the whole SERP is muddy on it.
Excel Copilot is Microsoft's product, not ChatGPT. Copilot lives inside Microsoft 365, uses OpenAI models under the hood, and bills through Microsoft. It is a different tool with different behavior.
When this page says "ChatGPT for Excel," it means OpenAI's own ChatGPT for Excel and Google Sheets add-in, a sidebar that lives inside your spreadsheet. If your test was actually Copilot, you were testing Microsoft, not OpenAI. Keep those three straight: Claude, ChatGPT, and Copilot are three tools, not two.
With that settled, here is what each side actually ships.
What each one actually does with a spreadsheet (July 2026)
The old line was "Claude makes files, ChatGPT edits in Excel." As of 2026 that is out of date. Both now work inside Excel and both can produce files. The real fork is which job each is better at.Claude
ChatGPTNative file creation
Creates real .xlsx with working formulas, per Anthropic's help center, on all plans incl. Free
Builds workbooks too, but leans on its in-sheet add-inIn-Excel add-in
Claude for Excel add-in (paid plans)
ChatGPT for Excel and Google Sheets sidebar (Free, Go, Plus, Pro)Works in place in a live sheet
Yes, via the add-in
Yes, its home turfNumerical calculation habit
Tends to run code to compute
Strong, but more likely to predict a valueVBA / macros
Limited
Limited; OpenAI's own docs say "may not be fully supported"$20 tier
Claude Pro, per claude.com/pricing
ChatGPT Plus, per OpenAI's tiers pageTwo facts worth pinning. Anthropic's file-creation help page states Claude can "create Excel spreadsheets (.xlsx)" and that "code execution and file creation is available to all Claude users (Free, Pro, Max, Team, and Enterprise)." OpenAI's add-in help page describes a "spreadsheet-native AI experience that lives in a sidebar" and can "build, update, and explain spreadsheets directly, including large, multi-tab files."
So neither is missing a basic capability. The difference is quality and fit, which is where the receipts come in.
Claude gets the numbers right, and here is the mechanical reason
The most repeated verdict from people doing real number work is that Claude's calculations come out clean.
An accountant, u/MrNariyoshiMiyagi on r/ClaudeAI, ran both tools on real client tax work and put the Excel verdict plainly: "Claude was phenomenal. The calculations were clean, the new Act was applied correctly, and the MS Excel formatting was genuinely brilliant." Same user, same prompt on ChatGPT, "made a complete mess of the numbers."
That is one person, but the reason it happens is not luck. It is mechanical.
For a calculation, Claude tends to execute code to compute the result, the same way a person would open a calculator instead of guessing. ChatGPT can do this too, but a language model that predicts a total token by token can be confidently, subtly wrong. Code that sums a column is just correct. The Reddit thread "Why does Claude always calculate sums correctly with code" is a whole discussion of exactly this behavior.
The Ship Lean rule for it: when the answer has to be right, you want the tool that runs the math, not the one that predicts it. For spreadsheets full of numbers that other people will trust, that is the edge that matters.
A second receipt points the same way from a different angle. A researcher, u/harpbelle on r/ChatGPT, hit ChatGPT inventing broken macro code and switched: Claude "was so much better, very good at self-diagnosing and troubleshooting until it gets it right." For building a clean, formatted .xlsx from scratch with formulas that actually work, Claude is the pick most number-heavy users land on.
Where ChatGPT genuinely wins for spreadsheets
This is not a Claude infomercial. For a large share of real spreadsheet work, ChatGPT is the better tool, and pretending otherwise would be dishonest.
The clearest counter-receipt comes from a construction estimator, u/Medium-Resort-8048 on r/ClaudeAI, doing exactly the kind of numbers-plus-documents work this page is about: "I'm a commercial construction estimator. ChatGPT is my go to. I put it on thinking mode to be more accurate. I upload the full plans and ask too many questions. Claude would run out of tokens by lunch. As of right now Chat is the way to go."
Read that last line twice. His problem was not accuracy. It was the token wall.
That is ChatGPT's real advantage for spreadsheet work: it keeps going. For a job that means uploading big files and hammering it with questions all day, ChatGPT's more generous practical limits win the whole task, even on prompts where Claude might produce a slightly cleaner single answer.
ChatGPT also wins on in-place editing. Its add-in lives inside the live sheet, so for "clean this column, add a pivot, fix these formulas right here," you are working in Excel, not copying a generated file back and forth. Another user described building full workbooks in ChatGPT, "front page dashboards, charts, pivot tables without any issue," by working in stages.
And a third, u/thiscarecupisempty, uses ChatGPT to "parse excel data and read 20-30 pages of existing documents" and reformat, saying "it does these things flawlessly." When the spreadsheet work is wrapped in reading PDFs, pulling web data, and iterating fast, ChatGPT's all-in-one versatility carries the day.
The real divide: build a file vs work in a sheet
Strip away the brand loyalty and the split is about where the work happens.Build a clean file from scratch: Claude. A formatted .xlsx, working formulas, a summary chart from a PDF's tables. Its code-execution habit keeps the math honest.
Work inside a live sheet all day: ChatGPT. The in-Excel add-in, faster iteration, no token wall on big files.Both can cross over. Claude has an add-in that edits in place; ChatGPT can generate a workbook. But if you route by their defaults, you get the best of each with the least fighting.
One caution that applies to both, because it is the failure mode people forget: neither tool fact-checks its own output. A confident formula can reference the wrong column; a clean-looking total can be built on a bad assumption. Whichever you use, spot-check the numbers that matter before anyone acts on them. The AI builds the sheet; you own the judgment.
VBA, macros, and the honest limits
If your work is heavy VBA or macros, temper expectations for both.
OpenAI's own add-in documentation flags it directly: advanced features "such as VBA and macros may not be fully supported." Claude is not dramatically stronger here either. For automating Excel with macros, both are draft-and-verify assistants, not hands-off builders. Expect to test and fix what they produce.
This is worth naming because a chunk of "AI for Excel" searches are really "write me a macro," and that is the weakest use case for either tool right now.
How this was compared
This verdict is built from public, linked sources, not a private benchmark.
The capability facts come from the official pages, linked inline: Anthropic's file-creation help article and pricing, and OpenAI's ChatGPT for Excel help article and tiers page. Model names and prices were checked on July 17, 2026; both companies change these often, so their pages are the source of truth if a number here has aged.
The user picks come from public Reddit threads where people did real spreadsheet work with each tool, linked where quoted, deliberately including receipts on both sides so this is a comparison and not a sales page.The 60-second version, from my Shorts: Claude Hands You a Real Spreadsheet, Not a Chat Answer.My working rules for spreadsheets and AI (either tool)
These matter more than the brand you pick.
Understand your format first. A lot of this work is not optional uploading: LinkedIn hands you a CSV, X hands you an XLSX or CSV. There is no way around the download-and-upload loop, so decide how you want the data to come back before you ask. A chart? Short, concise insights? A specific output format? That preference, once you know it, is the beginning of a repeatable procedure.
One job per session. Do not dump twenty spreadsheets about twenty different things and ask ten questions across ten topics. It gets confusing fast, for you and for the model. The exception that does make sense: same-domain consolidation, like exporting all your platform analytics and asking for one combined file.
Do the manual reps before you build the skill. I cannot tell you how many times I tried to build the automation before I had even tested the workflow. You will change your mind. Just interact with the AI and ask for the thing until it comes back the way you want, for a few weeks. Then build the repeatable version around what you actually settled on.
Trust but verify, and use one AI to check the other. The fully autonomous agent does not exist for most tasks, and that is fine, because you need the QA step. Run the AI in parallel with your manual process until it matches your answers. And when a number matters, paste the output into the other tool: if ChatGPT did the analysis, ask Claude to QA it. AI is already smarter than us at plenty of things and it can still be wrong. The balance is the skill.
The bigger point for the Excel person: being technical does not matter anymore. Beautiful charts and competent data interpretation are the commodity now, because AI does them at scale. The moat is taste, your procedure, and knowing how to direct the tool. Start with one simple, boring extraction you do all the time, do everything else manually until you trust it, and evolve from there.
Pick by the spreadsheet work you do
There is no universal winner, only a right match for the job.You need clean calculations and correct totals: Claude. It runs code to compute, which is why the math comes out cleaner.
You are building a formatted spreadsheet or financial model from scratch: Claude. Real .xlsx, working formulas, good formatting out of the box.
You live inside a big spreadsheet all day, editing in place: ChatGPT. The in-Excel add-in and no token wall win the whole task.
Your work wraps spreadsheets in PDFs, web research, and fast iteration: ChatGPT. All-in-one versatility beats a slightly cleaner single answer.
You do serious spreadsheet work daily and can justify it: both, routed by task. Claude for the build-and-calculate jobs, ChatGPT for the live-editing and high-volume ones.If you can only run one and your spreadsheets are full of numbers other people will trust, start with Claude. If your spreadsheet work is one part of a mixed, high-volume workload, start with ChatGPT. Then read the $20 tier breakdown before you pay, because the plan you choose caps how much of this you can actually do, and for heavy agentic use the $100 tier decision is the next fork. If you are picking on prose quality too, Claude vs ChatGPT for writing covers that side.
Published and last reviewed July 17, 2026. Capability facts and pricing checked that day against Anthropic's and OpenAI's official help pages, linked inline. User quotes are from public Reddit threads, linked at each quote. Both products change often; the official pages are the source of truth if a detail here has aged.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: yes, Claude Cowork scheduled tasks now run remotely, which means they fire on their schedule even when your laptop is asleep and the app is closed. The one exception is a task that needs your local files or apps, which still has to run on your awake machine.
That single fact is what makes recurring Cowork tasks actually dependable, and it is also the thing half the articles on this page are still wrong about.
If you searched "claude cowork scheduled tasks," "do cowork tasks run when my computer is closed," or "how to schedule a task in claude cowork," this page settles it: the closed-laptop question first, then the 60-second setup, then what to schedule so the feature earns its keep.I walk through this on camera in 8 Claude Routines That Run My Business While I Sleep (20 min).Yes, they run with your laptop closed (with one exception)
This is the whole ballgame, so it goes first.
A lot of pages you will find on this exact search still say scheduled tasks "only run while your computer is awake." That was true at launch. It is not true now.
Anthropic's current help center is explicit. From the scheduled tasks help article: "Scheduled tasks run remotely, so they run on their cadence even when your computer is asleep or the Claude Desktop app is closed." Their cross-device help page says the same thing a second way: scheduled tasks "no longer need your computer to be awake."
So the morning brief lands whether or not you opened your laptop. That is the difference between a nice demo and a system you can rely on.
Here is the exception, printed in the same help article, because it matters: "If a scheduled task requires local files or apps, it will only run locally." A remote run happens on Anthropic's servers, so it can reach your connectors (email, calendar, Slack, Drive) and files saved to your Claude account, but it cannot reach a folder sitting on your hard drive. Point a task at your Downloads folder and it becomes a local task again, which means your machine has to be awake for it.
The clean rule: connectors and cloud files run in the cloud, your local folders run on your desk. My shorthand for it is "cloud in, cloud out; local files, local machine." Design your recurring tasks around connectors and you get the closed-laptop magic. Design them around a local folder and you are back to leaving the lid open.
How to set one up (about 60 seconds)
You do not write any code. Two paths, both in the app.
Open the Scheduled section in the left sidebar, then click New task in the upper right. From there:Set up manually. Enter a task name, the prompt (what you want it to do), an approval mode, a frequency, and optionally a model and a folder. Click Save.
Create with Claude. Claude asks you a few multiple-choice questions, then proposes a task name, a schedule, and what it will do. You click Schedule to confirm.Frequency options, straight from the help article: hourly, daily, weekly, on weekdays, or manually. That covers almost every real recurring chore.
One detail that trips people up: each scheduled run is its own fresh Cowork session. It does not remember your last chat. So everything the task needs to know goes in the prompt, in a connected tool, or in a file it can open. Write the instruction as if the task has never met you, because on every run, it hasn't.
You manage the whole thing from that same Scheduled screen: review upcoming and past runs, edit the instructions or schedule, pause a task, resume it, delete it, or run it on demand when you do not want to wait for the next cycle.
The first thing I scheduled: my inbox
Hands down, inbox. It organizes, archives the junk, unsubscribes, and pings me about the stuff that matters: this person reached out, this bill needs to be paid.
But here is the part worth copying, because it is the trust path, not the task: I ran it manually first. I described what I wanted without fully understanding what I actually needed, let it run, and iterated. Okay, I need this part. I do not need that. I like this format better. Once it did the job right repeatedly, and only then, I put it on a schedule and stopped babysitting it. It eventually graduated off Cowork entirely onto my always-on machine, but Cowork was the proving ground.
Run it manually until the output stops surprising you. Then schedule it. That order is the whole trick.
What to actually schedule (the recurring stuff that compounds)
The mistake is scheduling something clever. Schedule something boring that you do every week anyway. This is the category that turns Cowork from a tool into staff, and it is the one most articles skip.
Anthropic's own launch note framed the sweet spot well: "You set it up once, a morning brief, weekly spreadsheet updates, Friday team recaps, and Claude handles it automatically." Real users converged on the same handful of jobs.The recurring job
How often
Runs closed?Morning email + calendar brief
Daily / weekdays
Yes (connectors)Weekly metrics or content review
Weekly
Yes (connectors/cloud files)Monday expense report from forwarded receipts
Weekly
Yes (connectors)7am digest of your industry sources
Daily
Yes (web + connectors)Sort/rename a local Downloads folder
Weekly
No (local files, needs awake machine)Monthly LinkedIn profile health check
Weekly / monthly
Yes (web)Notice the split maps exactly to the exception above. The jobs that run with the lid closed all pull from connectors or the web. The one that needs an awake machine is the one pointed at a local folder.
For the full list of what people run, 20 real Claude Cowork use cases has the receipts, and every recurring one there is a scheduled-task candidate.The 60-second version, from my Shorts: 3 FREE Claude Tasks That Run Your Business On Autopilot.The best scheduled task: a weekly review that does the research first
The single recurring task worth copying is the one I run myself, documented in the Cowork use cases breakdown: a weekly content and metrics review. Mine answers three questions: how did the content perform, are we trending the right direction, and why or why not.
Here is that real weekly audit, the actual artifact from my system, verdict-first and graded against my own playbooks:The part that makes it good, the rule I gave in that post: have it do the research BEFORE proposing changes, because sometimes the honest answer is "nothing is broken, keep going, put in the volume." A scheduled review that jumps straight to "here are 5 changes" every week trains you to thrash. One that checks the data first and is allowed to say "no change needed" is the one you keep.
That is the difference between a scheduled task that adds noise and one that removes it.
The gotchas nobody warns you about
Scheduled tasks are genuinely good, but three real snags show up in the community threads, and they are worth knowing before you rely on one overnight.
Fresh session means re-auth on some tasks. One user set up a daily task that scrapes an authenticated site in the browser and hit a wall: "because scheduled sessions start in a new chat window it prompts for browser access every time, which defeats the purpose of scheduled/automated runs." Anything that needs a live browser login is a weak fit for unattended scheduling right now.
Silent overnight failures. Another user chaining tasks flagged the honest hard part: "what happens when a scheduled agent fails silently at 3am. Still figuring out reliable error propagation." Anthropic's help article does not document retry behavior, so treat a critical task as "check the result," not "trust it blind." Read the past-runs list on the Scheduled screen for anything you actually depend on.
It burns your plan faster than chat. Anthropic says this directly in its consumption guide: "A single Cowork task or Claude Code debug session can consume many more tokens than chat." A task running every morning is doing multi-step work every morning. On the $20 Pro plan, a few heavy daily tasks will meet the weekly cap. That is plan sizing, not a bug, and it is the exact signal covered in Claude Max vs ChatGPT Pro.
What I refuse to schedule: content
The place scheduling has burned me every single time is content. I used to fantasize about a fully automated content pipeline: I put everything in, it generates, and my only job is to review at the end. It does not work that way. And honestly, it is not as fun. It is mindless.
I tried scripting entire 20 to 30 minute videos. You sound robotic reading it, and you lose track of what you are even talking about. What actually works is the inverse: the system generates the outline and the research, and I riff the substance. I wing it in a controlled way. I enjoy it more and I know what I am talking about. My newsletter gets repurposed from that long-form, so it is already in my voice. Same pattern on Substack: I pick the topic, I riff, I approve the outline and the research.
The rule that came out of all that burn: creating content has to be more me; the system's job is to help me create it faster. That is the difference between AI slop and having a ghostwriter partner who understands your vision. Schedule the deterministic stuff (pull the analytics, scrape the competitors, prep the research) and keep the substance human.
Cowork or Claude Code for scheduling?
If you write code, you might wonder whether to rig this up in a terminal instead. For recurring office chores, you almost never should.
My take, from Claude Cowork vs Claude Code: scheduled tasks are cleaner in Cowork than anything I have rigged up in the terminal, and I live in Claude Code all day. You define the cadence once and it just runs, no cron file, no script, no install.
The line that holds: Cowork does your office work, Claude Code builds software. A weekly report is office work. Keep it in Cowork.
What I would schedule first
If you have never made one, do not overthink the first task.Easiest win: a daily or weekday morning brief that reads your email and calendar and hands you the day. It runs on connectors, so it fires with your laptop closed.
The compounding win: a weekly review that checks your real numbers and is explicitly allowed to say "nothing is broken." Research first, changes second.
The reminder that works for me: my content day is Friday, so a Friday task kicks off the first steps on its own. It pulls candidate topics and starts the research without waiting for me. I cannot forget content day, because the work is already sitting there when I show up. If you batch anything weekly, give the batch day a task that runs the boring first hour for you.
Skip for now: anything pointed at a local folder or needing a live browser login, until you are okay leaving your machine awake for it.Build one, watch its first two runs on the Scheduled screen, then add a second. Boring tasks, delegated once, compound forever.
If you want ready-made prompts that make good first scheduled tasks, grab the 15 workday AI prompts at /start. Most of them run fine on the $20 tier and turn into recurring tasks in one click.
How this was checked
Product facts here are from Anthropic's official help center, linked inline at each claim: the scheduled tasks article, the cross-device article, and the consumption guide. User reports are public Reddit threads, linked where quoted.
A few things Anthropic does not document, so this page does not claim them: exact timezone/DST handling, retry behavior on a failed run, and whether scheduled-task completion pushes a notification. Where users report a gotcha, it is labeled as a community observation, not official behavior.
Published and last reviewed July 17, 2026. Product facts checked against Anthropic's support center on that date. Cowork's scheduled tasks are available on all paid plans and, per the current help center, run remotely even when your computer is asleep, with local-file tasks as the documented exception. These products change often; the linked official pages are the source of truth.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
For writing, Claude is the better pick when the words have to sound like a specific person and follow your instructions exactly. Pick ChatGPT when you need volume, versatility, or writing bundled with research and data in one tool. Most people who write seriously and can justify it end up keeping both and routing by task.
Whether you typed "claude vs chatgpt for writing" or "chatgpt vs claude for writing," that is the short answer. The rest of this page is the evidence: what real users say after months on each, the current model names and prices as of July 2026, a decision table, and a verdict by the kind of writer you are.
The Quick Answer, by What You Write
Skip ahead if you already know both tools. Here is the split most side-by-side users land on.What you are writing
Better pick
WhyIn your own voice or brand voice
Claude
Holds tone with less prompting; reads warmer by defaultLong documents and reports
Claude
Keeps the thread and follows document-wide instructionsStrict-format or constrained writing
Claude
Obeys "one sentence only" and similar rules more reliablyResearch-heavy or data-driven pieces
ChatGPT
Stronger at pulling, calculating, and citing inside the draftHigh volume, fast iteration
ChatGPT
More generous practical limits, more general-purposeWriting plus files, PDFs, and math
ChatGPT
Better all-in-one for mixed tasksBoth tools are good now. This is not one being broken and one being magic. It is a difference in defaults, and defaults matter when you write every day.
The Models You Are Actually Comparing (July 2026)
Comparisons go stale fast because the model names change. Here is what each side ships right now.Claude
ChatGPTWriting flagship
Opus (warmest voice), with Sonnet, Haiku, and Fable
GPT-5.5 (Instant and Thinking)$20 tier
Claude Pro, per claude.com/pricing
ChatGPT Plus, per OpenAI's tiers pageDefault writing feel
Natural, holds voice, follows constraints
Competent, versatile, more generic before tuningVoice control feature
Styles
Custom instructionsAnthropic's own pricing page lists "Write, edit, and create content" as a core Claude capability, and the Pro tier runs $17 per month billed annually or $20 monthly:If either name is new to you, start with the plain-English money comparison first: Claude Pro vs ChatGPT Plus breaks down the $20 tiers, and Claude Max vs ChatGPT Pro covers the $100 ones. This page is only about the writing.
Claude Writes More Like a Person Out of the Box
The most repeated verdict from people who paid for both and wrote with both is that Claude sounds less like AI by default.
In a four-month "paid for both" comparison on r/ClaudeAI, one user put the writing call plainly: for long-form writing, analysis, and structured documents, "claude wins... its not close".
The same post named the reason writers care most about, which is obedience: "tell it 'respond in 1 sentence' and it actually does. gpt-5 negotiates." That single line explains a lot of the frustration writers report with ChatGPT drafts that quietly ignore a length or format instruction.
A reply from u/Ashtonator28 in that thread became a clean way to remember the split: "anthropic accidentally built a brilliant coworker. openai accidentally built a very competent butler," per the r/ClaudeAI thread.
Voice is the word that matters here. Claude's Opus model reads a little warmer, and it tends to hold a consistent tone across a long piece instead of drifting into the flat, hedge-everything register that gives AI writing away.The 60-second version, from my Shorts: Claude vs ChatGPT for Writing (Use Both Like This).Where ChatGPT Genuinely Wins for Writers
This is not a Claude infomercial. ChatGPT wins real writing situations, and pretending otherwise would be dishonest.
The honest counterpoint comes from a real-estate and tax professional, u/Free-Writer-1123 on r/claude, who pushed back on the Claude hype after using both on actual client work: "ChatGPT has been the most-versatile with higher correct rate of answer. Better at reviewing text from PDF... I keep reading how awesome Claude is, but I have much different user experience than most."
That is the pattern. When the writing is wrapped around other work, reading a PDF, running the numbers, pulling research, then dropping it into a draft, ChatGPT's all-in-one versatility often wins the whole task even if Claude would win the prose alone.
ChatGPT also wins on sheer volume. Its practical limits run more generous, so if your job is to produce a lot of drafts fast and iterate, you will hit fewer walls. For a marketer pushing out ten variations of ad copy before lunch, that matters more than a slightly warmer sentence.
And the gap in voice narrows with effort. A saved custom instruction or a strong reusable style prompt pulls ChatGPT much closer to a personal voice. The difference is mostly at the default setting, with no tuning, where Claude starts ahead.
Instruction-Following Is the Real Divide
Strip away the vibes and the most measurable difference for writers is this: Claude does what you told it, and ChatGPT negotiates.
Writers feel this constantly. You ask for three sentences and get five. You say "no bullet points" and the next draft is a bulleted list. You set a tone and it slips back to corporate-neutral by paragraph four.
The reports lean one way. An accountant, u/MrNariyoshiMiyagi, ran both on real client work and found Claude cleaner on a strict task: "Claude was phenomenal. The calculations were clean... ChatGPT, on the same prompt, made a complete mess." That is numbers, not prose, but it is the same underlying trait: Claude holds the constraint you set.
For writing, the constraint is usually voice, length, or format. If your drafts live or die on those, the tool that obeys them saves you the most editing time.
What Neither Tool Does For You
Both of these get quoted as if they replace a writer. They do not.
Claude reads warmer, but it still produces confident filler when it has nothing real to say. ChatGPT is versatile, but it will happily invent a citation or a statistic in a research-heavy draft. Neither one fact-checks itself.
The teacher u/FATJIZZUSONABIKE, in the same r/claude discussion, praised Claude's output in that thread as "so detailed and inspired that I've barely had to adapt anything," but the operative words are "barely" and "adapt." Even the best case is a strong first draft you still own and edit.
The rule that survives every model update: the AI writes the draft, you write the judgment. Whichever tool you pick, the last pass is yours.
How This Was Compared
This verdict is built from public, linked sources, not a private benchmark. The picks come from multiple "paid for both" threads on r/ClaudeAI, r/claude, and r/ChatGPTPro where users wrote real work with each tool over weeks or months, plus the official product pages for current model names and pricing.
Model names and prices were checked on July 12, 2026 against claude.com/pricing and OpenAI's official tiers page. Both companies change these often, so their pages are the source of truth if a number here has aged.
No coding benchmarks were used, on purpose. SWE-bench and Elo scores say nothing about whether a paragraph sounds like you.
Pick By The Writer You Are
There is no universal winner, only a right match for your work.You write in a distinct voice (brand, personal, fiction): Claude. It holds tone with the least prompting and reads warmer out of the box.
You write long, structured documents: Claude. It keeps the thread and follows document-wide rules.
Your drafts must obey strict format or length rules: Claude. It stops negotiating with your instructions.
Your writing is wrapped in research, data, or PDFs: ChatGPT. The all-in-one versatility wins the whole task.
You produce high volume and iterate fast: ChatGPT. More generous practical limits, fewer walls.
You write seriously every day and can justify it: both, and route by task. Claude for the voice-critical drafts, ChatGPT for the research-heavy and high-volume ones. Many writers who tried to pick one ended up keeping both for exactly this reason.If you can only run one and your work is mostly writing where tone and instruction-following matter, start with Claude. If your writing is one part of a mixed workload full of files, data, and research, start with ChatGPT. Then read the $20 tier breakdown before you pay, because the plan you choose changes how much writing you can actually get done.
Published and last reviewed July 12, 2026. Model names and pricing checked that day against claude.com/pricing and OpenAI's official tiers page. User quotes are drawn from public Reddit threads, linked inline. These products change often; the official pages are the source of truth.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: the $100+ tiers are for people whose AI builds things while they do something else.
If your AI use is chat, even heavy chat, keep your $20 plan. If you run agents (coding sessions, automated workflows, research swarms), you will meet the caps, and the upgrade question answers itself.
I pay for both stacks: Claude Max at the 20x step, and ChatGPT Pro at $100 after dropping down from $200. The one-liner I stand behind: Claude Max is for building. ChatGPT Pro is for volume. The upgrade trigger is agents, not chat.
Searching "claude max vs chatgpt pro" or "chatgpt pro vs claude max"? Here is the comparison from someone actually paying for both, with the current numbers as of September 2026, because most articles on this SERP are stale on a big one.
Comparing Claude's own two tiers instead? That is a different question, and it has its own page: Claude Pro vs Claude Max.
The fact most comparisons get wrong: the ladders match on price, not on access
These two ladders were the same shape when this page went up in July, and this page said so. As of September 3, 2026, they match on price but not on access, and the thing that broke it is worth more than the price comparison.Claude Max
ChatGPT Pro$100/month
5x the usage of Pro ($20)
5x the usage of Plus ($20)$200/month
20x the usage of Pro
20x the usage of PlusWhat's included
Everything in Pro: Claude Code, Cowork, Design, Research, projects, plus higher output limits and early feature access
Pro models, Codex, Deep research, image creation, memory, file uploads"Unlimited"?
No. Session + weekly caps
No. 5x/20x metered, per OpenAI's own help pagePublished exact quotas?
No numeric quotas on the official page
Yes, now: GPT-6 Pro is published per weekNewest model, $100 tier
Fable up to 50% of weekly limits
50 GPT-6 Pro messages/week, shared with Sol ProNewest model, $200 tier
Fable up to 50% of weekly limits (same)
200 GPT-6 Pro messages/week, plus separate Sol ProAnnual billing?
No. Monthly only
No. Not currently supportedContext window
200k on every individual plan (1M is an API option, not a subscription one)
200kAds?
None
None. Ads are Free and Go onlyChatGPT Pro was famously "the $200 unlimited plan." That era is over: OpenAI's help center describes Pro as a 5x or 20x usage ladder, the same shape as Claude Max. Articles still selling "unlimited GPT" are describing last year's product.
Look at the two newest-model rows though. That is the split.
ChatGPT Pro now separates its own two tiers by 4x on the newest model. Claude Max does not. On Claude, Fable is capped at 50% of your weekly limits whether you pay $100 or $200, so the tier only changes how big that weekly pool is. On ChatGPT, the tier changes the frontier-model allowance directly: 50 messages a week against 200.
The pricing ladders still mirror each other. What you get at the top of each one no longer does.
Neither company lets you pay annually at the top tier. OpenAI states it plainly: "Currently, we do not support annual billing or the option to pay for multiple months in advance for ChatGPT Go, Plus, or Pro subscriptions." Anthropic's pricing FAQ says of the two Max options: "Both options are billed monthly."
So it is $100 or $200 every month on both sides, with no discount for committing. Claude Pro has an annual option at $17 a month.
The top tiers do not.
And you cannot negotiate your way past a cap. OpenAI: "There is no setting to increase or bypass a model's usage allowance."
One real feature gate exists on the Claude side. Fable 5 and Fable 5.1 are included as standard on Max up to 50% of your weekly limits, cost separate usage credits on Pro, and are unavailable on Free. That is the only place in this comparison where money buys a capability rather than capacity.
Also worth knowing: for general usage neither company publishes exact numeric quotas, whether you search for "claude max vs chatgpt pro" or "gpt pro vs claude max" and land on someone claiming otherwise. OpenAI has now published hard numbers for its Pro models specifically, which is new, and the section below has them. Everything else stays an opaque meter.
What changed on September 3: GPT-6 Astra and a 4x split inside ChatGPT Pro
OpenAI shipped GPT-6 Astra on September 3, 2026. In the ChatGPT model picker it is branded GPT-6 Pro, and OpenAI published the per-week allowance for it, which it does not do for general usage on either side of this comparison.Three things follow from that table, from OpenAI's help center:Pro $100: 50 GPT-6 Pro messages a week, in one shared pool with GPT-5.6 Sol Pro. Switching models does not buy you more.
Pro $200: 200 GPT-6 Pro messages a week, a separate 170/day for Sol Pro, and an automatic fallback to GPT-5.6 Thinking at Medium when you hit the wall.
Astra usage is included in your existing allowance, with credits purchasable on top.Claude Max has no equivalent. Fable sits at 50% of weekly limits on both Max steps, so Anthropic's tiers still differ only by capacity.
That is the cleanest way to say what changed: ChatGPT Pro's tiers now differ in what you can reach. Claude Max's tiers still only differ in how much. I broke the ChatGPT side down properly in ChatGPT Pro $100 vs $200, and the plan-by-surface question is in which ChatGPT plans get GPT-6 Astra.
Has Astra moved my own split? Not yet. It landed yesterday and I have run it on a few real tasks; it is genuinely more autonomous than anything I have used from OpenAI, but Claude is still where my building happens. Ask me again in a month.
The "20x" argument you should know about before you pay $200
This is the part I would want told to me before spending $200 on the Claude side, and it is currently the loudest complaint in the Claude community.
Anthropic's multipliers are per session, not per week. Its own pricing page says so:You get everything in Pro, plus more usage: choose from 5x or 20x the usage of Pro per 5-hour session.Read that carefully. The 20x is a 5-hour session multiplier. Paid plans then have weekly caps layered on top, and those weekly caps do not scale by the same factor.
So the practical weekly throughput a Max 20x subscriber gets is not 20 times a Pro subscriber's. There is now litigation over it: Kahn v. Anthropic, filed June 14, 2026 in the US District Court for the Northern District of California, alleges the weekly figure lands closer to six to eight times for Max 20x and nearer three and a half for Max 5x. Those numbers are an unproven allegation in an active case, not a finding, and I am not going to pretend to know the real multiplier.
But the mechanism is not in dispute, because Anthropic publishes it.
I am not telling you to skip Max. I pay for the 20x tier and it earns its place in my week. I am telling you that "20x" describes a five-hour window, and if you budget your week around it you will be disappointed. OpenAI's numbers, for all the metering, are currently the more literal ones.
The real upgrade trigger: agents burn tokens, chat does not
Here is the pattern across every "why am I hitting limits" thread, and my own bill.
Chat barely moves the meter. Agents devour it. One user on the $200 Max tier, u/jayplay90, reported: "I barely did anything and it ate 24% of me weekly usage on Claude max 20x." Other users report burning half a weekly allowance in a single day without writing any code.
That is not malfunction. That is what agentic work is: one instruction fans out into hundreds of model calls.
My own upgrade moment was exactly this. It started when I realized Claude Code does far more than code: it edits my YouTube videos. I was spending 20 to 60 minutes per video cutting repeated takes by hand.
Now an agent transcribes the raw footage, reads the transcript, decides which takes are keepers, plans the cuts, executes them with FFmpeg, and drops the finished video in a folder.
I say "edit" and step away. Out of my last 14 videos, maybe 3 needed tweaks. Here is this week's actual output folder:Building that system is what pushed me up the ladder. Refining a workflow means testing, rebuilding, and re-running it over and over. Add competitor research with three or four sub-agents running simultaneously, and you burn through a $100 allowance fast.
I kept hitting the cap, got tired of waiting for resets mid-build, and moved to the $200 tier. Straight from my phone:One discipline that stretches any tier further: which model you run matters more than which tier you buy. A $100 Max user who rarely hits limits, u/xAdakis, put it plainly: heavy users burning out their caps "are more than likely using Opus, which is more token heavy... I rarely hit my limits using Sonnet all the time for heavy coding and agentic workflows." Run the efficient model as your daily driver and save the heavyweight for the work that deserves it.
The dial I had been ignoring: Opus 5 and effort settings
This page went up on July 12. Opus 5 shipped on July 24, described in Anthropic's announcement as a step change for the Opus tier powering long-running agents. On the OpenAI side, the GPT-5.6 family led until GPT-6 Astra shipped on September 3 (see the section above); OpenAI previewed an Ultrafast mode for GPT-5.6 Sol on August 13.
So does that change the "Claude builds, GPT reviews" split? Not the split. Something else.
Opus 5 is genuinely good, and I only really understood how good after a week or two with it. The thing I had wrong was the effort setting. I was running it on medium. Medium to high is a big difference.
A real example. My YouTube descriptions were coming out slightly off. In the past I would have reached for Fable, the top model, for any skill tweaking. This was minor, so I went back and forth with Opus on high instead, and it caught the problem and fixed the skill.
Another time I had Opus generate a plan and handed it to Fable for QA. Fable agreed with it.
Run high, and Opus is arguably matching the top model on some of this work. What actually changed in my week is that I moved my content builds, video scripts and LinkedIn posts, from Opus medium to Opus high.
The division still holds: Claude executes, GPT reviews and researches. What changed is the effort dial, and it is the knob most people have not touched. Same model, same subscription, noticeably different output.
Where each one earns its $100+
Claude Max: the builder and the writer. For work that needs judgment and instruction-following, the receipts keep coming from non-developers. One accountant, u/MrNariyoshiMiyagi, ran both $100 tiers side by side on real client work: "Claude was phenomenal. The calculations were clean... ChatGPT, on the same prompt, made a complete mess of the numbers." And for content, my own verdict is simple: if you are creating in your own voice, Claude just writes better. Both models feel human to talk to now; Opus runs a little warmer.
ChatGPT Pro: the volume machine and the orchestrator. The practical limits run more generous, which is why users paying $200 on both sides report leaning on GPT for sheer volume.
That trade shows up cleanly in one user's summary on r/Anthropic, comparing the two $100 tiers: "Usage on chatGPT is much more generous with the latest price cuts. But claude is definitely less hands on for me, i find GPT requires a lot of [steering]."
More volume, more steering. That is the honest shape of it.
In my stack, ChatGPT's side (Codex) is also the better scheduler of boring, repetitive automations: my inbox triage and LinkedIn engagement queues run there because they execute consistently. And it is a phenomenal coder in its own right. I pay its bill too, at the $100 step rather than $200:I dropped that tier deliberately, to free up budget for a third tool, and I have not wanted it back. The full reasoning is in ChatGPT Pro $100 vs $200, including what the downgrade actually costs you now that GPT-6 Pro is metered 50 against 200.
The move nobody's comparison covers: they work best as a team. My setup runs a custom MCP connection (plain English: a bridge I built so the two AIs can talk to each other), and quote me on this: when Claude builds something, it sends the work across for QA automatically, like a sparring partner. It beats using a sub-agent of the same model, because the second opinion comes from a genuinely different brain. u/cc_apt107, who pays for both ChatGPT Pro and Claude Max, landed on the same division back in 2025 and it still holds: "Claude is MUCH more 'coachable' and context aware... I find myself using ChatGPT 5 more than Codex. It is really, really good at enhancing Claude's plans... But I rarely let GPT touch code."
Claude executes. GPT reviews and researches. Every serious both-payer I have found converges on some version of that split.
The cheaper play most people should try first
Before either $100 tier: $20 on both. Separate allowances, two toolsets, $40 total. It is one of r/claude's most-upvoted plan questions for a reason, and I wrote the full breakdown of the $20 tiers in Claude Pro vs ChatGPT Plus.
A month later, with people actively shopping these $100 tiers, I still stand behind it. Unless you have a serious workflow and the budget, just try the $20. When you hit the limit, you will feel it, and then upgrading takes about ten seconds.
Let the wall make the decision, not the comparison table.
The same conclusion shows up in the community. In an r/ClaudeAI thread on how people handle the gap between the $20 plan and Max, the most popular serious answer was a multi-model strategy: keep the $20 Claude Pro plan for the heavy lifting and supplement it with a second cheap plan rather than jumping a tier.
The honest counterpoint from that same $20-on-both thread, because it is real: one user found that "switching between models mid-workflow just to save $60 sounds good on paper but the context switching kills productivity more than the token limit does." If your work lives in ONE deep tool all day, one big tier beats two small ones.
Two practical receipts before you buy:Do not subscribe through the Apple App Store. Community threads consistently warn the in-app price runs ~$125 for the $100 tier because of Apple's cut. Subscribe on the web.
The one-month sprint is legit. Upgrading for a single month to build something specific, then dropping back down, is a strategy real users run. These are monthly toggles, not marriages.Your habits matter more than the price tag
Since there is no annual discount on either side, the only lever left is how efficiently you spend what you buy. This turned out to be worth more than the tier difference for me.
Two things did most of the work.
I run a context meter in my status line so I can eyeball how loaded a session is. And I have a hook that fires around 500k of context: it writes a next-steps markdown file at the repo root, I close the window, open a fresh one, say "read the next-steps file," and resume with zero context but the same knowledge.
Before those habits, I was burning 20 to 30% of my top-model allowance in a single day. I am not sitting at 500k of context paying for it on every message anymore.
The subscription price matters less than your habits. I measured this properly across nine weeks of my own logs, and the numbers are in my real Claude Code usage.
Verdict: pick by what your week looks likeYou chat, plan, draft, think: stay on $20. Truly. The caps you would pay to remove are not the ones limiting you.
You are starting to run agents (Claude Code, Cowork tasks, scheduled automations): $100 Claude Max, and run the efficient model as your default driver.
You build daily and hate waiting for resets: $200 Max. That was my trigger: serious building means testing fast, and waiting for a weekly reset mid-build is the most expensive thing on this page.
Your volume is research, review, and everyday everything: ChatGPT Pro at $100 carries shocking volume.
AI runs your whole operation: both, split the labor: Claude builds and writes, GPT reviews, researches, and runs the boring reliable stuff. That is my actual setup, and it is the one configuration none of the benchmark articles can review, because you have to live in it.Two adjacent paths worth naming before you commit: Google's Gemini tiers compete hard on price for chat-first users, and if your usage is truly spiky, both companies' pay-as-you-go APIs can beat a subscription. For teams, Claude Team and ChatGPT Business are the per-seat versions of this same decision.
Choosing between the two Claude Max steps rather than across vendors? That split is Claude Max 5x vs 20x. What do agents actually do all day for a non-developer? That list exists: 20 real Claude Cowork use cases. Torn between the two agent doors on your new plan? That fork is Claude Cowork vs Claude Code. And if you want somewhere to start before spending anything: the 15 workday AI prompts at /start run fine on the $20 tiers.
One small thing that saves a question: ads do not touch either of these tiers. OpenAI's ads test covers Free and Go only. The full picture is in does ChatGPT have ads.
Published July 12, 2026. Last reviewed and updated September 4, 2026: retired the "identical ladders" claim, which stopped being true on September 3 when GPT-6 Astra shipped and ChatGPT Pro split its own two tiers 4x on the newest model. Added the GPT-6 Pro allowance section, a section on what Anthropic's 5x/20x multiplier actually measures, and corrected my own spend to ChatGPT Pro at $100 after downgrading from $200. Pricing re-verified against claude.com/pricing and OpenAI's help center that day. Previously updated August 14, 2026 with the no-annual-billing parity, the Fable tiering row, the no-bypass rule, Opus 5 and the GPT-5.6 family, the effort-dial section, and the habits section. Community reports linked throughout.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: people use Claude Cowork for the boring work that eats their week: sorting files, triaging email, building the same report every Monday, prepping taxes, tracking home repairs, tailoring resumes.
Not science fiction. Chores.
This is a complete, organized list of 20 real use cases, pulled from what actual users report doing (linked, so you can verify every one) plus the ones I run myself. The community's own words, from the biggest use-case thread: "No '50x your productivity' hype please, just real, everyday use cases." Agreed. If you searched "claude cowork use cases," "cowork examples," or "what do people use Claude Cowork for," this is the reference. Steal freely.I walk through this on camera in The Only 4 Claude Cowork Use Cases Worth Keeping (I Deleted the Rest) (12 min).The map: 7 categories of real Cowork workCategory
The chores
Best first pick?Files and documents
Downloads cleanup, renaming by content, pulling data out of PDFs
✅ easiest winInbox and comms
Email triage, drafts in your voice, calendar wrangling
Start day 2Recurring reports
Monday expense report, weekly metrics, morning brief
The compounding oneResearch and monitoring
Competitor watch, topic digests, multi-document synthesisMoney and admin
Receipts to spreadsheet, tax-prep summaries
Seasonal heroCareer
Resume tailoring, job radar, LinkedIn upkeep
Quiet compounderPersonal ops
Home maintenance log, personal CRM, insurance questions
Sleeper hitsEverything below runs in plain English. No terminal, no code. Cowork is included on every paid Claude plan and, as of July 2026, its scheduled tasks run remotely, even with your laptop asleep.
Files and documents: the easiest first win
1. Clean the Downloads folder. The most-reported first use case, and for good reason. One user on r/ClaudeAI had Cowork organize their Downloads folder in 5 minutes and called it a day of work saved. A blogger stress-tested it on 2,200 files: sorted into 11 folders in about 20 minutes, including renaming 50 "Untitled" images and 17 generically named documents by reading their contents.
Honest number from that test: about 70% of the auto-names were right. Great assistant, still worth a skim.
2. Rename files by what is inside them. Not by filename. It opens each file, reads it, names it properly. Users in the big use-case thread describe exactly this workflow for the "sort the chaos" problem.
3. Pull specific data out of a pile of PDFs into one table. Contracts, invoices, statements in a folder → one clean Markdown or Excel table with the fields you asked for. Same thread, multiple reports.
4. Turn messy project folders into master documents. One engineer, in an r/ClaudeAI thread, had Cowork scan a folder from a production-line commissioning, categorize everything, and produce a master bill of materials for the maintenance team. Same move works on any documentation dump.
5. The insurance folder. One user dropped four insurance policies, including a 240-page health policy, into a folder and now just asks questions against it, per the thread. Every household has this folder waiting to exist.
One caution from experienced users: go folder by folder, not "reorganize my whole computer."
Inbox and comms: the daily grind, delegated
6. Email and calendar triage. One r/ClaudeAI user put it in the most relatable way possible: the fancy agentic stuff was over his head, so he just uses Cowork for handling emails and organizing calendars. That is the right starting altitude. This is also my own #1: my inbox gets checked multiple times a day by the system, and it is exactly the kind of chore I hand Cowork.
7. Status updates drafted in your voice. One user has Cowork read the project folder and draft the update so all that is left is pressing send. The trick: it writes FROM your files, so the update is accurate, not generic.
8. Gmail into a tracking spreadsheet. Sales and customer-success people report having Cowork categorize emails (opportunity, lead, rejection) and update their sheet or CRM. If your pipeline lives in your inbox, this is the unlock.
9. Synthesize the noise. Teams chats, meeting notes, emails, scattered docs, dumped and synthesized on a schedule into one readable brief. One user built a Getting Things Done dashboard aggregating overdue Asana tasks and unread email.
Recurring reports: where Cowork stops being a tool and becomes staff
This is the category most articles miss, and it is the best one. Type /schedule inside any task, set the schedule once, done. It runs even when your computer is asleep.
10. The Monday expense report. One user forwards anything resembling a receipt; Cowork stores them all week and sends an expense report every Monday morning. The same user's summary of the whole product: "It really is like a junior assistant."
11. The 7am digest. Same user gets a daily email summarizing highlights from the subreddits they care about. Swap in your industry sources.
12. Weekly analytics review. This one is mine, and my advice is: do not make it complicated. Mine answers three questions: how did my content perform, are we trending the right direction, and why or why not.
Here is a real one from this week, verbatim from my system:The one thing I have learned: have it do the research BEFORE proposing changes, because sometimes the honest answer is "nothing is broken, keep going, put in the volume." Not every flat week needs a strategy pivot.
13. The auto-updating deliverable. One user maintains a revenue model that updates itself with the latest forecasts, slides included. Cowork writes real Excel, PowerPoint, and Word files, so "the deck" can just always be current.
Research and monitoring: your standing scouts
14. Watch for specific signals. One user has Cowork scan public sources for phrases that signal frustration with a product category, logging date, source, and exact wording. That is a lead machine or a product-research machine, depending on your job.
15. Deep-mine your own archive. Podcast host Aakash Gupta ran Cowork across years of his own transcripts to find where guests directly contradicted each other, then had it build the "10 most quotable insights" into a Keynote deck, edited directly in the app. If you have an archive (calls, docs, posts), it is sitting on unmined gold.
16. Conference prep. Attendee list + a data source + Cowork = bios and talking points for everyone you want to meet, read on the plane.
Money and admin: the seasonal hero
17. Tax prep from raw transactions. A user in a Claude community group downloaded a year of bank and credit-card transactions, handed them to Cowork, and got back an annual summary for tax prep, in minutes instead of days. A LinkedIn user reported the same move saving roughly 5 hours. Pair it with use case #10 and next year's version builds itself weekly.
Career: the quiet compounder (even if you are not job hunting)
18. The recurring LinkedIn health check. My favorite move for employed people. Have Cowork pull your LinkedIn profile on a schedule and grade it: is your headline current, is your bio dry, when did you last post? Simple report, monthly or biweekly.
Why bother? Recruiters find active profiles. Inbound happens while you sleep.
And this is precisely the task you would forget: Cowork won't.
19. The job-search folder. Users run a folder system: resume, cover letters, and Cowork tailors per job description. Fair warning from that thread: review before sending, it can oversell you. A lighter version even if you are happy where you are: a monthly "what roles like mine are trending" radar.
Personal ops: the sleeper hits
20. The home maintenance tracker. The single best "I did not know it could do that" story in the threads: a homeowner had Cowork sift 10 years of email for every repair (plumber, electrician, stucco), log dates, contractors, and costs, then build a maintenance schedule synced to Google Calendar. Household operations, fully delegated. Honorable mentions from the same community: a personal CRM for keeping up with your network, and the insurance folder from #5.The 60-second version, from my Shorts: Claude Cowork can plan your entire trip: the 3-step setup.The move that beats all 20: reverse-prompt it
Here is what I tell everyone who asks "but what would I use it for?"
Do not guess. Tell Cowork your situation: who you are, what you are good at, what you hate doing, your goals, your roadblocks. Then ask IT for the best opportunities to take work off your plate. Save that context as an "about me" file so it remembers, and build the top two or three suggestions as scheduled tasks.
My content chores will not match yours. The method transfers: boring tasks, delegated once, compound forever.
If you think visually: riff out loud and have Cowork sketch and document your workflows. The old version of this was an afternoon in Canva. Now the documentation is the byproduct of a conversation.
The honest limitsIt burns your allowance faster than chat. Agentic work is multi-step work. On the $20 Pro plan, heavy daily Cowork use will meet the weekly cap. That is plan sizing, not failure; if you are hitting it weekly, that is the exact signal covered in Claude Max vs ChatGPT Pro.
It is a strong assistant, not an infallible one. The 2,200-file test above hit ~70% naming accuracy. Skim its work, especially anything outbound.
Chat memory does not carry over into Cowork (outside projects), so give tasks their context.Facts current as of July 2026, checked against Anthropic's Cowork page: included on all paid plans, on desktop (macOS, Windows, Linux, ChromeOS) and the web with mobile in beta, scheduled tasks run remotely, and connectors cover email, calendar, Slack, Drive, and more.
Still deciding between the two cockpits? That is its own decision: Claude Cowork vs Claude Code. And if you want ready-made starting prompts for your workday, grab the 15 workday AI prompts at /start: most of them make excellent first Cowork tasks.
Published and last reviewed July 12, 2026. Every use case links to its source thread or article; community-reported details verified against the linked threads at publish time.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: If you do not write code, use Cowork. But the reason everybody gives for that answer is now wrong, so it is worth thirty seconds of your time to get the real one.
The rule I give people: Chat answers. Cowork does your office work. Claude Code builds software.
Searching for "cowork vs code", "Claude Code vs Cowork", or "when to use Claude Cowork vs Claude Code"? You are in the right place. The fork is not the interface. It is what you are handing over.
The correction: the terminal fork is dead
The earlier version of this page told you the answer came down to one question: do you want to click, or do you want to type commands.
That was the whole thesis and it is no longer true.
Claude Code ships in the desktop app. From Anthropic's own Cowork FAQ: Claude Code "is available in the terminal, in the desktop..." It is not a terminal product with a GUI cousin. It is a product with two front doors.
Here is what that looks like on my own machine.My Claude desktop app, September 5, 2026. Claude Code and Cowork are sibling tabs in the same settings window. There is no terminal in this picture.
So if you have been avoiding Claude Code because you do not want to live in a command line, that reason has expired. You may still not want it. Just not for that reason.
The difference in one tableClaude Cowork
Claude CodeWhat you hand it
A job to finish
A codebase to buildWhere it lives
Claude desktop app, with a phone beta
Terminal and desktop appWhere tasks run
Anthropic's servers; the desktop app bridges your local files
SameBuilt for
Files, docs, email, reports, browser work
Software: code, projects you are building, custom systemsSetup required
None. It is in the app on every paid plan
None on desktop; a terminal install if you want the CLIBest at
Delegated and recurring work, scheduled tasks
Complex builds, custom scripts, full step-by-step visibilityIncluded on
Pro, Max, Team, Enterprise
All paid plansSame AI engine?
Yes
YesYou are not choosing what to buy. Both come with your paid plan. You are choosing which door to open.
Same engine, different job
People treat Chat, Cowork, and Claude Code like a power ladder: Chat is the starter tier, Code is the pro tier, Cowork sits in the middle. That model is wrong and it is the single biggest source of confusion I see.
Anthropic could not be more direct about it:"Cowork and Code run on the same engine. Both are Claude Code underneath."And the replacement fork, in Anthropic's own words: "Cowork is for delegating to Claude. Code is for building software with Claude. Most knowledge workers will live in Chat and Cowork."
This gets debated constantly. u/SilverConsistent9222 on r/Anthropic put the trap well: "They're often discussed as if one is an upgrade over the other. That's not really accurate. They operate in different environments."
The loudest version of the confusion is a thread titled "What's the point of Cowork when you have Claude Code?", which pulled 316 upvotes and 149 comments on r/ClaudeAI. That question gets asked because the interface answer never satisfied anyone.
The cleanest rule of thumb I have seen came out of u/Talley-Ho's thread on the same question: codebase goes to Code, everything else file-based goes to Cowork.What Cowork gives you out of the box
This is the list that should have been on this page all along, and almost no ranking page reproduces it. Per Anthropic's current Cowork documentation:Access local folders with no upload step
Navigate your logged-in browser through Claude in Chrome
Reach desktop apps via computer use
Delegate parallel work, several jobs at once
Run tasks on a scheduleOne honest caveat, because Claude Code Desktop has been catching up fast: several of these are no longer strictly exclusive to Cowork. The distinction that holds is that in Cowork they are the product, ready without configuration, while in Claude Code they are things you can set up. Check Anthropic's current docs for both before you treat any single row as a hard gate.
And one that post-dates this page's last rewrite: since August 26, 2026, Cowork has its own browser built into the desktop app.
That computer-use row is the one to sit with, because it is the capability that has nothing to do with avoiding a terminal.
To be precise about my own setup: I have a shorts pipeline that logs into my Mac mini, takes screenshots, clicks around and opens real apps to run experiments. That one runs in Claude Code, not Cowork. I mention it because it is the kind of work Cowork's computer use puts within reach of someone who never opens a terminal, not because Cowork ran it.
Scheduled tasks are the least-hyped item on that list and probably the most useful. You set the schedule once and it just runs.Real people use this for unglamorous, valuable things. A product manager, u/RusticGroundSloth on r/ClaudeCowork, described his setup:"there's a 'director' cowork project that pulls all of my Teams and Email from the last 24 hours every morning at 6:00 a.m... Claude Cowork has essentially become my project/program manager and is saving me HUGE amounts of time... I realize a bunch of this could probably be done in Claude Code but it was extremely easy to set up in Cowork and I haven't had to think about it since."That last part is the whole point. Could have been done in Code. Was not worth doing in Code.The 60-second version, from my Shorts: You don't need Claude Code. Cowork does 90% of it.Everything Cowork hands a non-coder lives behind this one plus button: files, a screen recording that becomes a skill, your skills, your connectors, plugins. None of it is a terminal.
Where Claude Code still wins
Code is for building things. That is the line, and it holds.
If you use Cowork every day, you are going to do most things. You will get your job done, you will build skills, you will run genuinely powerful workflows. That covers the large majority of knowledge work.
But if you want to code, hence the name, that is where Claude Code starts. Build a game. Build an app.
And "an app" is broader than it sounds. I have a skill that edits my videos. That is not an app in the App Store sense, but it is a real piece of software with many moving parts, and that is Claude Code work.
The other thing Code gives you is visibility into every step the agent takes. That is why developers will not give it up, and it is a legitimate reason for a non-developer to graduate too, eventually.
Which is why you should ignore a whole category of online takes. Developers keep posting some version of what u/gatsbtc1 wrote: "Everything cowork can do, so can code. And code is so much more robust."
True. And irrelevant. Everything a car does, a manual transmission race car also does. You still should not learn to heel-toe shift to get groceries.
How I actually split them, and why I am the wrong model
Here is the honest version, including the part that does not flatter me.
I do not switch between them mid-job. I want to be straight about that, because "tell me about a job that moved from one to the other" is a great question and I do not have a story for it. I use both. I do not toggle.
For me, it is Claude Code, because it does everything I need and I am already fluent in it. Going from Claude Code back to Cowork would feel like a downgrade from what I am used to.
But that is a fact about my history, not about the products.
The reason I got good at Claude Code is that I decided to, so I could make YouTube videos about it. And getting there was genuinely rough. I would see the hype, watch technical people using it, and think: I am pretty technical, why can I not get into this?
I did not even know what Claude Code was. I knew it was in the terminal. Did I have to install something? I did not know what a CLI meant.
This was over a year ago, maybe six months after it first came out, when I was mostly living in n8n.
When it finally clicked, it clicked hard. It just made sense to me. But I am comfortable staring at a terminal window, and I do not recommend that path for everybody.
Here is the part I would want you to take from this: if I were starting today, and Cowork existed, I would probably have started there. The transition I made was painful and it was only worth it because I had a specific reason.
You probably do not need to repeat it.
The intention behind having both seems clear enough: if Cowork is your main driver, Claude Code is there for when something gets serious. It is the same shape as ChatGPT Work and Codex on the other side of the market. Nobody thinks that pairing is strange.
The usage gotcha nobody warns beginners about
Anthropic prints this in plain sight on its Cowork documentation and almost no comparison article mentions it: Cowork consumes your limits faster than Chat does.
Agentic work is multi-step work. One Cowork task that sorts a folder of 30 files is not one message; it is dozens of steps, each drawing on your allowance.
Claude Pro caps you two ways: a rolling five-hour session limit plus a separate weekly cap. On the $20 plan, heavy Cowork use will hit that, and it will feel like the product is broken.
It is not. The $20 tier is sized for chat-first use with some agent work on the side.
On the narrower question people keep asking, whether Cowork and Claude Code burn your plan at different rates, there is a 27-upvote thread asking exactly that. I am not going to give you a ratio, because Anthropic publishes no multiplier and every number circulating is somebody's estimate from their own usage. What is documented is that both draw from the same pool as your chat.
If either one becomes your daily workhorse, that is what Claude Pro vs Max is about.
If you are choosing between Chat and Cowork instead
A lot of people who land here are actually asking the earlier question: when does a job leave the chat window at all?I walk through that one on camera in the Chat vs Cowork video (11 min).Anthropic's own distinction is a good one: Chat is a conversation you steer turn by turn; Cowork is a delegation. Chat gives you text you will read. Cowork gives you a finished file or an action taken.
Which one should you open?You do not write code: Cowork. It is already in your Claude desktop app on every paid plan.
Your work is documents, email, spreadsheets and reports: Cowork. That is what it was built for.
You want a chore to just happen every morning: Cowork scheduled tasks.
You want Claude to click around real apps or your browser without setting anything up: Cowork. Claude Code can be driven to do this too, but Cowork is where it is a built-in feature rather than something you wire together.
You are building software, a game, an app, or a complex reusable skill: Claude Code.
You want every step the agent takes to be visible: Claude Code, terminal or desktop, your choice now.
You avoided Claude Code because of the terminal: that reason is gone. Try the desktop tab before you decide.How I checked this
The shared-engine architecture, the desktop availability of Claude Code, the Cowork-only capability list and the Chat-versus-Cowork distinction come from Anthropic's current Cowork and model documentation, read September 5, 2026. The built-in browser date comes from Anthropic's own announcement. Plan inclusion comes from claude.com/pricing, read the same day. The settings screenshot is my own machine on that date.
I run both daily, so the preference above is mine. The capability lists themselves are Anthropic's documentation, not my testing. What is not lived: I have not measured usage consumption between the two, and I have deliberately left that number out rather than repeat an estimate.
If you want somewhere concrete to start today, grab the 15 workday AI prompts at /start: they are built for exactly the kind of tasks Cowork eats for breakfast.
For adjacent decisions: Claude Cowork vs ChatGPT Work for the cross-vendor version, Claude Cowork use cases for what to actually run first, and Claude Cowork scheduled tasks for the recurring-work setup.
Published July 9, 2026. Last reviewed and updated September 5, 2026: replaced this page's central thesis, which claimed the fork was clicking versus typing commands, after Anthropic made Claude Code available in the desktop app. Added the Cowork-only capability list, the built-in browser that shipped August 26, 2026, an honest account of why I personally live in Claude Code and would not recommend that path, and the usage question with no invented ratio. Moved the Chat versus Cowork video out of the introduction to the section it actually answers. Product facts checked that day against Anthropic's own documentation and pricing page, linked inline.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Quick answer: I pay for both, every month, and use both every day.
If you can only justify one $20 subscription for work, take ChatGPT Plus: it is more general, you will use it for more things, and its limits bend instead of break. Take Claude Pro instead when your work is depth: long documents, serious writing, analysis where the AI must do exactly what you said.
The honest one-liner: Claude wins quality per message. ChatGPT wins messages. Real workloads want both.
If you searched "claude pro vs chatgpt plus" or "chatgpt plus vs claude pro," here is the comparison I wish existed when I started paying for these: current numbers from the official pages (September 2026), what real users say after months on each, and a verdict by the kind of person you are, not a fake universal winner.
The two plans in one table (September 2026, official pages)Claude Pro
ChatGPT PlusPrice
$17/mo billed annually ($200 up front), or $20 monthly
$20/mo, no annual optionUsage limit style
5-hour session limit + a separate weekly cap
Reasoning allowances per model, not published as a single numberWhen you hit the limit
Hard stop until reset, or turn on usage credits
Continues with another available reasoning modelWhat shares the allowance
One bucket: chat + Claude Code + Cowork
Chat; Work and Codex have separate usage and credit rulesModels
Opus, Sonnet, Haiku. Fable 5 and Fable 5.1 are NOT included - Pro pays usage credits for them
GPT-5.6 Sol at Instant, Medium and High. No Extra High, no ProNewest model included?
No. Fable runs on usage credits
No. GPT-6 Astra is in Work and Codex, not the chat pickerIncluded agent tools
Claude Code, Cowork, Design, Research, unlimited projects
Agent mode, custom GPTs, scheduled automations, ChatGPT WorkUpgrade path
Max: $100 (5x) or $200 (20x), and Max is where Fable 5 and Fable 5.1 become included
Pro tiers: $100 and $200, and Pro is where GPT-6 Pro appears in chatOn the missing numbers. That Plus row used to carry a message count. OpenAI no longer publishes one: its current help page describes reasoning allowances per model without a figure, and the only hard number it publishes anywhere in the ChatGPT ladder is for the Pro tiers (50 GPT-6 Pro messages a week at $100, 200 at $200). Anthropic does not publish a numeric quota either. Anyone quoting you an exact message cap for a $20 plan in 2026 is quoting a number the vendor stopped standing behind.
Correction (July 31, 2026). This table used to list Fable as a Claude Pro model. That is no longer true, and it is the single most important change on the Claude side this year. Per Anthropic's Fable 5 plan page, the promotion that included Fable 5 in Pro's weekly usage limits ended July 19, 2026 at 11:59:59 PM PT. On Pro today, “Fable 5 and Fable 5.1 aren’t included in your plan's usage limits” - you pay usage credits on top of the $20. On Max both are included as standard, up to 50% of your weekly limits.
That reshapes the $20-vs-$20 question this whole page is about, so it gets its own section below.
Two structural details still matter more than any benchmark. First, the limit style: Claude stops you; ChatGPT demotes you. A hard stop mid-workday feels very different from quietly getting a smaller model. Second, the bucket: on Claude, every product drains one allowance, so a heavy agent session eats your chat capacity too.
The two ceilings, in plain language
People searching for "claude pro vs chatgpt plus limits" are usually trying to work out why they got stopped, so here is the mechanism.
Claude Pro runs two separate ceilings at once. There is a rolling five-hour session cap, and on top of that a weekly cap that applies across all models. You can be nowhere near your weekly limit and still get stopped by the session one.
That is why being cut off feels random. It usually is not - you hit whichever ceiling came first.
One more thing worth knowing before you buy: Cowork burns your allowance faster than chatting because multi-step tasks are compute-intensive. The feature most likely to sell you on Claude is also the one most likely to run you into a wall.
The complaint you will actually feel: Claude's weekly cap
The single loudest pain in the Claude community right now is the $20 tier's weekly limit. In a heavily upvoted r/ClaudeAI thread with close to 300 comments, "Claude Pro feels amazing, but the limits are a joke," u/iameastblood says it plainly: "Even though I use it quite sparingly, my weekly limit is already mostly drained by mid-week (currently sitting at 74% used)... it feels like I'm paying for a premium service I can barely use."
The consensus across those roughly 300 comments is uncomfortable but consistent: for heavy users, the $20 Pro plan works like a trial tier, and serious daily use points to Max at $100 and up.
Here is the de-blame part, because the complaints usually read like user error and they are not: the $20 tiers are designed differently on purpose. Claude Pro sells you a smaller amount of a very deep tool. ChatGPT Plus sells you a large amount of a very general tool. Neither is a scam. They are different bets, and you should pick the bet that matches your work.
Where each one actually wins (from people doing real work)
The pattern across every long-term comparison I have read, and my own daily use, is consistent:
Claude wins on depth and obedience. The best summary I have seen comes from a four-month "paid for both" comparison on r/ClaudeAI: for long-form writing, analysis, and structured documents, "claude wins... its not close." The same post nails instruction-following: "tell it 'respond in 1 sentence' and it actually does. gpt-5 negotiates." And a reply from u/Ashtonator28 became my favorite one-line review of both companies: "anthropic accidentally built a brilliant coworker. openai accidentally built a very competent butler."
ChatGPT wins on versatility and volume. A real-estate and tax professional, u/Free-Writer-1123 on r/claude, pushed back on the Claude hype after using both on actual client work: "ChatGPT has been the most-versatile with higher correct rate of answer. Better at reviewing text from PDF. Better at calculating for loans... I keep reading how awesome Claude is, but I have much different user experience than most." In the same thread, a teacher (u/FATJIZZUSONABIKE) countered that Claude's lesson plans were "so detailed and inspired that I've barely had to adapt anything."
Both are right. That is the whole point: depth versus breadth.
My own week looks like this. ChatGPT earns its subscription on the general layer: quick scripts, research, everyday questions, and its scheduled automations, which are genuinely excellent and feel like Cowork on steroids. I also keep its Codex coding agent wired in as a sparring partner that reviews and stress-tests work my main tools produce. Claude earns its subscription on the deep layer: the writing, the systems, the long careful work where instruction-following is everything.
Full disclosure, since this post is about the $20 tiers: these days I pay well above them, Claude Max at the 20x step and ChatGPT Pro at $100, because AI runs my entire operation. The split logic in this post is exactly how I route work between them; the bigger tiers just give the same split bigger buckets. Straight from my phone:The move nobody prices out: $40 across both beats $100 on one
Here is the configuration question that keeps showing up in the forums: one $100 plan, or two $20 plans?
For most working professionals, two $20 plans win, and not just on price. The reason is the limits math. The two allowances are completely separate, so splitting your workload means you almost never hit a wall on either side. I stopped hitting limits entirely once I split my work this way: the general, high-volume load goes to ChatGPT, the deep work goes to Claude, and each subscription only carries half my day.
You also get insurance. When one tool has a bad day, and they all have bad days, the other is right there.
The $100+ tiers make sense at the point where ONE of the tools has become your primary workhorse and its half of the split is still hitting caps. That is a great problem to have, and you will know when you have it. If that fork is where you are, the within-Claude version of this decision is Claude Pro vs Claude Max.
Neither $20 plan includes its newest model now
This section used to be about Claude only. As of September 2026 it is about both, and the symmetry is the most useful thing on this page.
Claude Pro gates Fable behind credits. The promotion that included Fable 5 in Pro's weekly limits ended July 19, 2026. Re-checked on claude.com/pricing September 4, 2026, the Fable row still reads "Usage credits" for Pro:ChatGPT Plus gates Astra out of chat. OpenAI shipped GPT-6 Astra on September 3, 2026. Its help center says two things: in chat the model is branded GPT-6 Pro and is listed for Pro $100, Pro $200, Business and Enterprise, and separately that Plus does not get the picker's Pro reasoning option at all:Two different mechanics, one outcome. At $20 a month, neither company hands you its newest model the way you would assume.
Where Plus does get Astra
This is the part the launch coverage got wrong, so it is worth being precise.
ChatGPT Plus is not excluded from GPT-6 Astra. OpenAI's help center says "Plus plans include GPT-6 Astra in ChatGPT Work and Codex as it rolls out." It is only the chat model picker that is gated.
OpenAI's own announcement reads differently, which is why everyone got confused:GPT-6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users.Both statements are true, one is about the model and one is about the surface. The full breakdown is in which ChatGPT plans get GPT-6 Astra.
Plus also keeps GPT-5.6 Sol at Medium and High reasoning in chat, which is a strong model. It just does not get Extra High or Pro.
So what does this change?
Less than the headlines suggest, and that is the honest answer.
If you were choosing a $20 plan based on which one gives you the shiniest model, stop. At this price neither does, and the difference between "usage credits" and "wrong surface" is not worth agonizing over.
Choose on the tools and the failure mode instead. That is what actually shapes your week.
There is a second thing I did not appreciate until I was paying for both. ChatGPT is straight up more generous about resetting limits. Do not be surprised if once or twice a month those limits just get reset on you. That is my experience across both accounts, not a published policy, so do not budget around it.
None of that makes Claude the wrong pick. Everything in the "where each one wins" section above still holds - depth, instruction-following, and design are still why I keep it. The honest 2026 read: at $20 versus $20, ChatGPT gives you more room and more forgiveness; Claude gives you a better deep-work tool with a harder ceiling. Neither gives you its newest model.The 60-second version, from my Shorts: ChatGPT and Claude both cost $20. Here's how to pick ONE..Verdict: pick by the person you areFirst AI subscription, general work (email, documents, planning, everyday questions): ChatGPT Plus. It is the better all-rounder and the friendlier on-ramp, and you will use it for personal life too: meal plans, budgets, all of it.
Your work is writing, analysis, or anything where the output must follow instructions exactly: Claude Pro, and accept that the weekly cap is the price of depth at $20.
AI does real work for you every single day: both, $40 total. Split the load, double the limits, route each task to its specialist.
You are already slamming into caps on a split setup: that is the upgrade signal for a $100+ tier on whichever side does your heavy lifting.The number to remember is not a message count, it is a failure mode: Claude stops you, ChatGPT demotes you. One shared weekly bucket that runs dry, against per-model allowances that fall back to a smaller model and keep going. Match that to your actual day and the choice makes itself.
If your $20 decision turns into a $100 decision, that fork is ChatGPT Pro $100 vs $200 on the OpenAI side and Claude Max vs ChatGPT Pro across both.
Whichever plan you pick, put it to work the same day: the 15 workday AI prompts at /start run on either tool and cover the exact tasks these subscriptions are for.
If the next question is what to do with these subscriptions once you have them, start with Claude Cowork vs Claude Code for the Claude side, and my breakdown of when you need automation at all for the bigger picture.
Published July 3, 2026. Last reviewed and updated September 4, 2026: GPT-6 Astra shipped September 3 and ChatGPT Plus does not get it in the chat picker, so the "newest model" section is now symmetric across both plans. Also replaced the stale GPT-5.5 message-count rows with what OpenAI currently publishes, corrected the Plus reasoning levels (Medium and High, not Extra High or Pro), and corrected my own ChatGPT spend to the $100 Pro tier. Verified against claude.com/pricing and OpenAI's help center article 20001354 that day. Previously reviewed July 31, 2026, when Fable 5 left Pro's included usage limits. These numbers change often; both companies' pages are the source of truth.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Claude Projects and Custom GPTs solve the same problem: you keep re-pasting the same background into a chat every morning. Claude Projects give you a workspace with saved instructions and reference files that every chat inside it already knows. Custom GPTs turn that same setup into a shareable assistant that other people can use too.
The short decision rule: if the assistant is just for you, use whichever tool you already pay for. If you need to hand it to other people, Custom GPTs are easier to share. And there is a decent chance you do not need either one yet. More on that below.
Quick comparisonClaude Projects
Custom GPTsWhat it is
A workspace inside Claude with saved instructions and files
A configured assistant inside ChatGPTBuilt for
Your own recurring work
Assistants you hand to other peopleSetup
Custom instructions plus project knowledge files
Instructions, knowledge files, optional actionsSharing
Inside a Claude Team or Enterprise workspace only
Direct link, workspace, or the public GPT StoreExtra powers
Skills, reusable instruction packs Claude loads when relevant
Actions, which let the GPT call other apps' APIsCost to build
Included on Claude's free plan, with limits
Requires a paid ChatGPT planCost to use
Free with limits, more on paid plans
Free users can use existing GPTsDoes Claude have an equivalent to Custom GPTs?
Yes, mostly. Claude Projects are the equivalent for personal use.
A Project holds two things: custom instructions (how Claude should behave in this context) and project knowledge (the files you would otherwise re-upload every time). Every new chat inside the project starts with all of that already loaded. Projects also support Skills now, which are reusable instruction packs Claude pulls in when a task calls for them, like a house style for documents or a specific report format.
What Claude does not have is a store. You cannot publish a Project to a public gallery or send a coworker a link to "your assistant." Sharing only works inside a Claude Team or Enterprise workspace.
So the honest version of the answer: for "an assistant configured for my own work," Claude matches Custom GPTs. For "an assistant I can hand to anyone," it does not.
What is a Claude Project, in plain English?
Think of it as a folder that remembers.
Say you are an HR manager. You create a project called "Policy Questions," upload the employee handbook and your benefits summary, and write instructions like "answer questions using only these documents, quote the relevant section, and flag anything the documents do not cover." From then on, every chat in that project answers from your actual handbook instead of generic HR advice.
A teacher might keep one project per course: syllabus, rubric, and a note about reading level. A project manager might keep one per client: status report template, stakeholder names, the tone the client expects.
The win is not that Claude gets smarter. It is that you stop spending the first five minutes of every session rebuilding context. I run my own recurring work this way, and the setup pays for itself in the first week. If you want to see what a working example looks like end to end, here is my Claude SEO workflow. Different job, same pattern: instructions once, files once, then every chat starts warm.The 60-second version, from my Shorts: The free Claude folder that remembers you (3 setups beginners never turn on).What is a Custom GPT, in plain English?
A Custom GPT is a pre-configured version of ChatGPT with its own name, instructions, and knowledge files. You build it once through a conversational setup screen, no code involved, and it behaves the same way every time anyone opens it.
Two things make Custom GPTs genuinely different from a Claude Project:Sharing. You can send a GPT to a coworker as a link, share it across your company workspace, or publish it in the GPT Store. If you build a "Job Description Drafter" for your HR team, the whole team gets the exact same assistant without configuring anything.
Actions. A GPT can be wired to call outside services, so it can look something up or send data somewhere instead of just chatting. In practice, setting up actions requires API details most people will never touch. That is fine. The sharing alone is the reason most teams pick Custom GPTs.One cost note: anyone can use existing GPTs for free, but building your own requires a paid ChatGPT plan.
Which one fits how you work? Three questions
You can make this decision in five minutes. Ask these in order.
1. Which subscription do you already pay for?
This settles it for most people. Do not switch from ChatGPT to Claude, or the reverse, to get this one feature. Both products have a good version of "saved context plus instructions." The tool you already use, with the history and habits you already have, wins by default.
2. Do you need to share the assistant, or just use it yourself?
Just you: Claude Projects or ChatGPT Projects, whichever side you are on. Done.
Your team needs the same assistant: Custom GPTs win clearly. A link your coworkers can open beats a setup doc you have to walk five people through. Claude can share projects too, but only if everyone is in the same paid Team workspace, which is a bigger ask than "click this link."
3. Where does your work context live?
If your context is documents you can upload, like handbooks, templates, rubrics, and past examples, both tools handle it well. If your context lives inside other systems you would need to connect, neither one solves that cleanly out of the box, and you should solve the document version first anyway.
What about ChatGPT Projects vs Claude Projects?
If you are on ChatGPT and the assistant is just for you, skip Custom GPTs entirely. ChatGPT has its own Projects feature: chats grouped in a folder, with files and instructions attached, and memory scoped to that project.
For solo use, ChatGPT Projects and Claude Projects are close to interchangeable. Both hold files. Both hold instructions. Both keep your chats organized by context instead of one endless sidebar. People argue about which model writes better, and that is a real preference, but the projects features themselves are not the deciding factor.
The clean way to remember the whole lineup:Projects (Claude or ChatGPT) = context for you
Custom GPTs = a configured assistant for other peopleDo you actually need either one?
Honestly, maybe not yet.
If you use AI a few times a week for varied tasks, a project is a filing cabinet for things you do not file. What gets you most of the value is one well-written, reusable prompt: who you are, what you need, what format you want back, saved in a doc and pasted when needed. No setup, no subscription decision, works in any AI tool.
I keep 15 reusable prompts that cover most workdays on my start page, and that is where I would point anyone who has not built the habit yet.
The graduation rule is simple: when you catch yourself pasting the same prompt plus the same two or three files more than three times a week, move it into a Project. The repetition is the signal. Before that, the prompt is enough.
When a Project stops being enough
There is a ceiling, and it is worth knowing where it is before you hit it.
A Project still requires you to show up, open the chat, paste the new input, and carry the output somewhere. If you are running the same multi-step routine on a schedule, like every Monday you collect updates, format a report, and send it to the same people, that is not a chat problem anymore. That is a workflow.
The signs you have outgrown the chat window:the input arrives on a schedule, not when you feel like it
the output always goes to the same place
you are the only step in the middleWhen that describes your task, look at the workflows I have documented to see what the next level looks like. And if you are wondering whether you need actual automation tools or are fine staying in the chat, I wrote a sibling decision page for exactly that: ChatGPT vs n8n, on whether you need automation at all. If what you actually want is Claude doing the task instead of chatting about it, the fork is Claude Cowork vs Claude Code.
Most people do not need that level. But knowing the ceiling exists keeps you from forcing a chat tool to do a scheduler's job.
FAQ
Does Claude have an equivalent to Custom GPTs?
Yes. Claude Projects are the closest equivalent: saved instructions plus uploaded reference files that every chat in the project can use. The main thing missing is public sharing. There is no Claude version of the GPT Store.
Can you share a Claude Project the way you share a Custom GPT?
Only inside a Claude Team or Enterprise workspace. There is no public link or store. If you need to hand a configured assistant to people outside your workspace, a Custom GPT is the easier path.
Should I switch from ChatGPT to Claude just to get Projects?
No. ChatGPT has its own Projects feature that covers the same solo use case: files, instructions, and chats grouped in one place. Switching subscriptions for this one feature is not worth it.
Do I need a paid plan to use Claude Projects or Custom GPTs?
Claude includes Projects on the free plan with usage limits. ChatGPT lets anyone use existing Custom GPTs for free, but building your own requires a paid plan.This post is part of Claude at Work, the hub with every plan decision, task comparison, and setup guide for using Claude at your job without code.
Most people asking this question do not need n8n.
Here is the test. If a task happens when you ask for it, and you are sitting there while it happens, a good reusable prompt covers it. n8n earns its place in exactly one situation: the same task has to run on a schedule or a trigger, without you, across two or more apps. That describes far fewer tasks than YouTube makes it sound.
I use both tools every week, so this is not tool loyalty. It is more like "do not buy a forklift to carry groceries."
What is the actual difference between ChatGPT and n8n?
They are not competitors. They do completely different jobs.
ChatGPT (or Claude) is a thinking tool. You give it something, it gives you something back. Summarize these meeting notes. Draft this awkward email. Turn 40 survey responses into the five complaints that actually matter. You are present for every exchange, and that is fine, because the task only exists when you ask for it.
n8n is plumbing. It moves data between apps when something happens. A form gets submitted, so a row lands in a spreadsheet and a Slack message goes out. Nobody is sitting there. That is the entire point of it.
The confusion comes from demo videos where n8n has an AI step in the middle, so it looks like "ChatGPT, but automated." True as far as it goes. But that AI step does the same job ChatGPT does in your browser tab. The real question is never which tool is smarter. The question is whether the task needs to happen without you.
How do you know if you need n8n? Ask 3 questions
Run any task you are thinking about through these:Does it repeat on a schedule or a predictable trigger? Every Monday morning. Every time a form comes in. Every new invoice.
Does it cross two or more apps? Form to spreadsheet to email. Inbox to tracker to Slack.
Does it need to run when you are not watching? Overnight, during meetings, while you are on leave.Three yeses: automation software is worth a look. Two or fewer: a saved prompt almost certainly covers it, and it covers it today, for free, with nothing to maintain.
Real examples:"Summarize my meeting notes into action items." You are present, one app, on demand. Prompt.
"Help me draft replies to difficult parent emails." On demand, you review every word anyway. Prompt.
"Every Friday at 4pm, pull this week's form responses, summarize the complaints, and email me the digest." Scheduled, three apps, runs alone. That is a real n8n job.
"When a new applicant submits the intake form, score them against the role requirements and add the score to my tracking sheet." Triggered, multiple apps, unattended. Also a real n8n job.Notice the pattern. The prompt tasks are about judgment and wording. The automation tasks are about moving the same data the same way, over and over, with nobody watching.
ChatGPT vs n8n: quick comparisonThe task
UseSummarize notes, docs, or transcripts on demand
ChatGPT promptDraft or rewrite emails in your voice
ChatGPT promptTurn messy feedback into clear themes
ChatGPT promptPrep for a meeting from an agenda and past notes
ChatGPT promptMove form responses into a sheet automatically
n8nSend a weekly report without touching it
n8nWatch an inbox and route requests to the right person
n8nScore or sort new entries the same way every time
n8n with an AI stepAnything you do once or twice, ever
Neither. Just do it.If your whole list lands in the top half of that table, you have your answer. Save your prompts and skip the software.
What does n8n really cost if you cannot code?
This is the part the tutorials skip, so let me be the honest friend here.
Money. n8n Cloud starts around $25 a month. The "free" self-hosted version needs a server, usually $5 to $10 a month, plus you become responsible for installing it, updating it, and backing it up. If words like Docker mean nothing to you, self-hosting is not free. It is a part-time hobby.
Heads up: some links in this post are affiliate links — a small kickback to me at no cost to you. I only recommend tools I've actually run.
Setup time. Your first real workflow takes an afternoon, not the 8 minutes the video showed. Most of that time is not building. It is connecting accounts: API keys, permission screens, and figuring out why Google says no. Plan 3 to 4 hours for workflow number one.
Debugging. Workflows break silently. A password-like credential expires. An app changes how it sends data. The workflow you built in March fails quietly in June, and you find out because the Friday report never showed up. Then you are staring at a red error node and a message written for engineers.
Maintenance. Every workflow you build is a small machine you now own. A few simple ones might need an hour a month. But it never goes to zero, and you are the repair person.
None of this means avoid n8n. It means n8n has to earn that overhead. A saved prompt has zero of these costs, which is why it should always be your first move.
When is a reusable prompt all you need?
If you are present when the task happens, the prompt is the automation.
The trick is writing it once, properly, and saving it, instead of improvising a new mediocre prompt every time. A reusable prompt spells out four things: who the AI should act as, what input you will paste in, the exact format you want back, and one example of a good output. That takes 15 minutes to write and pays you back every week after.
Then keep your prompts somewhere you can grab them: a doc, a notes app, whatever you will actually open. If you want a starting set, I keep 15 reusable work prompts here, built for exactly this kind of on-demand work.
The honest comparison: a prompt like this saves you 20 minutes every time you use it, costs nothing, and cannot break while you sleep. That is a high bar for automation software to clear.
When does n8n actually make sense?
Volume and absence. Those are the two things that flip the answer.
If 30 form submissions arrive every week and each one needs the same three steps, you are not doing judgment work anymore. You are being a conveyor belt. Same if the task has to fire at 6am or while you are on vacation. No prompt fixes "I was not there."
If a task passes the 3-question test, the pattern that holds up is boring: a trigger, a step that gathers the data, an AI step for any judgment call, a human approval for anything customer-facing, then route the result. I wrote up that exact pattern in the n8n AI agent workflow if you get to that point.
Also, an option nobody mentions: you do not have to build it yourself. If the task clearly qualifies but the setup sounds miserable, ask IT, or pay a freelancer for a day. A capable n8n freelancer can build a clean first workflow in a day. The building is the cheap part. The maintaining is what you are really signing up for, so decide who owns that before anything gets built.
One more fork in the road: if you or a technical coworker work in code all day, the comparison changes shape. That version of the decision is in Claude Code vs n8n.The 60-second version, from my Shorts: More Automations = Better Results. Or Does It?.What I would do first
Skip the software question for one week and do this instead:List every task you repeat weekly. Most people find 8 to 12.
Run each one through the 3 questions: schedule or trigger, 2+ apps, runs without you.
For everything that fails the test, write and save the prompt today. That is 15 minutes per task.
For the one or two that pass, score them with the automation priority audit before you sign up for anything. It forces the time-saved versus time-spent math that the excitement skips.When I run this with people, the usual result is nine prompt tasks and one genuine automation candidate. That ratio is normal. It is also good news: you can fix most of your repetitive work this afternoon, without buying, hosting, or maintaining anything.
The boring answer wins here. Prompts first. n8n only when a task proves it deserves a machine.
FAQ
Is n8n better than ChatGPT?
Neither is better. They do different jobs. ChatGPT answers when you ask it something. n8n moves work between apps on a schedule or trigger, without you present. Most people only need the first one.
Can ChatGPT replace n8n?
For on-demand tasks where you are present, yes. A saved prompt covers summarizing, drafting, and rewriting. ChatGPT cannot replace n8n for tasks that must run unattended across multiple apps.
Do I need to know how to code to use n8n?
You can build simple workflows without code, but it gets technical fast: API credentials, error logs, and data formats. Budget real time for setup and debugging, or pay someone to build it.
Is n8n free?
The software can be self-hosted free, but you pay for a server, plus your time for setup, updates, and fixing broken workflows. n8n Cloud starts around $25 a month and removes the server work.This post is part of the n8n AI agents hub: definitions, tutorials, workflow patterns, and the build-vs-run decision pages in one place.