Key takeaways
- Anthropic says Claude Code averages about $13 per developer per active day, or $150 to $250 a month.
- In Gartner Peer Insights research, 6% of organisations pay over $2,000 per developer a month for agent tokens.
- At $13 a day, Claude Max 5x pays for itself after about 8 active days a month.
- Several spending caps stay off until you set them, and organisation-wide caps stop every user at once.
- Cache reads and writes make up 80% of the cost of a sample Opus 5.5 session.
As of 5 October 2026, Anthropic puts the average Claude Code cost per month at $150 to $250 per developer, measured across enterprise deployments. That works out at about $13 per active day, and 90% of users stay under $30 a day.1 Spend per engineer has a long tail, though. In Gartner Peer Insights research reported by Computer Weekly, 23% of tech leaders spend $200 to $500 per developer per month on coding-agent tokens, and 6% of organisations pay more than $2,000.2
For an engineer who uses an agent on most working days, a seat plan costs less than API billing at Anthropic’s averages, provided the work fits inside the plan’s limits. The caps that protect a budget are hard limits per person and per workload, plus a separate budget for the most expensive models.
Below are the published spend figures, plan prices and limits for Claude Code, Codex and Cursor, the cost of a single task, and worked budgets for teams of 1, 5 and 25 engineers. After those come the caps that stop spend, the settings that cut token use and a policy template. Prices and limits are as of 5 October 2026.
What teams spend per engineer on coding agents
On 24 June, Gartner predicted that AI coding costs will overtake the average developer’s salary by 2028, as token consumption rises and vendors move from seat licences to consumption-based pricing.3 The same day, Gartner analyst Nitish Tyagi told The Register that bills were rising from $20 or $100 to $2,000 to $5,000 per developer per month, and could reach $20,000 in extreme cases.4
The published figures measure different things, so read the third column before comparing them.
| Source | Figure | What it measures |
|---|---|---|
| Anthropic, Claude Code docs1 | About $13 per active day and $150 to $250 per month; 90% of users under $30 per active day | Average Claude Code cost per developer across enterprise deployments |
| Gartner Peer Insights, via Computer Weekly, 24 June 20262 | 23% spend $200 to $500 per developer per month; 6% spend over $2,000 | Spend on coding-agent tokens reported by tech leaders and organisations |
| Cursor docs5 | $60 to $100 per month for daily agent users; often $200 or more for power users | Total usage on Cursor’s individual plans |
| Databricks, 28 September 20266 | 60% more spend per developer, overnight | A control group given GPT Astra with no cost controls |
The Databricks figure shows how fast spend can move. When the company gave GPT Astra to a control group with no cost mitigations, the average developer spent 60% more than before. Databricks says an overnight rise of that size is very hard to plan for across more than 10,000 users.6 It now gives every user four budgets:
- A monthly maximum across all models.
- A daily runaway limit, which the user can raise in Slack.
- A quality frontier budget, which reserves a fixed share of the monthly budget for GPT Astra and Claude Fable. Databricks says these cost 2 to 3 times the next tier down.
- An experimental budget for new, untested models.
Databricks doesn’t publish the amounts, and the policy template at the end of this post borrows its structure.
Codex pricing vs Claude Code and Cursor plans
All three vendors sell a $20 entry plan and a $200 plan for heavy individual use. OpenAI added a $500 tier, Pro 500, at its DevDay event on 29 September.7 The table lists prices per person per month in US dollars, before tax, as of 5 October 2026.
| Tier | Claude, with Claude Code89 | ChatGPT, with Codex1011 | Cursor125 |
|---|---|---|---|
| Entry | Pro, $20 ($17 billed annually) | Plus, $20 | Pro, $20 |
| Middle | Max 5x, $100 | Pro 100, $100 | Pro+, $60 |
| Heavy | Max 20x, $200 | Pro 200, $200 | Ultra, $200 |
| Top | None | Pro 500, $500 | None |
| Team seat | Standard $25 or Premium $125 ($20 or $100 billed annually) | Business, $25 ($20 billed annually) | Teams Standard $40 or Premium $120 |
| Enterprise | $20 per seat billed annually, plus usage at API rates | Priced by sales | Custom, with pooled usage |
| Past the allowance | Usage credits at API rates | Credits, or an API key at API rates | On-demand usage at API rates |
Each vendor meters included usage differently. OpenAI estimates that a five-hour window on Plus or Standard Business holds 15 to 160 local Codex messages on GPT-6.1 Sol, or 5 to 45 on GPT-6 Astra, depending on the task. Pro plans have no five-hour limit, though weekly limits may apply. Fast mode uses the allowance at 2.5 times the standard rate, and GPT-6 Astra Ultrafast at 8 times.10
At DevDay, OpenAI said Pro 500 includes 25 times the usage of Plus.7 New Pro 200 subscribers get a lower allowance, and earlier subscribers keep the old one until 29 October.11 The Next Web reported that the cut halves Pro 200’s Work and Codex usage.13
Cursor gives each plan a monthly allowance in two pools, one for its own models and one for third-party models billed at API prices, then switches to on-demand usage. On Teams and Enterprise, third-party models also carry a Cursor Token Rate of $0.25 per million tokens.5
Claude Code usage limits since 14 September
Claude subscriptions have a session limit that resets every five hours and a weekly limit on top. Max gives 5 or 20 times Pro’s usage per five-hour session, and a Team Premium seat gives 5 times a Standard seat.8 Anthropic’s pricing page doesn’t state these allowances in tokens or dollars, so a seat can’t be converted into an API budget exactly.
From 13 May to 13 September, Anthropic raised weekly Claude Code limits by 50% on Pro, Max, Team and seat-based Enterprise plans. Since 14 September they have been 25% above their level before the promotion.14 That is about a sixth less than during the promotion (1.25 divided by 1.5 is 0.83), so a team that sized its seats over the summer can now hit weekly limits it didn’t hit in August. The five-hour limits didn’t change.
Fable models have their own rule. On Max and Team Premium seats, Fable 5 and Fable 5.1 can use up to half of the weekly limit, and they use it faster than other models. On Pro and Team Standard seats, Fable runs only on usage credits.15
What one coding task costs
Specific Labs’ Real-SWE benchmark prices attempts at real work. Its published results cover 10 tasks from private production codebases, run eight times per model in each model’s own harness at high reasoning effort. Its estimated cost per attempt:16
| Model | Harness | Estimated cost per attempt |
|---|---|---|
| Gemini 3.8 Flash | Gemini CLI | $2.50 |
| GPT-5.6 Sol | Codex CLI | $2.65 |
| Grok 4.6 | Grok Build | $2.67 |
| Muse Spark 1.3 | Muse Code | $2.74 |
| Kimi K3 | Kimi Code | $3.90, with incomplete usage data |
| GPT-6 Astra | Codex CLI | $4.67 |
| GLM 5.3 | Claude Code | $5.12 |
| Claude Fable 5.1 | Claude Code | $6.96 |
The most expensive model costs 2.8 times as much per attempt as the cheapest. No model passed half of its attempts, so the cost of a finished task is more than twice these figures. Ten tasks is a small sample, so we use the costs and leave the model ranking aside.
Databricks measured cost per session from its gateway traces, comparing early adopters with their own usage a week earlier and reweighting sessions by type. Opus 5.5 averaged $4.23 a session against $5.94 on Opus 4.8, 29% less. GPT-6 Sol averaged $2.34 against $4.52 on GPT-5.6 Sol, 48% less, after a 50% price cut.6 At $4.23 a session, Anthropic’s $13 average day is about three Opus 5.5 sessions, though the two figures come from different groups of users.
Most tokens in an agent session are context that is sent again on every turn and served from the prompt cache. Anthropic’s cost docs show a sample session with 1,200 fresh input tokens, 5,300 output tokens, 940,000 cache reads and 50,000 cache writes.1 We priced those tokens on four Claude models at API rates as of 5 October 2026, with five-minute cache writes:17
| Model | Session cost | Share from cache reads and writes | Session cost relative to Opus 5.5 | Per-token price relative to Opus 5.5 |
|---|---|---|---|---|
| Claude Fable 5.1 | $1.14 | 76% | 2.07x | 2.5x |
| Claude Opus 5.5 | $0.55 | 80% | 1x | 1x |
| Claude Sonnet 5.5 | $0.37 | 85% | 0.67x | 0.5x |
| Claude Haiku 4.5 | $0.18 | 85% | 0.34x | 0.25x |
Cache reads and writes are 80% of the Opus 5.5 total, so anything that breaks the cache costs more than a longer answer does. Opus 5.5 and Sonnet 5.5 both charge $0.20 per million cached tokens read, which is why moving this session to Sonnet saves about a third of the cost. The per-token prices suggest a half.
Claude Code cost per month for teams of 1, 5 and 25
These budgets compare Claude seat plans with API billing, because Anthropic is the only vendor here that publishes an average daily cost. They are illustrations built from the cited figures. We have not measured our own spend for this post. The assumptions:
- A1. A typical engineer costs $200 a month on API billing, the midpoint of Anthropic’s $150 to $250 average. Anthropic’s averages are token costs, which is what API billing charges.1
- A2. A heavy engineer costs $600 a month, which is $30 per active day over 20 working days. Anthropic says 90% of users stay under $30 a day, so this understates the top tenth.
- A3. One engineer in ten is heavy, rounded up. That is one in a team of 5 and three in a team of 25.
- A4. Seats are billed monthly. One person uses Max 5x at $100 or Max 20x at $200, and teams use Team Premium at $125 per seat.89 Team plans need at least two people.
- A5. Seat allowances aren’t published in dollars, so the seat cost is a floor. Engineers who pass their allowance buy usage credits at API rates. The last column shows how much the team can spend on credits before seats cost more than the API.
| Team | API billing | Seat plan | Seat cost | Credits before seats cost more |
|---|---|---|---|---|
| 1 typical engineer | $200 | Max 5x | $100 | $100 |
| 1 typical engineer | $200 | Max 20x | $200 | $0 |
| 1 heavy engineer | $600 | Max 20x | $200 | $400 |
| 5 engineers, 1 heavy | $1,400 | 5 Team Premium seats | $625 | $775 |
| 25 engineers, 3 heavy | $6,200 | 25 Team Premium seats | $3,125 | $3,075 |
For 25 engineers, the API figure is 22 × $200 + 3 × $600 = $6,200, and the seats are 25 × $125 = $3,125. Annual billing lowers the seats to $2,500.
Per engineer, a seat pays for itself once the engineer would have spent its price on the API. At Anthropic’s average of $13 per active day:
| Plan, monthly billing | Price | Active days a month to break even |
|---|---|---|
| Claude Pro | $20 | 1.5 |
| Team Standard | $25 | 1.9 |
| Max 5x | $100 | 7.7 |
| Team Premium | $125 | 9.6 |
| Max 20x | $200 | 15.4 |
Anthropic’s $150 to $250 monthly average implies about 12 to 19 active days at $13 a day. That clears the break-even point for every plan in the table, except Max 20x for engineers at the low end of the range.
API billing has a ceiling of its own. Anthropic’s usage tiers cap monthly API spend at $500 on Start, $1,000 on Build and $200,000 on Scale. An organisation that reaches its tier’s cap is paused until the first day of the next month, unless it gets a higher limit sooner.18 The five-person team above would hit the Build cap before the month ended, and a single heavy engineer would pass the Start cap.
If one of the three heavy engineers in the team of 25 spends $2,000 in a month, the API total rises from $6,200 to $7,600. On seats, that engineer draws usage credits instead, up to whatever per-member limit you set.
Enterprise is a third route, at $20 per seat billed annually plus usage at API rates.8 On our assumptions, 25 engineers would cost $500 plus $6,200, or $6,700 a month. That is more than Team seats here. Enterprise admins can also set the default model, restrict which models each role can use and cap effort per role.19
On these numbers, seats cost less for anyone who uses an agent on most working days. API billing suits occasional users and automation. We would buy seats for daily users, run CI and scheduled jobs on API keys in a workspace of their own, and give the heavy tail usage credits with a per-member limit.
Which spending caps are hard, by vendor
By a hard cap we mean a limit that makes requests fail once spend reaches it. Alerts and budgets that only send email don’t count. Simon Willison argued on 3 October that hard caps should be the default for anything billed by usage, because coding agents make it easy to build things that run up bills.20 As of 5 October 2026, every product here has a hard stop somewhere, but several leave it off until you set it, and some apply it to the whole organisation at once.
| Product | What stops spend | On by default | Notes |
|---|---|---|---|
| Claude Pro and Max21 | Plan limits, then prepaid usage credits with an optional monthly limit | Plan limits yes; credits stay off until enabled | The credit limit can be set to unlimited, and auto-reload refills the balance |
| Claude Team1 | Seat allowance, then usage credits with limits per organisation, group or member | Seat allowance yes; credit limits only once set | Unless an admin enables credits, the allowance is the ceiling |
| Claude Enterprise22 | Spend limits per organisation, group or member | Only once set | Anthropic’s guide: “No limit anywhere = no limit” |
| Claude API18 | Tier cap, plus your own organisation or workspace limit | Tier cap yes; your own limit no | Tier cap returns 429 until the 1st of the next month; your own limit returns 400 |
| ChatGPT Plus and Pro11 | Plan allowance, then optional credits | Yes | OpenAI says a limit can’t be raised through a setting |
| ChatGPT Business23 | Per-seat limits, then pooled workspace credits | Yes, until an owner buys credits | Auto-recharge refills the pool; credit alerts only notify |
| OpenAI API24 | OpenAI’s tier-based usage limit, plus your own hard spend limit per organisation or project | Tier limit yes; your hard limit no | 429 when reached; enforcement isn’t instantaneous |
| Cursor Pro, Pro+ and Ultra25 | Spend limit on on-demand usage | On-demand stays off until enabled | Usage can briefly pass the limit |
| Cursor Teams2627 | Team-wide monthly spend limit | On-demand is on by default | Per-member limits need Enterprise |
| AWS projects28 | Project spend limit that pauses the project | No, and in limited release | Designed for experiments and sandbox workloads |
The Claude Code spending cap on API billing is a workspace limit. The first time someone signs in to Claude Code with a Claude Console account, the Console creates a “Claude Code” workspace for that usage.1 Give it a spend limit below the organisation’s. A limit you set returns a 400 error, the tier cap returns a 429 with no retry-after header, and Claude Code’s workspace can return a 429 that does carry one.18 Code that retries on 429 needs to tell them apart.
Organisation-wide caps stop everyone at once. For that reason, Anthropic’s Enterprise guide recommends starting with group and per-user limits.22
Team defaults differ. Cursor Teams turns on-demand usage on by default and keeps per-member limits for Enterprise. Unless an admin turns on “Only Admins Can Edit Usage Settings”, any team member can change the team-level limit.2627
Enforcement lags. OpenAI says its hard limits aren’t instantaneous, Cursor says usage can briefly pass a limit, and Anthropic says usage can go slightly over before it pauses.242622 How a backfill bug ran up a $1,100 LLM vision bill covers why provider caps make a late backstop and which guardrails to add in your own code.
Auto-reload removes the stop. Claude’s prepaid usage credits and ChatGPT Business’s pooled credits act as a cap only until someone turns on auto-reload or automatic recharge.2123
Cloud providers need their own controls. For Claude Code on Amazon Bedrock and other clouds, Anthropic’s docs point to the cloud’s budget controls, or to a self-hosted gateway with per-user spend limits.1 AWS’s new project spend limits pause a project and stop its resources, but AWS is still releasing them to a limited number of customers.28
How to reduce Claude Code token usage
Gartner advises routing simple, high-frequency tasks to smaller models, training developers to send only relevant context, and reviewing the heaviest workflows in sprint retrospectives.3 The vendors’ docs turn that advice into settings.
Default model and effort. As of 5 October 2026, Claude Code’s default model is Opus 5.5 on every Claude plan and on the API, at medium effort.19 Within one model, effort moves cost by several times. On Artificial Analysis’s test suite, Opus 5.5 cost $1.34 per task at medium, $1.82 at high, $3.46 at xhigh and $5.98 at max.29 Those are general reasoning tasks, so treat the ratio as a guide: max cost 4.5 times medium. Thinking tokens are billed as output, and Anthropic says the default thinking budget can run to tens of thousands of tokens per request.1
Admins on any plan can cap effort with the maxEffortLevel managed setting and restrict models with availableModels.19 Codex has the equivalent model_reasoning_effort setting, and OpenAI’s advice is to start at the default effort and raise it only when a task needs deeper planning.3031 Anthropic’s docs say Sonnet handles most coding tasks well at a lower price than Opus, though the session table above puts the saving nearer a third than a half.1
Caching. Anthropic charges 1.25 times the input price to write to its five-minute cache and 2 times for the one-hour cache. A cache read costs a tenth of the input price, or less on Opus 5.5 and Fable 5.1.17 Claude Code keeps a one-hour cache on subscriptions, which drops to five minutes once you draw on usage credits. API keys default to five minutes. The first message after a longer break reprocesses the whole context.1 A developer on API billing who steps away for ten minutes pays to rebuild the cache when they come back. Codex credit billing has no separate charge for cache writes.10
Subagent ceilings. Subagents and agent teams send their own requests on top of the main conversation. Anthropic says agent teams use about 7 times the tokens of a standard session when teammates run in plan mode.1 Each Claude Code subagent definition can set its own model, effort and maxTurns, so routine searches can run on a smaller model with a turn limit.32 For scripted runs, --max-budget-usd stops a print-mode run at a dollar amount. Subagent spend counts toward it, and no new subagents start once it is reached.33 Codex caps concurrent spawned agents with agents.max_concurrent_threads_per_session.30
A frontier-model sub-budget. Claude Fable 5.1 lists at 2.5 times the per-token price of Opus 5.5, and GPT-6 Astra Ultrafast uses a ChatGPT allowance at 8 times the standard rate.1711 Databricks puts its premium tier at 2 to 3 times the cost of the next one down, and gives it a fixed share of each user’s monthly budget.6 Anthropic does something similar on Max, where Fable can take at most half of the weekly limit.15
Context hygiene. Claude Code sends the conversation with every request, so stale context costs money on every turn. Anthropic’s docs recommend /clear between unrelated tasks, a CLAUDE.md under 200 lines with workflow instructions moved into skills, and turning off MCP servers you aren’t using. They also suggest CLI tools such as gh in place of MCP servers, and hooks that cut test output down before Claude reads it.1 Scheduled tasks send the full context each time they fire. OpenAI’s Codex guidance is the same. It suggests nesting AGENTS.md files, limiting MCP servers and giving the agent only the files it needs.10
To see which of these matter for your team, log tokens per model and per kind of work. Why our LLM pipeline reported 0 dropped shows a ledger that records input, cached, output and reasoning tokens for every call, which is the data a monthly spend review needs.
How the figures were gathered
- Prices, limits and controls come from each vendor’s pricing pages, help centre articles and documentation as they stood on 5 October 2026. Prices are in US dollars before tax, with monthly billing unless stated.
- The spend figures from Anthropic, Cursor and Databricks are self-reported. The 23% and 6% figures come from Computer Weekly’s report of Gartner Peer Insights research, and neither it nor Gartner’s press release gives a sample size.
- Real-SWE’s costs are Specific Labs’ estimates over 10 tasks with eight attempts each. Artificial Analysis’s cost per task is for its general test suite, and we use it only for the ratio between effort levels.
- The session table and the worked budgets are our arithmetic on those inputs, with the assumptions labelled. We have not measured our own spend per engineer for this post.
- Not covered: GitHub Copilot, Gemini CLI plans, enterprise discounts, taxes and regional pricing.
A spend policy template for coding agents
Copy this, replace the values in brackets, and review it after the first month.
- Seats. Engineers who use an agent on most working days get [seat type]. Everyone else uses [lighter seat or API key]. Seat types are reviewed each quarter against usage.
- Defaults. The default model is [model] at [effort level], enforced in managed settings. Higher effort is a per-task choice.
- Hard limits per person. Each engineer has a monthly limit of [amount] and a daily limit of [amount]. A team lead can raise the daily limit for the rest of that day.
- Frontier budget. Up to [share] of each monthly limit can go on the most expensive models and speed modes, such as Fable, GPT-6 Astra and Ultrafast.
- Automation. CI jobs, scheduled agents and scripts run on API keys in their own workspace or project, with a hard monthly limit and a budget per run.
- No unlimited settings. An unlimited credit limit or auto-reload needs sign-off from [role].
- Alerts. The engineer and the budget owner get alerts at 50%, 75% and 90% of every limit.
- Monthly review. Spend is reviewed by engineer, model and kind of work. Anyone over [amount] talks with their lead about what the spend produced before the limit changes.
-
Anthropic, “Manage costs effectively”, Claude Code docs, https://code.claude.com/docs/en/costs ↩↩↩↩↩↩↩↩↩↩↩↩
-
Computer Weekly, “Gartner: AI coding agents will cost more than real developers”, 24 June 2026, https://www.computerweekly.com/news/366645054/Gartner-AI-coding-agents-will-cost-more-than-real-developers ↩↩
-
Gartner, “Gartner Predicts AI Coding Costs Will Surpass Average Developer’s Salary by 2028 as Token Consumption Surges”, 24 June 2026, https://www.gartner.com/en/newsroom/press-releases/2026-06-24-gartner-predicts-ai-coding-costs-will-surpass-average-developer-salary-by-2028-as-token-consumption-surges ↩↩
-
The Register, “AI coding agents could soon cost more than the developers using them”, 24 June 2026, https://www.theregister.com/ai-and-ml/2026/06/24/ai-coding-agents-could-soon-cost-more-than-the-developers-using-them/5260864 ↩
-
Cursor, “Models & Pricing”, Cursor Docs, https://cursor.com/docs/account/pricing ↩↩↩
-
Databricks, “How Databricks rolls out frontier models to 12,000 employees on Day 1”, 28 September 2026, https://www.databricks.com/blog/how-databricks-rolls-out-frontier-models-14000-employees-day-1 ↩↩↩↩
-
Simon Willison, “OpenAI DevDay 2026 live blog”, 29 September 2026, https://simonwillison.net/2026/Sep/29/openai-devday-2026-live-blog/ ↩↩
-
Anthropic, “Plans & Pricing”, https://claude.com/pricing ↩↩↩↩
-
Anthropic, “What is the Max plan?”, Claude Help Center, https://support.claude.com/en/articles/11049741-what-is-the-max-plan ↩↩
-
OpenAI, “Pricing”, Codex docs, https://developers.openai.com/codex/pricing ↩↩↩↩
-
OpenAI, “About ChatGPT Pro tiers”, OpenAI Help Center, https://help.openai.com/en/articles/9793128-about-chatgpt-pro-tiers ↩↩↩↩
-
Cursor, “Pricing”, https://cursor.com/pricing ↩
-
The Next Web, “OpenAI halves Pro 200 usage and launches a $500 ChatGPT plan at DevDay”, 29 September 2026, https://thenextweb.com/news/openai-devday-pro-200-usage-cut-pro-500-plan ↩
-
Anthropic, “Claude Code May–August 2026 weekly limits promotion”, Claude Help Center, https://support.claude.com/en/articles/15910845-claude-code-may-august-2026-weekly-limits-promotion ↩
-
Anthropic, “Claude Fable models on your plan”, Claude Help Center, https://support.claude.com/en/articles/15424964-claude-fable-models-on-your-plan ↩↩
-
Specific Labs, “Introducing Real-SWE”, September 2026, https://withspecific.com/benchmarks/real-swe ↩
-
Anthropic, “Pricing”, Claude API docs, https://platform.claude.com/docs/en/about-claude/pricing ↩↩↩
-
Anthropic, “Rate limits”, Claude API docs, https://platform.claude.com/docs/en/api/rate-limits ↩↩↩
-
Anthropic, “Model configuration”, Claude Code docs, https://code.claude.com/docs/en/model-config ↩↩↩
-
Simon Willison, “We’re going to need default hard budget caps on pretty much everything”, 3 October 2026, https://simonwillison.net/2026/Oct/3/default-hard-budget-caps/ ↩
-
Anthropic, “Manage usage credits for paid Claude plans”, Claude Help Center, https://support.claude.com/en/articles/12429409-extra-usage-for-paid-claude-plans ↩↩
-
Anthropic, “Claude Enterprise consumption guide”, Claude Help Center, https://support.claude.com/en/articles/14782391-claude-enterprise-consumption-guide ↩↩↩
-
OpenAI, “Flexible pricing for the Enterprise, Edu, and Business plans”, OpenAI Help Center, https://help.openai.com/en/articles/11487671 ↩↩
-
OpenAI, “Spend limits”, OpenAI API docs, https://developers.openai.com/api/docs/guides/spend-limits ↩↩
-
Cursor, “Usage-based charges”, Cursor Docs, https://cursor.com/help/account-and-billing/overages ↩
-
Cursor, “Spend limits”, Cursor Docs, https://cursor.com/help/account-and-billing/spend-limits ↩↩↩
-
Cursor, “Team Pricing”, Cursor Docs, https://cursor.com/docs/account/teams/pricing ↩↩
-
AWS, “Create a spend limit in AWS Settings”, AWS Account Management docs, https://docs.aws.amazon.com/accounts/latest/reference/create-spend-limit.html ↩↩
-
Artificial Analysis, model pages for Claude Opus 5.5 at medium, high, xhigh and max effort, https://artificialanalysis.ai/models/claude-opus-5-5 ↩
-
OpenAI, “Configuration Reference”, Codex docs, https://developers.openai.com/codex/config-reference ↩↩
-
OpenAI, “Models”, Codex docs, https://developers.openai.com/codex/models ↩
-
Anthropic, “Create custom subagents”, Claude Code docs, https://code.claude.com/docs/en/sub-agents ↩
-
Anthropic, “CLI reference”, Claude Code docs, https://code.claude.com/docs/en/cli-reference ↩
Frequently asked questions
How much does Claude Code cost per month?
As of 5 October 2026, Claude plans that include Claude Code cost $20 a month for Pro, $100 or $200 for Max, and $25 or $125 per Team seat with monthly billing. Anthropic says Claude Code averages about $13 per developer per active day in token costs across enterprise deployments, or $150 to $250 per developer per month.
Is Codex cheaper than Claude Code?
On list prices they match at $20, $100 and $200 a month, and ChatGPT adds Pro 500 at $500. The vendors meter usage differently, so compare the cost of finished tasks on your own work. On the Real-SWE benchmark, estimated cost per attempt ranged from $2.50 to $6.96 across eight models and harnesses.
What are the Claude Code usage limits?
Claude subscriptions have a session limit that resets every five hours and a weekly limit. Since 14 September 2026, weekly Claude Code limits on Pro, Max, Team and seat-based Enterprise plans are 25% above their level before the May to September promotion, which had raised them by 50%.
How do I set a spending cap on Claude Code?
On API billing, set a spend limit on the Claude Code workspace in the Claude Console. On Team and Enterprise, set usage-credit limits per organisation, group or member. On Pro and Max, keep usage credits off or give them a monthly limit. For scripted runs, the --max-budget-usd flag stops a run at a dollar amount.
How do I reduce Claude Code token usage?
Clear the session between unrelated tasks, keep CLAUDE.md short, turn off MCP servers you don’t use, give subagents a smaller model and lower the effort level for routine work. On Artificial Analysis’s tests, Opus 5.5 cost $1.34 per task at medium effort and $5.98 at max.
Is Claude Max cheaper than the API?
For daily use, usually. At Anthropic’s average of $13 per active day, Max 5x at $100 pays for itself after about 8 active days a month and Max 20x at $200 after about 15, provided your work fits inside the plan’s limits. Occasional users pay less on the API.
Building something like this?
9io is a small team of senior engineers with a fractional CTO, and we work by the hour. Send us a note about your product. The reply comes from the person who'd do the work.