• bitcoinBitcoin(BTC)$80,279.00-1.12%
  • ethereumEthereum(ETH)$2,578.18-1.69%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$749.48-1.49%
  • rippleXRP(XRP)$1.38-2.57%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$108.13-4.66%
  • tronTRON(TRX)$0.3395370.30%
  • zcashZcash(ZEC)$1,460.06-4.85%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.02-1.50%
  • HyperliquidHyperliquid(HYPE)$90.09-4.06%
  • dogecoinDogecoin(DOGE)$0.085434-2.83%
  • moneroMonero(XMR)$529.12-7.98%
  • whitebitWhiteBIT Coin(WBT)$81.80-1.74%
  • RainRain(RAIN)$0.0136321.74%
  • USDSUSDS(USDS)$1.00-0.02%
  • chainlinkChainlink(LINK)$12.01-3.17%
  • cardanoCardano(ADA)$0.220467-4.66%
  • leo-tokenLEO Token(LEO)$8.90-0.12%
  • stellarStellar(XLM)$0.190159-2.10%
  • uniswapUniswap(UNI)$8.69-4.07%
  • bitcoin-cashBitcoin Cash(BCH)$244.46-4.64%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • daiDai(DAI)$1.000.00%
  • nearNEAR Protocol(NEAR)$3.51-7.53%
  • litecoinLitecoin(LTC)$56.94-2.92%
  • USD1USD1(USD1)$1.00-0.02%
  • avalanche-2Avalanche(AVAX)$9.5311.69%
  • CantonCanton(CC)$0.105143-7.11%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.37-0.14%
  • MemeCoreMemeCore(M)$1.5823.55%
  • hedera-hashgraphHedera(HBAR)$0.0799610.72%
  • suiSui(SUI)$0.82-0.52%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-1.74%
  • crypto-com-chainCronos(CRO)$0.058269-1.59%
  • BittensorBittensor(TAO)$253.27-1.77%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • tether-goldTether Gold(XAUT)$4,366.81-0.18%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • okbOKB(OKB)$115.25-1.68%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.04%
  • aaveAave(AAVE)$137.41-4.97%
  • mantleMantle(MNT)$0.60-2.71%
  • AsterAster(ASTER)$0.74-5.80%
  • OndoOndo(ONDO)$0.403846-1.10%
  • EthenaEthena(ENA)$0.1923408.76%
  • pax-goldPAX Gold(PAXG)$4,359.41-0.18%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Hidden costs in AI deployment: Why Claude models may be 20-30% more expensive than GPT in enterprise settings

May 1, 2025
in AI & Technology
Reading Time: 5 mins read
A A
Hidden costs in AI deployment: Why Claude models may be 20-30% more expensive than GPT in enterprise settings
ShareShareShareShareShare

It is a well-known fact that different model families can use different tokenizers. However, there has been limited analysis on how the process of “tokenization” itself varies across these tokenizers. Do all tokenizers result in the same number of tokens for a given input text? If not, how different are the generated tokens? How significant are the differences?

In this article, we explore these questions and examine the practical implications of tokenization variability. We present a comparative story of two frontier model families: OpenAI’s ChatGPT vs Anthropic’s Claude. Although their advertised “cost-per-token” figures are highly competitive, experiments reveal that Anthropic models can be 20–30% more expensive than GPT models.

API Pricing — Claude 3.5 Sonnet vs GPT-4o

As of June 2024, the pricing structure for these two advanced frontier models is highly competitive. Both Anthropic’s Claude 3.5 Sonnet and OpenAI’s GPT-4o have identical costs for output tokens, while Claude 3.5 Sonnet offers a 40% lower cost for input tokens.

Source: Vantage

The hidden “tokenizer inefficiency”

Despite lower input token rates of the Anthropic model, we observed that the total costs of running experiments (on a given set of fixed prompts) with GPT-4o is much cheaper when compared to Claude Sonnet-3.5.

Why?

The Anthropic tokenizer tends to break down the same input into more tokens compared to OpenAI’s tokenizer. This means that, for identical prompts, Anthropic models produce considerably more tokens than their OpenAI counterparts. As a result, while the per-token cost for Claude 3.5 Sonnet’s input may be lower, the increased tokenization can offset these savings, leading to higher overall costs in practical use cases. 

This hidden cost stems from the way Anthropic’s tokenizer encodes information, often using more tokens to represent the same content. The token count inflation has a significant impact on costs and context window utilization.

Domain-dependent tokenization inefficiency

Different types of domain content are tokenized differently by Anthropic’s tokenizer, leading to varying levels of increased token counts compared to OpenAI’s models. The AI research community has noted similar tokenization differences here. We tested our findings on three popular domains, namely: English articles, code (Python) and math.

DomainModel InputGPT TokensClaude Tokens% Token Overhead
English articles7789~16%
Code (Python)6078~30%
Math114138~21%

% Token Overhead of Claude 3.5 Sonnet Tokenizer (relative to GPT-4o) Source: Lavanya Gupta

When comparing Claude 3.5 Sonnet to GPT-4o, the degree of tokenizer inefficiency varies significantly across content domains. For English articles, Claude’s tokenizer produces approximately 16% more tokens than GPT-4o for the same input text. This overhead increases sharply with more structured or technical content: for mathematical equations, the overhead stands at 21%, and for Python code, Claude generates 30% more tokens.

This variation arises because some content types, such as technical documents and code, often contain patterns and symbols that Anthropic’s tokenizer fragments into smaller pieces, leading to a higher token count. In contrast, more natural language content tends to exhibit a lower token overhead.

Other practical implications of tokenizer inefficiency

Beyond the direct implication on costs, there is also an indirect impact on the context window utilization.  While Anthropic models claim a larger context window of 200K tokens, as opposed to OpenAI’s 128K tokens, due to verbosity, the effective usable token space may be smaller for Anthropic models. Hence, there could potentially be a small or large difference in the “advertised” context window sizes vs the “effective” context window sizes.

Implementation of tokenizers

GPT models use Byte Pair Encoding (BPE), which merges frequently co-occurring character pairs to form tokens. Specifically, the latest GPT models use the open-source o200k_base tokenizer. The actual tokens used by GPT-4o (in the tiktoken tokenizer) can be viewed here.

JSON
 
{
    #reasoning
    "o1-xxx": "o200k_base",
    "o3-xxx": "o200k_base",

    # chat
    "chatgpt-4o-": "o200k_base",
    "gpt-4o-xxx": "o200k_base",  # e.g., gpt-4o-2024-05-13
    "gpt-4-xxx": "cl100k_base",  # e.g., gpt-4-0314, etc., plus gpt-4-32k
    "gpt-3.5-turbo-xxx": "cl100k_base",  # e.g, gpt-3.5-turbo-0301, -0401, etc.
}

Unfortunately, not much can be said about Anthropic tokenizers as their tokenizer is not as directly and easily available as GPT. Anthropic released their Token Counting API in Dec 2024. However, it was soon demised in later 2025 versions.

Latenode reports that “Anthropic uses a unique tokenizer with only 65,000 token variations, compared to OpenAI’s 100,261 token variations for GPT-4.” This Colab notebook contains Python code to analyze the tokenization differences between GPT and Claude models. Another tool that enables interfacing with some common, publicly available tokenizers validates our findings.

The ability to proactively estimate token counts (without invoking the actual model API) and budget costs is crucial for AI enterprises. 

Key Takeaways

  • Anthropic’s competitive pricing comes with hidden costs:
    While Anthropic’s Claude 3.5 Sonnet offers 40% lower input token costs compared to OpenAI’s GPT-4o, this apparent cost advantage can be misleading due to differences in how input text is tokenized.
  • Hidden “tokenizer inefficiency”:
    Anthropic models are inherently more verbose. For businesses that process large volumes of text, understanding this discrepancy is crucial when evaluating the true cost of deploying models.
  • Domain-dependent tokenizer inefficiency:
    When choosing between OpenAI and Anthropic models, evaluate the nature of your input text. For natural language tasks, the cost difference may be minimal, but technical or structured domains may lead to significantly higher costs with Anthropic models.
  • Effective context window:
    Due to the verbosity of Anthropic’s tokenizer, its larger advertised 200K context window may offer less effective usable space than OpenAI’s 128K, leading to a potential gap between advertised and actual context window.

Anthropic did not respond to VentureBeat’s requests for comment by press time. We’ll update the story if they respond.

Daily insights on business use cases with VB Daily

YOU MAY ALSO LIKE

How Long Can You Expect Your Old Cassette Tapes To Last?

How To Record Audio On Your iPhone

If you want to impress your boss, VB Daily has you covered. We give you the inside scoop on what companies are doing with generative AI, from regulatory shifts to practical deployments, so you can share insights for maximum ROI.

Read our Privacy Policy

Thanks for subscribing. Check out more VB newsletters here.

An error occured.

Credit: Source link
ShareTweetSendSharePin

Related Posts

How Long Can You Expect Your Old Cassette Tapes To Last?
AI & Technology

How Long Can You Expect Your Old Cassette Tapes To Last?

September 20, 2026
How To Record Audio On Your iPhone
AI & Technology

How To Record Audio On Your iPhone

September 20, 2026
OpenClaw Releases 2026.9.5 With Atomic Updates, Plugin Hot Reload, Conversation Sharing, and Expanded GPT Live
AI & Technology

OpenClaw Releases 2026.9.5 With Atomic Updates, Plugin Hot Reload, Conversation Sharing, and Expanded GPT Live

September 19, 2026
Trump Proposes Renaming Artificial Intelligence, Announces AI Force – Unite.AI
AI & Technology

Trump Proposes Renaming Artificial Intelligence, Announces AI Force – Unite.AI

September 19, 2026
Next Post
‘We are not the enemy of the people’: Journalists honored at White House Correspondents dinner

'We are not the enemy of the people': Journalists honored at White House Correspondents dinner

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
In Havana, church meals become a lifeline for many

In Havana, church meals become a lifeline for many

September 19, 2026
Witkoff and Kushner travel to Russia and Ukraine

Witkoff and Kushner travel to Russia and Ukraine

September 16, 2026
27,000 Adecco Group Employees Gain Access To Salesforce’s AI Teammate – Unite.AI

27,000 Adecco Group Employees Gain Access To Salesforce’s AI Teammate – Unite.AI

September 15, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!