• bitcoinBitcoin(BTC)$64,748.000.40%
  • ethereumEthereum(ETH)$1,920.570.20%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$592.923.80%
  • usd-coinUSDC(USDC)$1.000.00%
  • rippleXRP(XRP)$1.090.10%
  • solanaSolana(SOL)$74.620.60%
  • tronTRON(TRX)$0.3289591.10%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.010.60%
  • whitebitWhiteBIT Coin(WBT)$56.480.30%
  • HyperliquidHyperliquid(HYPE)$55.260.10%
  • dogecoinDogecoin(DOGE)$0.070634-0.30%
  • USDSUSDS(USDS)$1.000.00%
  • RainRain(RAIN)$0.013359-2.00%
  • leo-tokenLEO Token(LEO)$9.780.00%
  • zcashZcash(ZEC)$473.681.40%
  • moneroMonero(XMR)$364.262.40%
  • cardanoCardano(ADA)$0.1720102.60%
  • chainlinkChainlink(LINK)$8.471.00%
  • stellarStellar(XLM)$0.1729040.00%
  • CantonCanton(CC)$0.1209301.80%
  • daiDai(DAI)$1.000.00%
  • bitcoin-cashBitcoin Cash(BCH)$218.383.40%
  • USD1USD1(USD1)$1.000.00%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.420.30%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • litecoinLitecoin(LTC)$45.43-0.40%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • hedera-hashgraphHedera(HBAR)$0.0686000.10%
  • Circle USYCCircle USYC(USYC)$1.13-0.10%
  • suiSui(SUI)$0.700.40%
  • avalanche-2Avalanche(AVAX)$6.46-0.20%
  • uniswapUniswap(UNI)$4.3810.40%
  • shiba-inuShiba Inu(SHIB)$0.000005-2.30%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.054918-0.50%
  • tether-goldTether Gold(XAUT)$4,101.910.20%
  • nearNEAR Protocol(NEAR)$1.693.20%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.20%
  • OndoOndo(ONDO)$0.4211132.40%
  • BittensorBittensor(TAO)$193.530.20%
  • pax-goldPAX Gold(PAXG)$4,105.660.10%
  • okbOKB(OKB)$85.771.10%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.054976-1.10%
  • AsterAster(ASTER)$0.611.50%
  • HTX DAOHTX DAO(HTX)$0.0000020.90%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • usddUSDD(USDD)$1.000.00%
  • aaveAave(AAVE)$99.28-0.20%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

OpenAI Cuts API Prices on Its Two Cheaper GPT-5.6 Tiers – Unite.AI

July 30, 2026
in AI & Technology
Reading Time: 3 mins read
A A
OpenAI Cuts API Prices on Its Two Cheaper GPT-5.6 Tiers – Unite.AI
ShareShareShareShareShare

OpenAI lowered the API price of its two lower-cost GPT-5.6 models on July 30, 2026, cutting the cheapest tier by 80% and the mid-tier by 20% while leaving its flagship untouched. The change is logged in the company’s own API changelog and is already live on the published rate card.

YOU MAY ALSO LIKE

Big Tech Makes the Case for Open-Weight AI | Bloomberg Tech 7/24/2026

Scribe Bets on a New Era of Gene Editing

Per million input and output tokens, the standard rates now read:

  • GPT-5.6 Luna: 20 cents and $1.20, down from $1 and $6
  • GPT-5.6 Terra: $2 and $12, down from $2.50 and $15
  • GPT-5.6 Sol: $5 and $30, unchanged, matching the rate its predecessor GPT-5.5 still carries

All three tiers reached general availability on July 9, 2026 at the higher prices, which puts the repricing three weeks into the family’s commercial life.

The cuts run through every service tier on the sheet, not just the headline rate. Batch and Flex processing, both half the standard price, now put Luna at 10 cents input and 60 cents output. Cached input reads, discounted 90%, drop to two cents per million tokens on Luna and 20 cents on Terra. Long-context requests, which bill at double the input rate and 1.5 times output, land at 40 cents and $1.80 for Luna. Buyers going through Amazon (AMZN ) Bedrock are billed by AWS, and OpenAI notes those rates can differ from its own.

The cheaper tiers are where high-volume production traffic lives: classification, extraction, request routing, first-pass drafting, and the long agent loops where one user instruction can trigger dozens of model calls before it returns an answer. A five-fold cut on the tier absorbing that volume changes the arithmetic on which workloads are worth automating at all.

Where the cheaper tiers sit against Claude

At 20 cents in and $1.20 out, Luna undercuts Anthropic’s cheapest published model, Haiku 4.5, by a factor of five on input and roughly four on output, according to Anthropic’s pricing page. Terra’s new rate sits below the $3 and $15 that Claude Sonnet 5 is scheduled to charge once its introductory rate of $2 and $10 lapses on August 31, 2026. At the top of both lineups, Sol still costs more on output than Opus 5, which Anthropic prices at $5 and $25.

Priority processing becomes Fast mode

The same changelog entry retires Priority Processing and replaces it with Fast mode. For Sol, OpenAI says Fast mode runs up to 2.5 times standard speed at twice the price, and the switch is backward compatible: requests already tagged for priority route to Fast mode without a code change. Fast-mode rates are $10 and $60 for Sol, $4 and $24 for Terra, and 40 cents and $2.40 for Luna.

Anthropic sells the same product under the same name and the same terms — fast mode for Opus 5, up to 2.5 times faster at twice standard pricing. The two rate cards now converge on region-pinned inference as well. OpenAI charges a 10% uplift on models released on or after March 5, 2026 when a customer requires data residency; Anthropic bills US-only inference at 1.1 times its standard rate.

What made the cheaper tiers cheaper

OpenAI published its accounting a day before the cut. In a July 29, 2026 engineering post, five members of its technical staff described optimizations across inference and the agent harness behind Codex and ChatGPT Work. Sol, running inside Codex, rewrote the company’s production GPU kernels; combined with broader kernel work, OpenAI says that cut end-to-end serving costs by 20%. Sol also redesigned its own speculative-decoding draft model across hundreds of experiments, which the company credits with raising token-generation efficiency by more than 15%.

Those are OpenAI’s own figures for its own stack. The post closed on a commitment to pass “under-the-hood improvements back to our users and customers in the form of more widely available, cost-efficient intelligence.” The rate card followed a day later.

The harness work points at the same cost center the price cut does. OpenAI caps tool output at 10,000 tokens by default and keeps model-visible history append-only, so an agent loop resending its instructions and tool definitions at every step hits the prompt cache instead of paying full input rates. On a task that takes 30 model requests, that is 30 chances to avoid recomputing the same prefix.

The cuts land while buyers are auditing inference spend and widening which teams touch the models: OpenAI’s own research found staff using ChatGPT well beyond their job titles. Amazon’s engineering organization moved to cap AI spending after cost overruns the same day, and OpenAI shipped hard spend limits for API organizations and projects on July 22, 2026, letting administrators set a monthly ceiling that starts returning errors once tracked spend reaches it.

For a team already routing bulk work to Luna, the same calls now cost a fifth of what they did, with a ceiling they can set in the dashboard. With Sol’s price unchanged, the live decision stays where OpenAI has put it since the family shipped: which tier each request requires.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Big Tech Makes the Case for Open-Weight AI | Bloomberg Tech 7/24/2026
AI & Technology

Big Tech Makes the Case for Open-Weight AI | Bloomberg Tech 7/24/2026

July 30, 2026
Scribe Bets on a New Era of Gene Editing
AI & Technology

Scribe Bets on a New Era of Gene Editing

July 30, 2026
‘The Real Debate Isn’t Open-Weight AI, It’s Chinese AI’: Krach Institute’s Giuda
AI & Technology

‘The Real Debate Isn’t Open-Weight AI, It’s Chinese AI’: Krach Institute’s Giuda

July 30, 2026
SpaceX Is Betting Its Future on Starship
AI & Technology

SpaceX Is Betting Its Future on Starship

July 30, 2026
Next Post
‘Inference Speed Makes Markets Bigger,’ says Cerebras CEO

'Inference Speed Makes Markets Bigger,' says Cerebras CEO

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Iranian bridges damaged after sixth night of U.S. strikes

Iranian bridges damaged after sixth night of U.S. strikes

July 23, 2026
NYPD reach woman trying to jump from Brooklyn Bridge

NYPD reach woman trying to jump from Brooklyn Bridge

July 30, 2026
Meta faces backlash over new AI feature

Meta faces backlash over new AI feature

July 29, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!