• bitcoinBitcoin(BTC)$83,006.00-2.13%
  • ethereumEthereum(ETH)$2,645.68-2.66%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$761.68-2.14%
  • rippleXRP(XRP)$1.48-4.22%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$118.23-4.94%
  • tronTRON(TRX)$0.333580-0.02%
  • zcashZcash(ZEC)$1,547.32-6.99%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.060.00%
  • HyperliquidHyperliquid(HYPE)$88.92-4.77%
  • dogecoinDogecoin(DOGE)$0.092738-5.34%
  • chainlinkChainlink(LINK)$13.71-4.77%
  • moneroMonero(XMR)$528.64-5.40%
  • whitebitWhiteBIT Coin(WBT)$82.74-2.19%
  • USDSUSDS(USDS)$1.00-0.01%
  • cardanoCardano(ADA)$0.243215-5.49%
  • RainRain(RAIN)$0.012507-1.60%
  • leo-tokenLEO Token(LEO)$9.010.02%
  • stellarStellar(XLM)$0.207718-5.42%
  • nearNEAR Protocol(NEAR)$5.13-4.66%
  • bitcoin-cashBitcoin Cash(BCH)$308.10-10.89%
  • uniswapUniswap(UNI)$9.11-9.28%
  • CantonCanton(CC)$0.1379031.47%
  • litecoinLitecoin(LTC)$70.42-2.21%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • avalanche-2Avalanche(AVAX)$10.42-6.80%
  • suiSui(SUI)$1.19-6.36%
  • daiDai(DAI)$1.000.02%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.632.07%
  • USD1USD1(USD1)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.0974161.68%
  • quant-networkQuant(QNT)$265.3854.65%
  • BitwayBitway(BTW)$1.2921.74%
  • BittensorBittensor(TAO)$302.89-9.25%
  • shiba-inuShiba Inu(SHIB)$0.000006-6.03%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.064060-5.52%
  • tether-goldTether Gold(XAUT)$4,160.90-2.77%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • EthenaEthena(ENA)$0.265469-2.23%
  • MemeCoreMemeCore(M)$1.15-6.29%
  • OndoOndo(ONDO)$0.52-3.40%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$116.74-4.33%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.11%
  • aaveAave(AAVE)$147.59-5.83%
  • Pump.funPump.fun(PUMP)$0.0048809.45%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Fireworks AI Releases Ember-1: A Post-Trained Kimi K3 That Uses About 40% Fewer Tokens

September 28, 2026
in AI & Technology
Reading Time: 15 mins read
A A
Fireworks AI Releases Ember-1: A Post-Trained Kimi K3 That Uses About 40% Fewer Tokens
ShareShareShareShareShare

YOU MAY ALSO LIKE

You Can Now Preorder The Tiny Boox Picco Ereader

20 Agentic Use Cases of TypeSafe AI’s Jev

Fireworks AI has released Ember-1, a specialized model from Fireworks Research built by post-training Moonshot AI’s open-weight Kimi K3. Ember-1 learns to produce shorter reasoning traces while keeping task accuracy. This is different from lowering the reasoning effort setting at inference time. According to the Fireworks release post, Ember-1 delivers Kimi K3’s quality with about 40% fewer tokens.

Is it deployable? Yes, but only through the Fireworks serverless API as a Research Preview. Fireworks has not released Ember-1’s weights, training code, or exact training algorithms, so self-hosting is not an option today.

The Problem: Reasoning Models Think Too Much

Fireworks team reports that reasoning models like Kimi K3 sometimes spend more than 90% of generated tokens on internal reasoning. That cost compounds in multi-turn agentic workloads. Each turn replays prior reasoning back to the model. Context grows roughly quadratically with the number of turns. Long traces from early turns get re-read, and re-billed, on every later call.

Fireworks team explains how customers wanted K3’s coding capability at lower cost. Turning down K3’s reasoning effort did not solve it. Lower effort settings gave up too much quality. So the team trained the model to reason more efficiently instead.

How Fireworks Research Built Ember-1

Not all of K3’s reasoning is waste. Some of it is useful self-reflection, like revisiting an assumption or reacting to feedback. Ember-1 keep that behavior while cutting redundant reasoning and unproductive loops.

The training collection spans mathematics, coding, instruction following, conversation, search, tool use, and software engineering. It covers both standalone problems and extended multi-step interactions. Task and environment feedback guides on-policy planning and learning. Fireworks team ran more than 50 training experiments and over 200 evaluations. They also developed new training algorithms, which it has not published. All training ran on Fireworks Serverless Training. Fireworks states it used its own data and no customer data.

Benchmark Results

Fireworks compared Ember-1 with Kimi K3 at three reasoning effort levels. Cost was computed with public Kimi K3 API pricing. These are Fireworks’ own published evaluations.

Benchmark N K3 Low K3 High K3 Max Ember-1 Ember-1 vs K3 Max (cost)
Terminal Bench 2.1 89 76.4% 77.6% 80.9% 82.0% -51.9% / -23.1 USD
SWE-bench Verified 500 80.4% 86.0% 93.2% 92.2% -15.5% / -68.1 USD
SWE-Interact 75 6.7% 13.3% 21.3% 20.0% -32.5% / -60.8 USD
DeepSWE 1.1 113 55.8% 62.8% 66.4% 75.2% -23.7% / -126.9 USD
τ-2 Bench Airline 50 64% 64% 64% 66% -5.9% / -0.3 USD

Ember-1 leads K3 Max on Terminal Bench 2.1 and DeepSWE 1.1. It trails slightly on SWE-bench Verified and SWE-Interact. Across seven benchmarks and two customers’ production traffic, Fireworks says K3’s reasoning was shortened by 35 to 50% without sacrificing accuracy.

On Doximity’s Bedside Bench, a physician-validated set of 500 clinical cases, Ember-1 set a new cost-per-task Pareto frontier. That result comes from Fireworks’ new Specialized Intelligence Index.

Production A/B Test Results

Fireworks ran live A/B tests with 2 customers on production coding workloads. Both saw roughly 35% fewer tokens per task at comparable quality. In the published run, output tokens fell from 49.3K to 29.9K per task. Reasoning tokens dropped 71.3% and total tokens dropped 39%. The task score was essentially unchanged: 0.753 for Ember-1 versus 0.751 for K3. Average steps fell from 23.8 to 21.4. One customer now runs Ember-1 in production.

Ember-1 costs the same per token as Kimi K3 on Fireworks: $3.00 input, $0.30 cached input, and $15.00 output per 1M tokens. The savings come entirely from generating fewer tokens. At that output rate, the A/B figures work out to about $0.74 versus $0.45 in output cost per task (our calculation, output only).

Interactive Explainer

Credit: Source link

ShareTweetSendSharePin

Related Posts

You Can Now Preorder The Tiny Boox Picco Ereader
AI & Technology

You Can Now Preorder The Tiny Boox Picco Ereader

September 28, 2026
20 Agentic Use Cases of TypeSafe AI’s Jev
AI & Technology

20 Agentic Use Cases of TypeSafe AI’s Jev

September 28, 2026
Google Research Introduces an AI Video Co-Director: 4 Agentic Frameworks for Coherent, Minutes-Long Video Generation
AI & Technology

Google Research Introduces an AI Video Co-Director: 4 Agentic Frameworks for Coherent, Minutes-Long Video Generation

September 28, 2026
Which Is Better To Use?
AI & Technology

Which Is Better To Use?

September 28, 2026
Next Post
Micron: Remains A Strong Buy Ahead Of Q4 Earnings (NASDAQ:MU)

Micron: Remains A Strong Buy Ahead Of Q4 Earnings (NASDAQ:MU)

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Exclusive | Trump Rejects Iran Ceasefire, Expects Renewed Bombing After Midterms – WSJ

Exclusive | Trump Rejects Iran Ceasefire, Expects Renewed Bombing After Midterms – WSJ

September 26, 2026
Which Is Better To Use?

Which Is Better To Use?

September 28, 2026
Ponce Financial Group: Potential Preferred Stock Repurchase Makes It Look More Interesting

Ponce Financial Group: Potential Preferred Stock Repurchase Makes It Look More Interesting

September 22, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!