• bitcoinBitcoin(BTC)$84,016.00-0.44%
  • ethereumEthereum(ETH)$2,691.890.05%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$777.060.00%
  • rippleXRP(XRP)$1.571.91%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$121.984.41%
  • tronTRON(TRX)$0.338145-0.62%
  • zcashZcash(ZEC)$1,555.040.89%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.45%
  • HyperliquidHyperliquid(HYPE)$92.700.63%
  • dogecoinDogecoin(DOGE)$0.0990703.44%
  • moneroMonero(XMR)$558.67-1.77%
  • chainlinkChainlink(LINK)$13.895.01%
  • whitebitWhiteBIT Coin(WBT)$83.89-0.49%
  • USDSUSDS(USDS)$1.00-0.01%
  • cardanoCardano(ADA)$0.2589754.05%
  • RainRain(RAIN)$0.011585-3.83%
  • leo-tokenLEO Token(LEO)$8.83-1.35%
  • stellarStellar(XLM)$0.2201492.03%
  • bitcoin-cashBitcoin Cash(BCH)$343.562.26%
  • nearNEAR Protocol(NEAR)$4.968.47%
  • uniswapUniswap(UNI)$9.624.86%
  • litecoinLitecoin(LTC)$72.380.56%
  • CantonCanton(CC)$0.13007215.06%
  • suiSui(SUI)$1.2119.42%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • avalanche-2Avalanche(AVAX)$10.654.50%
  • daiDai(DAI)$1.000.01%
  • USD1USD1(USD1)$1.000.05%
  • hedera-hashgraphHedera(HBAR)$0.0953882.17%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.453.18%
  • BittensorBittensor(TAO)$317.567.28%
  • BitwayBitway(BTW)$1.3238.59%
  • shiba-inuShiba Inu(SHIB)$0.0000062.99%
  • crypto-com-chainCronos(CRO)$0.0663045.79%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • MemeCoreMemeCore(M)$1.21-0.31%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • EthenaEthena(ENA)$0.26668319.68%
  • OndoOndo(ONDO)$0.555.90%
  • tether-goldTether Gold(XAUT)$4,285.150.43%
  • okbOKB(OKB)$120.841.06%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • Circle USYCCircle USYC(USYC)$1.140.03%
  • aaveAave(AAVE)$153.624.51%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.00%
  • mantleMantle(MNT)$0.67-1.02%
  • polkadotPolkadot(DOT)$1.225.34%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Nvidia debuts Nemotron 3 with hybrid MoE and Mamba-Transformer to drive efficient agentic AI

December 15, 2025
in AI & Technology
Reading Time: 5 mins read
A A
Nvidia debuts Nemotron 3 with hybrid MoE and Mamba-Transformer to drive efficient agentic AI
ShareShareShareShareShare

Nvidia launched the new version of its frontier models, Nemotron 3, by leaning in on a model architecture that the world’s most valuable company said offers more accuracy and reliability for agents. 

Nemotron 3 will be available in three sizes: Nemotron 3 Nano with 30B parameters, mainly for targeted, highly efficient tasks; Nemotron 3 Super, which is a 100B parameter model for multi-agent applications and with high-accuracy reasoning and Nemotron 3 Ultra, with its large reasoning engine and around 500B parameters for more complex applications. 

YOU MAY ALSO LIKE

How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data

New Mexico Jury Rules Meta Misled State Residents About Data Privacy

To build the Nemotron 3 models, Nvidia said it leaned into a hybrid mixture-of-experts (MoE) architecture to improve scalability and efficiency. By using this architecture, Nvidia said in a press release that its new models also offer enterprises more openness and performance when building multi-agent autonomous systems. 

Kari Briski, Nvidia vice president for generative AI software, told reporters in a briefing that the company wanted to demonstrate its commitment to learn and improving from previous iterations of its models. 

“We believe that we are uniquely positioned to serve a wide range of developers who want full flexibility to customize models for building specialized AI by combining that new hybrid mixture of our mixture of experts architecture with a 1 million token context length,” Briski said.  

Nvidia said early adopters of the Nemotron 3 models include Accenture, CrowdStrike, Cursor, Deloitte, EY, Oracle Cloud Infrastructure, Palantir, Perplexity, ServiceNow, Siemens and Zoom.

Breakthrough architectures 

Nvidia has been using the hybrid Mamba-Transformer mixture-of-experts architecture for many of its models, including Nemotron-Nano-9B-v2.

The architecture is based on research from Carnegie Mellon University and Princeton, which weaves in selective state-space models to handle long pieces of information while maintaining states. It can reduce compute costs even through long contexts. 

Nvidia noted its design “achieves up to 4x higher token throughput” compared to Nemotron 2 Nano and can significantly lower inference costs by reducing reasoning token generation by up 60%.

“We really need to be able to bring that efficiency up and the cost per token down. And you can do it through a number of ways, but we’re really doing it through the innovations of that model architecture,” Briski said. “The hybrid Mamba transformer architecture runs several times faster with less memory, because it avoids these huge attention maps and key value caches for every single token.”

Nvidia also introduced an additional innovation for the Nemotron 3 Super and Ultra models. For these, Briski said Nvidia deployed “a breakthrough called latent MoE.”

“That’s all these experts that are in your model share a common core and keep only a small part private. It’s kind of like chefs sharing one big kitchen, but they need to get their own spice rack,” Briski added. 

Nvidia is not the only company that employs this kind of architecture to build models. AI21 Labs uses it for its Jamba models, most recently in its Jamba Reasoning 3B model.

The Nemotron 3 models benefited from extended reinforcement learning. The larger models, Super and Ultra, used the company’s 4-bit NVFP4 training format, which allows them to train on existing infrastructure without compromising accuracy.

Benchmark testing from Artificial Analysis placed the Nemotron models highly among models of similar size. 

New environments for models to ‘work out’

As part of the Nemotron 3 launch, Nvidia will also give users access to its research by releasing its papers and sample prompts, offering open datasets where people can use and look at pre-training tokens and post-training samples, and most importantly, a new NeMo Gym where customers can let their models and agents “workout.” 

The NeMo Gym is a reinforcement learning lab where users can let their models run in simulated environments to test their post-training performance. 

AWS announced a similar tool through its Nova Forge platform, targeted for enterprises that want to test out their newly created distilled or smaller models.  

Briski said the samples of post-training data Nvidia plans to release “are orders of magnitude larger than any available post-training data set and are also very permissive and open.”

Nvidia pointed to developers seeking highly intelligent and performant open models, so they can better understand how to guide them if needed, as the basis for releasing more information about how it trains its models. 

“Model developers today hit this tough trifecta. They need to find models that are ultra open, that are extremely intelligent and are highly efficient,” she said. “Most open models force developers into painful trade-offs between efficiencies like token costs, latency, and throughput.”

She said developers want to know how a model was trained, where the training data came from and how they can evaluate it.

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data
AI & Technology

How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data

September 25, 2026
New Mexico Jury Rules Meta Misled State Residents About Data Privacy
AI & Technology

New Mexico Jury Rules Meta Misled State Residents About Data Privacy

September 25, 2026
Apple’s HomePod Mini 2 Will Reportedly Come In New Colors, But Feature A Similar Design
AI & Technology

Apple’s HomePod Mini 2 Will Reportedly Come In New Colors, But Feature A Similar Design

September 25, 2026
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB
AI & Technology

Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB

September 25, 2026
Next Post
Broadcom: What So Many Analysts & Investors Got Wrong – Buy The Dip (NASDAQ:AVGO)

Broadcom: What So Many Analysts & Investors Got Wrong - Buy The Dip (NASDAQ:AVGO)

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Now Trump Says He’s Creating An AI Force

Now Trump Says He’s Creating An AI Force

September 19, 2026
Anthropic’s Claude Takes Bigger Role in Building AI

Anthropic’s Claude Takes Bigger Role in Building AI

September 20, 2026
1985: Tim Curry discusses playing the butler in ‘Clue’

1985: Tim Curry discusses playing the butler in ‘Clue’

September 23, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!