• bitcoinBitcoin(BTC)$75,407.00-2.01%
  • ethereumEthereum(ETH)$2,376.40-2.59%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$711.35-1.67%
  • rippleXRP(XRP)$1.26-10.82%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$96.33-4.03%
  • tronTRON(TRX)$0.3353780.22%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-4.27%
  • zcashZcash(ZEC)$1,242.778.85%
  • HyperliquidHyperliquid(HYPE)$78.080.39%
  • dogecoinDogecoin(DOGE)$0.078541-4.51%
  • USDSUSDS(USDS)$1.00-0.02%
  • RainRain(RAIN)$0.013095-7.75%
  • moneroMonero(XMR)$485.75-4.40%
  • whitebitWhiteBIT Coin(WBT)$77.44-2.33%
  • leo-tokenLEO Token(LEO)$8.85-0.30%
  • chainlinkChainlink(LINK)$10.64-6.26%
  • cardanoCardano(ADA)$0.191038-6.17%
  • stellarStellar(XLM)$0.172707-10.40%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • daiDai(DAI)$1.000.01%
  • USD1USD1(USD1)$1.00-0.03%
  • bitcoin-cashBitcoin Cash(BCH)$214.41-4.00%
  • litecoinLitecoin(LTC)$50.34-3.44%
  • uniswapUniswap(UNI)$6.10-6.50%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.30-2.34%
  • CantonCanton(CC)$0.090227-3.38%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • nearNEAR Protocol(NEAR)$2.451.24%
  • avalanche-2Avalanche(AVAX)$7.19-4.38%
  • hedera-hashgraphHedera(HBAR)$0.072168-7.77%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • shiba-inuShiba Inu(SHIB)$0.000005-7.40%
  • suiSui(SUI)$0.68-4.16%
  • crypto-com-chainCronos(CRO)$0.054933-4.01%
  • tether-goldTether Gold(XAUT)$4,338.641.04%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • MemeCoreMemeCore(M)$1.12-0.13%
  • BittensorBittensor(TAO)$214.91-4.94%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$108.86-2.15%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.15%
  • BitwayBitway(BTW)$0.757.84%
  • pax-goldPAX Gold(PAXG)$4,344.091.12%
  • AsterAster(ASTER)$0.68-2.53%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.057028-0.17%
  • mantleMantle(MNT)$0.54-1.98%
  • aaveAave(AAVE)$113.90-10.19%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

NVIDIA Vera Rubin NVL72 Posts First MLPerf Inference Preview Results – Unite.AI

September 16, 2026
in AI & Technology
Reading Time: 3 mins read
A A
NVIDIA Vera Rubin NVL72 Posts First MLPerf Inference Preview Results – Unite.AI
ShareShareShareShareShare

NVIDIA on September 16, 2026, published its MLPerf Inference v6.1 submission results, headlined by the first MLPerf Inference preview submission of its Vera Rubin NVL72 system, which the company said delivered up to 3.7x higher throughput than its GB300 NVL72 system on the Qwen3-VL benchmark.

YOU MAY ALSO LIKE

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down

MLPerf Inference: Datacenter is an MLCommons Association benchmark suite that measures how fast systems process inputs and produce results using a trained model. Under MLCommons’ published definitions, the Closed division — the category NVIDIA cited for its headline figures — requires using the same model as the reference implementation so that hardware platforms and software frameworks can be compared “apples-to-apples,” while systems in the Preview availability category must be submittable as Available in the next submission round.

DeepSeek-R1 and Qwen3-VL Preview Results

NVIDIA submitted Vera Rubin NVL72 preview results on DeepSeek-R1 and Qwen3-VL, which it described as two of the most demanding benchmarks in the v6.1 suite. On Qwen3-VL, the company reported up to 3.7x higher throughput than GB300 NVL72 across the offline, server and interactive scenarios, using vLLM with the NVIDIA Dynamo open source inference framework. On DeepSeek-R1, using the NVIDIA TensorRT-LLM library, NVIDIA reported throughput up to 2.5x higher than GB300 NVL72.

The cited Closed division figures were retrieved from mlcommons.org on September 16, 2026, and are drawn from submission entries 6.1-0106 and 6.1-0074, according to the citation published with the results. NVIDIA said the performance means each Vera Rubin NVL72 rack delivers significantly more tokens, serves more users and generates more revenue than a GB300 NVL72 rack while lowering cost per token.

Nebius also submitted Vera Rubin NVL72 preview results in the same round; NVIDIA described those results as demonstrating excellent performance.

Hardware-Software Codesign and Agentic Testing

NVIDIA credited full-stack codesign across hardware and software for the results. The company said Vera Rubin’s enhanced Tensor Cores and Transformer Engine accelerate both the prefill and decode stages of inference, while NVFP4 precision reduces the memory footprint of model weights, attention and KV cache, raising throughput with what NVIDIA described as minimal loss of output quality.

The submissions used disaggregated serving, which separates prefill and decode, together with large-scale expert parallelism across the mixture-of-experts layers used by models such as DeepSeek-R1 and Qwen3-VL. NVIDIA said the NVL72 scale-up domain, built on sixth-generation NVLink and NVLink Switch, delivers 10x higher packet rates and 3x lower latency than off-the-shelf Ethernet, providing the interconnect foundation for those techniques at rack scale.

The company also reported that Vera Rubin NVL72 delivered 30x better performance than GB300 NVL72 in preview testing on the SemiAnalysis AgentX benchmark, which is designed to measure agentic workloads. NVIDIA said the upcoming MLPerf Endpoints benchmark will bring standardized measurement to agentic inference workloads beyond what traditional throughput benchmarks capture.

GB300 NVL72 Scaling, Software Gains and Partner Results

In the same round, NVIDIA’s DeepSeek-R1 submission scaled from a single GB300 NVL72 rack of 72 GPUs to four racks totaling 288 GPUs, achieving 99% scaling efficiency in the offline scenario, with throughput growing nearly in proportion to the hardware added; the company cited entries 6.1-0073 and 6.1-0074. On the WAN 2.2 text-to-video benchmark, GB300 NVL72 reached 0.65 720p videos per second at 5.7 seconds per video, which NVIDIA said represents 9x higher throughput and 7.5x lower latency than a single node.

NVIDIA reported that GB300 NVL72 performance on Qwen3-VL improved up to 1.6x in v6.1 over its v6.0 results, crediting lower KV cache precision, additional kernel fusion, better kernels and disaggregated serving with vLLM and NVIDIA Dynamo. The company said optimization continued past the v6.1 submission deadline, and that post-submission results on GPT-OSS-120B and DLRMv3 show further gains, though those figures have not yet been verified by MLCommons.

Beyond its rack-scale platforms, NVIDIA submitted Jetson AGX Thor results using the TensorRT Edge-LLM library on the newly introduced Edge-Agentic benchmark with the Qwen3.6-27B model. The company said 19 partners participated in the round, eight of them submitting on multi-node Blackwell NVL72 systems: ASUS, Azure, Cisco, CoreWeave, Crusoe, Dell Technologies, Fujitsu, Giga Computing, HPE, Inventec, Lambda, MiTAC Computing, Nebius, Oracle Cloud Infrastructure, Quanta Cloud Technology, Red Hat, ScitiX, Supermicro and Wiwynn.

MLCommons notes on its benchmark page that published results are sometimes modified or invalidated after initial publication, with any changes recorded in its change log.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI
AI & Technology

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

September 16, 2026
MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down
AI & Technology

MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down

September 16, 2026
Samsung Brings One UI 9 To The Rest Of The Galaxy S26 Series
AI & Technology

Samsung Brings One UI 9 To The Rest Of The Galaxy S26 Series

September 16, 2026
NVIDIA, Google and Emerald AI Form AI Energy Management Alliance – Unite.AI
AI & Technology

NVIDIA, Google and Emerald AI Form AI Energy Management Alliance – Unite.AI

September 16, 2026
Next Post
Powerful waves from Hurricane Marie slam Southern California

Powerful waves from Hurricane Marie slam Southern California

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Q3 Estimated Tax Is Due September 15

Q3 Estimated Tax Is Due September 15

September 14, 2026
Roblox Pushes Deeper Into AI-Powered Gaming

Roblox Pushes Deeper Into AI-Powered Gaming

September 16, 2026
Philippine coast guard rescues passengers after ferry fire

Philippine coast guard rescues passengers after ferry fire

September 14, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!