• bitcoinBitcoin(BTC)$81,726.001.03%
  • ethereumEthereum(ETH)$2,650.582.09%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$767.791.27%
  • rippleXRP(XRP)$1.444.12%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$112.160.17%
  • tronTRON(TRX)$0.3392950.21%
  • zcashZcash(ZEC)$1,493.813.26%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.45%
  • HyperliquidHyperliquid(HYPE)$93.142.19%
  • dogecoinDogecoin(DOGE)$0.0908473.70%
  • moneroMonero(XMR)$567.73-0.19%
  • RainRain(RAIN)$0.0139894.16%
  • whitebitWhiteBIT Coin(WBT)$83.390.49%
  • USDSUSDS(USDS)$1.00-0.03%
  • chainlinkChainlink(LINK)$12.623.99%
  • cardanoCardano(ADA)$0.2305035.37%
  • leo-tokenLEO Token(LEO)$8.930.72%
  • stellarStellar(XLM)$0.2000974.05%
  • uniswapUniswap(UNI)$8.81-0.15%
  • bitcoin-cashBitcoin Cash(BCH)$256.323.01%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • nearNEAR Protocol(NEAR)$3.62-2.56%
  • daiDai(DAI)$1.00-0.01%
  • litecoinLitecoin(LTC)$57.953.40%
  • CantonCanton(CC)$0.1127303.43%
  • USD1USD1(USD1)$1.000.01%
  • avalanche-2Avalanche(AVAX)$9.7620.90%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.391.45%
  • hedera-hashgraphHedera(HBAR)$0.0820794.85%
  • suiSui(SUI)$0.878.26%
  • MemeCoreMemeCore(M)$1.4911.71%
  • shiba-inuShiba Inu(SHIB)$0.0000064.02%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • BittensorBittensor(TAO)$268.708.20%
  • crypto-com-chainCronos(CRO)$0.0602612.10%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • tether-goldTether Gold(XAUT)$4,374.47-0.23%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • okbOKB(OKB)$119.433.45%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.03%
  • aaveAave(AAVE)$143.413.85%
  • OndoOndo(ONDO)$0.4371649.92%
  • mantleMantle(MNT)$0.634.41%
  • AsterAster(ASTER)$0.772.50%
  • EthenaEthena(ENA)$0.20356924.04%
  • Pump.funPump.fun(PUMP)$0.004188-2.09%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

MLPerf 4.0 training results show up to 80% in AI performance gains

June 12, 2024
in AI & Technology
Reading Time: 4 mins read
A A
MLPerf 4.0 training results show up to 80% in AI performance gains
ShareShareShareShareShare

It’s time to celebrate the incredible women leading the way in AI! Nominate your inspiring leaders for VentureBeat’s Women in AI Awards today before June 18. Learn More


Innovation in machine learning and AI training continues to accelerate, even as more complex generative AI workloads come online.

YOU MAY ALSO LIKE

How To Block And Unblock A Number On Your Android Phone

Google Gemini Also Escaped Its Testing Environment And Hacked Three Companies

Today MLCommons released the MLPerf 4.0 training benchmark, once again showing record levels of performance. The MLPerf training benchmark is a vendor neutral standard that enjoys broad industry participation. The MLPerf Training suite measures performance of full AI training systems across a range of workloads. Version 4.0 included over 205 results from 17 organizations. The new update is the first MLPerf training results release since MLPerf 3.1 training in November 2023.

The MLPerf 4.0 training benchmarks include results for image generation with Stable Diffusion and Large Language Model (LLM) training for GPT-3. With the MLPerf 4.0 training benchmarks are a number of first time results including a new LoRA benchmark that fine-tunes the Llama 2 70B language model on document summarization using a parameter-efficient approach.

As is often the case with MLPerf results, when comparing even to just six months ago, there is significant gain.


VB Transform 2024 Registration is Open

Join enterprise leaders in San Francisco from July 9 to 11 for our flagship AI event. Connect with peers, explore the opportunities and challenges of Generative AI, and learn how to integrate AI applications into your industry. Register Now


“Even if you look at relative to the last cycle, some of our benchmarks have gotten nearly 2x better performance, in particular Stable Diffusion,” MLCommons founder and executive director David Kanter said in a press briefing. “So that’s pretty impressive in six months.”

The actual gain for Stable Diffusion training is 1.8x faster vs November 2023, while training for GPT-3 was up to 1.2x faster.

AI training performance isn’t just about hardware

There are many factors that go into training an AI model.

While hardware is important, so too is software as well as the network that connects clusters together.

“Particularly for AI training, we have access to many different lead levers to help improve performance and efficiency,” Kanter said. “For training, most of these systems are using multiple processors or accelerators and how the work is divided and communicated is absolutely critical.”

Kanter added that not only are vendors taking advantage of better silicon, they are also using better algorithms and better scaling to provide more performance over time.

Nvidia continues to scale training on Hopper

The big results in the MLPerf 4.0 training benchmarks all largely belong to Nvidia.

Across nine different tested workloads, Nvidia claims to have set new performance records on five of them. Perhaps most impressively is that the new records were mostly set using the same core hardware platforms Nvidia used a year ago in June 2023.

In a press briefing David Salvator, director of AI at Nvidia, commented that the Nvidia H100 Hopper architecture continues to deliver value.

“Throughout Nvidia’s history with deep learning in any given generation of product we will typically get two to 2.5x more performance out of an architecture, from software innovation over the course of the life of that particular product,” Salvator said.

For the H100, Nvidia used numerous techniques to improve performance for MLPerf 4.0 training. The various techniques include full stack optimization, highly tuned FP8 kernels, FP8-aware distributed optimizer, optimized cuDNN FlashAttention, improved math and comms execution overlap as well as intelligent GPU power allocation.

Why the MLPerf training benchmarks matter to the enterprise

Aside from providing organizations with standardized benchmarks on training performance, there is more value that the actual numbers provide.

While performance keeps on getting better all the time, Salvator emphasized that it’s getting better also with the same hardware.

Salvator noted that the results are a quantitative demonstration that shows how Nvidia is able to deliver new value on top of existing architectures. As organizations are considering building out new deployments, particularly on-premises, he said they are essentially make a big bet on a technology platform. The fact that an organization can get growing benefits for years after an initial technology debut is important.

“In terms of why we care so much about performance, the simple answer is because for businesses, it drives return on investment,” he said.

VB Daily

Stay in the know! Get the latest news in your inbox daily

By subscribing, you agree to VentureBeat’s Terms of Service.

Thanks for subscribing. Check out more VB newsletters here.

An error occured.

Credit: Source link
ShareTweetSendSharePin

Related Posts

How To Block And Unblock A Number On Your Android Phone
AI & Technology

How To Block And Unblock A Number On Your Android Phone

September 19, 2026
Google Gemini Also Escaped Its Testing Environment And Hacked Three Companies
AI & Technology

Google Gemini Also Escaped Its Testing Environment And Hacked Three Companies

September 19, 2026
What Is AI Agent Memory? Short-Term, Long-Term, Episodic, and Semantic Memory Explained – Unite.AI
AI & Technology

What Is AI Agent Memory? Short-Term, Long-Term, Episodic, and Semantic Memory Explained – Unite.AI

September 19, 2026
Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model
AI & Technology

Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model

September 19, 2026
Next Post
Should Elon Musk be paid  million to run #Tesla? #technology #shorts

Should Elon Musk be paid $55 million to run #Tesla? #technology #shorts

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
I Blow Through Our Monthly Budget In A Week

I Blow Through Our Monthly Budget In A Week

September 12, 2026
Sununu thanks supporters after projected GOP Senate primary win in New Hampshire

Sununu thanks supporters after projected GOP Senate primary win in New Hampshire

September 15, 2026
Dire warning that A.I. could end humanity

Dire warning that A.I. could end humanity

September 14, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!