• bitcoinBitcoin(BTC)$75,520.00-4.12%
  • ethereumEthereum(ETH)$2,395.24-5.96%
  • tetherTether(USDT)$1.00-0.04%
  • binancecoinBNB(BNB)$712.04-1.58%
  • rippleXRP(XRP)$1.28-11.25%
  • usd-coinUSDC(USDC)$1.00-0.02%
  • solanaSolana(SOL)$96.72-6.29%
  • tronTRON(TRX)$0.332433-1.97%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.01-2.76%
  • zcashZcash(ZEC)$1,110.83-5.71%
  • HyperliquidHyperliquid(HYPE)$76.61-4.99%
  • dogecoinDogecoin(DOGE)$0.079789-5.48%
  • RainRain(RAIN)$0.014077-1.79%
  • USDSUSDS(USDS)$1.00-0.03%
  • moneroMonero(XMR)$500.85-2.69%
  • whitebitWhiteBIT Coin(WBT)$77.68-4.80%
  • leo-tokenLEO Token(LEO)$8.88-1.32%
  • chainlinkChainlink(LINK)$10.90-6.34%
  • cardanoCardano(ADA)$0.194708-7.60%
  • stellarStellar(XLM)$0.175335-9.28%
  • Ethena USDeEthena USDe(USDE)$1.00-0.07%
  • daiDai(DAI)$1.000.01%
  • USD1USD1(USD1)$1.00-0.03%
  • bitcoin-cashBitcoin Cash(BCH)$215.08-4.76%
  • litecoinLitecoin(LTC)$51.05-4.28%
  • uniswapUniswap(UNI)$6.24-4.91%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.32-2.66%
  • CantonCanton(CC)$0.090356-7.80%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.074654-4.31%
  • avalanche-2Avalanche(AVAX)$7.26-4.94%
  • nearNEAR Protocol(NEAR)$2.31-7.66%
  • shiba-inuShiba Inu(SHIB)$0.000005-7.23%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.04%
  • suiSui(SUI)$0.68-6.91%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,291.88-0.23%
  • crypto-com-chainCronos(CRO)$0.054996-7.74%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • MemeCoreMemeCore(M)$1.121.90%
  • BittensorBittensor(TAO)$217.47-7.33%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$109.68-3.42%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.03%
  • BitwayBitway(BTW)$0.7017.29%
  • aaveAave(AAVE)$121.24-6.32%
  • pax-goldPAX Gold(PAXG)$4,293.96-0.27%
  • AsterAster(ASTER)$0.68-4.01%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056904-1.29%
  • mantleMantle(MNT)$0.54-6.05%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

NVIDIA AI Researchers Introduce ScaleFold: A Leap in High-Performance Computing for Protein Structure Prediction

April 22, 2024
in AI & Technology
Reading Time: 5 mins read
A A
NVIDIA AI Researchers Introduce ScaleFold: A Leap in High-Performance Computing for Protein Structure Prediction
ShareShareShareShareShare

In recent years, deep learning has been effective in high-performance computing. Surrogate models continue to advance, surpassing physics-based simulations in accuracy and utility. This AI-driven progress is evident in protein folding, exemplified by RoseTTAFold, AlphaFold2, OpenFold, and FastFold, democratizing protein structure-based drug discovery. AlphaFold, a breakthrough by DeepMind, achieved accuracy comparable to experimental methods, addressing a longstanding biological challenge. Despite its success, AlphaFold’s training process, relying on a sequence attention mechanism, is time and resource-intensive, hindering research speed. Efforts to enhance training efficiency, such as those by OpenFold, DeepSpeed4Science, and FastFold, target scalability, a pivotal challenge in accelerating AlphaFold’s training.

Incorporating AlphaFold training into the MLPerf HPC v3.0 benchmark underscores its importance. However, this training process presents significant challenges. Firstly, despite its relatively modest parameter count, the AlphaFold model demands extensive memory due to Evoformer’s unique attention mechanism, which scales cubically with input size. OpenFold addressed this with gradient checkpointing but at the expense of training speed. Additionally, AlphaFold’s training involves a multitude of memory-bound kernels, dominating computation time. Also, critical operations like Multi-Head Attention and Layer Normalization consume substantial time, limiting the effectiveness of data parallelism.

NVIDIA researchers provide a study that thoroughly analyzes AlphaFold’s training, identifying key impediments to scalability: inefficient distributed communication and underutilization of compute resources. The researchers propose several optimizations. They introduce a non-blocking data pipeline to alleviate slow-worker issues and employ fine-grained optimizations, such as utilizing CUDA Graphs to reduce overhead. Also, they design specialized Triton kernels for critical computation patterns, fuse fragmented computations, and optimize kernel configurations. This optimized training method, named ScaleFold, aims to enhance overall efficiency and scalability.

The researchers extensively examine AlphaFold’s training, identifying barriers to scalability across communication and computation. Solutions include a non-blocking data pipeline and CUDA Graphs to mitigate communication imbalances, while manual and automatic kernel fusions enhance computation efficiency. They addressed issues like CPU overhead and imbalanced communication, proposing optimizations such as Triton kernels and low-precision support. Asynchronous evaluation and caching alleviate evaluation time bottlenecks. AlphaFold training is scaled to 2080 NVIDIA H100 GPUs through systematic optimizations, completing pretraining in 10 hours.

Comparing ScaleFold to public OpenFold and FastFold reveals superior training performance. On A100, ScaleFold achieves step times of 1.88s for DAP-2, outperforming FastFold’s 2.49s and far surpassing OpenFold’s 6.19s without DAP support. On H100, ScaleFold exhibits step times of 0.65s for DAP-8, significantly faster than OpenFold’s 1.80s with NoDAP. Comprehensive evaluations demonstrate ScaleFold’s advancements, achieving up to 6.2X speedup compared to reference models on NVIDIA H100. MLPerf HPC 3.0 benchmarking on Eos further confirms ScaleFold’s efficiency, reducing training time to 8 minutes on 2080 NVIDIA H100 GPUs, a 6X improvement over reference models. Training from scratch on ScaleFold completes AlphaFold pretraining in under 10 hours, showcasing its accelerated performance.

The key contributions of this research are three-fold:

  1. Researchers identified the key factors that prevented the AlphaFold training from scaling to more compute resources.
  2. They introduced ScaleFold, a scalable and systematic training method for the AlphaFold model.
  3. They empirically demonstrated the scalability of ScaleFold and set new records for the AlphaFold pretraining and the MLPef HPC benchmark.

To conclude, This research addressed the scalability challenges of AlphaFold training by introducing ScaleFold, a systematic approach tailored to mitigate inefficient communications and overhead-dominated computations. ScaleFold incorporates several optimizations, including FastFold’s DAP for GPU scaling, a Non-Blocking Data Pipeline to address batch access inequalities, and a CUDA Graph to eliminate CPU overhead. Also, efficient Triton kernels for critical patterns, automatic fusion via torch compiler, and bfloat16 training are implemented to reduce step time. Asynchronous evaluation and an evaluation dataset cache alleviate evaluation time bottlenecks. These optimizations enable ScaleFold to achieve a training convergence time of 7.51 minutes on 2080 NVIDIA H100 GPUs in MLPerf HPC V3.0, demonstrating a 6X speedup over reference models. Training from scratch is reduced from 7 days to 10 hours, setting a new record in efficiency compared to prior works.


Check out the Paper. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. Join our Telegram Channel, Discord Channel, and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 40k+ ML SubReddit


YOU MAY ALSO LIKE

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

How To Improve The Audio Quality On Your iPhone

Asjad is an intern consultant at Marktechpost. He is persuing B.Tech in mechanical engineering at the Indian Institute of Technology, Kharagpur. Asjad is a Machine learning and deep learning enthusiast who is always researching the applications of machine learning in healthcare.


🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
AI & Technology

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

September 15, 2026
How To Improve The Audio Quality On Your iPhone
AI & Technology

How To Improve The Audio Quality On Your iPhone

September 15, 2026
Ferrovalle Taps INFORM for AI Smart Yard at Mexico City Rail Hub – Unite.AI
AI & Technology

Ferrovalle Taps INFORM for AI Smart Yard at Mexico City Rail Hub – Unite.AI

September 15, 2026
Google’s Latest Pixel Drop Will Keep You More Connected To Your VIPs
AI & Technology

Google’s Latest Pixel Drop Will Keep You More Connected To Your VIPs

September 15, 2026
Next Post
Chuck

Chuck

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Options Trading for Beginners 2026 (The Complete Guide)

Options Trading for Beginners 2026 (The Complete Guide)

September 14, 2026
Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster

Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster

September 10, 2026
Tech company discloses first-of-its-kind A.I. cyberattack against a government

Tech company discloses first-of-its-kind A.I. cyberattack against a government

September 15, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!