• bitcoinBitcoin(BTC)$84,031.00-0.91%
  • ethereumEthereum(ETH)$2,691.10-0.19%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$775.48-0.50%
  • rippleXRP(XRP)$1.581.50%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$122.183.44%
  • tronTRON(TRX)$0.337497-0.75%
  • zcashZcash(ZEC)$1,546.03-0.56%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-0.04%
  • HyperliquidHyperliquid(HYPE)$92.300.53%
  • dogecoinDogecoin(DOGE)$0.0991082.65%
  • moneroMonero(XMR)$556.02-2.54%
  • chainlinkChainlink(LINK)$13.963.69%
  • whitebitWhiteBIT Coin(WBT)$83.90-0.68%
  • cardanoCardano(ADA)$0.2628254.43%
  • USDSUSDS(USDS)$1.00-0.01%
  • leo-tokenLEO Token(LEO)$8.90-0.22%
  • stellarStellar(XLM)$0.221471-0.65%
  • RainRain(RAIN)$0.010728-11.02%
  • bitcoin-cashBitcoin Cash(BCH)$342.830.81%
  • nearNEAR Protocol(NEAR)$4.896.13%
  • uniswapUniswap(UNI)$9.543.65%
  • litecoinLitecoin(LTC)$72.871.72%
  • CantonCanton(CC)$0.13475717.03%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • suiSui(SUI)$1.1814.29%
  • avalanche-2Avalanche(AVAX)$10.885.39%
  • daiDai(DAI)$1.00-0.01%
  • USD1USD1(USD1)$1.000.02%
  • hedera-hashgraphHedera(HBAR)$0.0950561.04%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.452.62%
  • BittensorBittensor(TAO)$315.175.71%
  • shiba-inuShiba Inu(SHIB)$0.0000063.04%
  • crypto-com-chainCronos(CRO)$0.0662201.65%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • MemeCoreMemeCore(M)$1.22-0.66%
  • EthenaEthena(ENA)$0.26556615.30%
  • tether-goldTether Gold(XAUT)$4,284.07-0.12%
  • OndoOndo(ONDO)$0.541.83%
  • okbOKB(OKB)$121.571.16%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BitwayBitway(BTW)$0.86-10.48%
  • Circle USYCCircle USYC(USYC)$1.140.03%
  • aaveAave(AAVE)$154.815.43%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.07%
  • mantleMantle(MNT)$0.67-1.27%
  • polkadotPolkadot(DOT)$1.236.02%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

LinkedIn Released Liger (Linkedin GPU Efficient Runtime) Kernel: A Revolutionary Tool That Boosts LLM Training Efficiency by Over 20% While Cutting Memory Usage by 60%

August 25, 2024
in AI & Technology
Reading Time: 5 mins read
A A
LinkedIn Released Liger (Linkedin GPU Efficient Runtime) Kernel: A Revolutionary Tool That Boosts LLM Training Efficiency by Over 20% While Cutting Memory Usage by 60%
ShareShareShareShareShare

LinkedIn has recently unveiled its groundbreaking innovation, the Liger (LinkedIn GPU Efficient Runtime) Kernel, a collection of highly efficient Triton kernels designed specifically for large language model (LLM) training. This new technology represents an advancement in machine learning, particularly in training large-scale models that require substantial computational resources. The Liger Kernel is poised to become a pivotal tool for researchers, machine learning practitioners, and those eager to optimize their GPU training efficiency.

Introduction to Liger Kernel

YOU MAY ALSO LIKE

Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding

How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data

The Liger Kernel has been meticulously crafted to address the growing demands of LLM training by enhancing both speed and memory efficiency. The development team at LinkedIn has implemented several advanced features in the Liger Kernel, including Hugging Face-compatible RMSNorm, RoPE, SwiGLU, CrossEntropy, FusedLinearCrossEntropy, and more. These kernels are efficient and compatible with widely used tools like Flash Attention, PyTorch FSDP, and Microsoft DeepSpeed, making them highly versatile for various applications.

Key Features and Benefits

One of the most remarkable aspects of the Liger Kernel is its ability to increase multi-GPU training throughput by more than 20% while reducing memory usage by up to 60%. This dual benefit is achieved through kernel fusion, in-place replacement, and chunking techniques that optimize the computational processes involved in LLM training. The kernel is designed to be lightweight, with minimal dependencies, requiring only Torch and Triton, which eliminates the common headaches associated with managing complex software dependencies.

The Liger Kernel’s efficiency is further exemplified by its ability to handle larger context lengths, larger batch sizes, and massive vocabularies without compromising performance. For example, while traditional Hugging Face models may encounter out-of-memory (OOM) errors at 4K, the Liger Kernel can scale up to 16K, substantially boosting model capacity and capability.

Applications and Use Cases

The Liger Kernel is particularly beneficial for those working on large-scale LLM training projects. For instance, when training the LLaMA 3-8B model, the Liger Kernel can achieve up to a 20% increase in training speed and a 40% reduction in memory usage. This is especially useful for training on datasets like Alpaca, where computational efficiency can significantly impact the overall cost and time required for model development.

In more advanced scenarios, such as the retraining phase of a multi-head LLM like Medusa, the Liger Kernel can reduce memory usage by an impressive 80% while improving throughput by 40%. These improvements are crucial for researchers and practitioners aiming to push the boundaries of what is possible with LLMs, enabling them to experiment with larger models and more complex architectures without hardware limitations.

Technical Overview

The Liger Kernel integrates several key Triton-based operations that enhance the performance of LLM training. Among these are RMSNorm, RoPE, SwiGLU, and FusedLinearCrossEntropy, each contributing to the kernel’s overall efficiency. For instance, RMSNorm normalizes activations using their root mean square. This process has been optimized within the Liger Kernel to achieve a threefold increase in speed and peak memory reduction.

Similarly, RoPE (Rotary Positional Embedding) and SwiGLU (Swish Gated Linear Units) have been implemented with in-place replacement techniques that significantly reduce memory usage and increase computational speed. The CrossEntropy loss function, critical for many LLM tasks, has also been optimized to reduce peak memory usage by over four times while doubling the execution speed.

Ease of Use and Installation

Despite its advanced capabilities, the Liger Kernel is designed to be user-friendly & easily integrated into existing workflows. Users can patch their existing Hugging Face models with the optimized Liger Kernels using just one line of code. The kernel’s lightweight design also ensures it is compatible with multi-GPU setups, including PyTorch FSDP and DeepSpeed, without requiring extensive configuration or additional libraries.

The Liger Kernel can be installed via pip, with both stable and nightly versions available. This ease of installation, combined with the kernel’s minimal dependencies, makes it accessible to a wide range of users, from seasoned machine learning practitioners to curious novices looking to enhance their training efficiency.

Future Prospects and Community Involvement

LinkedIn is committed to continually improving the Liger Kernel and welcomes contributions from the community. By fostering collaboration, LinkedIn aims to gather the best kernels for LLM training and incorporate them into future versions of the Liger Kernel. This approach ensures that the kernel remains at the forefront of technological innovation in LLM training.

Conclusion

LinkedIn’s release of the Liger Kernel marks a significant milestone in the evolution of LLM training. The Liger Kernel is set to become an indispensable tool for anyone involved in large-scale model training by offering a highly efficient, easy-to-use, and versatile solution. Its ability to drastically improve both speed and memory efficiency will undoubtedly accelerate the development of more advanced and capable LLMs, paving the way for breakthroughs in artificial intelligence. 


Check out the GitHub. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter and join our Telegram Channel and LinkedIn Group. If you like our work, you will love our newsletter..

Don’t Forget to join our 49k+ ML SubReddit

Find Upcoming AI Webinars here


Asif Razzaq is the CEO of Marktechpost Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.

🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding
AI & Technology

Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding

September 25, 2026
How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data
AI & Technology

How To Stop Meta Training Its AI Models On Your Smart Glasses’ Visual Data

September 25, 2026
New Mexico Jury Rules Meta Misled State Residents About Data Privacy
AI & Technology

New Mexico Jury Rules Meta Misled State Residents About Data Privacy

September 25, 2026
Cricut’s New DIY Machines Let You Print And Cut Your Own Stickers
AI & Technology

Cricut’s New DIY Machines Let You Print And Cut Your Own Stickers

September 25, 2026
Next Post
Venezuelans rally in Miami to protest Maduro’s election claims

Venezuelans rally in Miami to protest Maduro's election claims

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Iran launches new missile attacks

Iran launches new missile attacks

September 19, 2026
California entrepreneur warns billionaire wealth tax could trigger ‘giant sucking sound’ of business exits

California entrepreneur warns billionaire wealth tax could trigger ‘giant sucking sound’ of business exits

September 19, 2026
Survivors return to flood-stricken communities in Nepal

Survivors return to flood-stricken communities in Nepal

September 19, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!