• bitcoinBitcoin(BTC)$77,073.00-0.53%
  • ethereumEthereum(ETH)$2,483.78-0.27%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$721.420.60%
  • rippleXRP(XRP)$1.455.38%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$101.280.62%
  • tronTRON(TRX)$0.338287-0.54%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.66%
  • zcashZcash(ZEC)$1,137.462.26%
  • HyperliquidHyperliquid(HYPE)$79.720.61%
  • dogecoinDogecoin(DOGE)$0.0832420.22%
  • USDSUSDS(USDS)$1.000.00%
  • moneroMonero(XMR)$519.452.06%
  • whitebitWhiteBIT Coin(WBT)$79.77-0.56%
  • RainRain(RAIN)$0.013197-12.00%
  • chainlinkChainlink(LINK)$11.471.95%
  • leo-tokenLEO Token(LEO)$8.970.12%
  • cardanoCardano(ADA)$0.2072200.19%
  • stellarStellar(XLM)$0.1979174.40%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • daiDai(DAI)$1.000.01%
  • bitcoin-cashBitcoin Cash(BCH)$224.411.53%
  • USD1USD1(USD1)$1.00-0.02%
  • uniswapUniswap(UNI)$6.738.27%
  • litecoinLitecoin(LTC)$52.74-1.22%
  • CantonCanton(CC)$0.095388-0.10%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.340.01%
  • hedera-hashgraphHedera(HBAR)$0.0797764.59%
  • avalanche-2Avalanche(AVAX)$7.592.35%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • nearNEAR Protocol(NEAR)$2.400.57%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.26%
  • suiSui(SUI)$0.720.08%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • crypto-com-chainCronos(CRO)$0.057731-1.27%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,288.660.56%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$225.99-2.00%
  • MemeCoreMemeCore(M)$1.11-1.31%
  • okbOKB(OKB)$112.79-0.47%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.15-0.04%
  • aaveAave(AAVE)$128.773.35%
  • BitwayBitway(BTW)$0.700.97%
  • AsterAster(ASTER)$0.700.67%
  • pax-goldPAX Gold(PAXG)$4,290.630.44%
  • mantleMantle(MNT)$0.56-0.17%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0575080.82%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

This AI Paper from Intel Presents a SYCL Implementation of Fully Fused Multi-Layer Perceptrons (MLPs) on Intel Data Center GPU Max

March 30, 2024
in AI & Technology
Reading Time: 4 mins read
A A
This AI Paper from Intel Presents a SYCL Implementation of Fully Fused Multi-Layer Perceptrons (MLPs) on Intel Data Center GPU Max
ShareShareShareShareShare

In the field of Artificial Intelligence (AI), Multi-Layer Perceptrons (MLPs) are the foundation for many Machine Learning (ML) tasks, including partial differential equation solving, density function representation in Neural Radiance Fields (NeRFs), and ray tracing simulation using Neural Ray Tracing.

Fully connected layers, in which every neuron in a layer is connected to every other neuron in the layer above and below, are a defining characteristic of MLPs. In MLPs, every neuron’s output is independent of the output of its nearby neurons in the same layer, in contrast to certain other topologies. Because of this property, MLPs can be used for fully fusing processes, which is essential for some computational workloads.

In recent research, a team of researchers from Intel Corporation and Ecole Polytechnique has focussed on effectively building narrow MLPs on Intel GPUs. Narrow MLPs feature a tiny, fixed number of neurons per layer and a shallow depth, i.e., the number of layers. Narrow MLPs are universal approximators that have significance in a wide range of applications despite their narrow width. Their narrow breadth, however, limits their performance, leading to low memory bandwidth utilization and arithmetic intensity during training and inference.

Combining the layers into a single kernel is a popular solution to these problems, as it allows for the use of quicker memories such as caches, shared memory, and register files. This method, called ‘fully-fused MLPs,’ was previously utilized with CUDA to construct Nvidia GPUs.

The team has shared that the goal of this study is to create fully-fused MLPs with a fixed layer width of 2^i neurons and arbitrary depth using SYCL for Intel GPUs (where i varies from 4 to 7). These MLPs are effective universal approximators in spite of the fixed layer width. Utilizing the XMX technology in Intel’s Data Centre GPU Max 1550, the implementation is based on Intel’s joint matrix SYCL extensions.

Models requiring high data throughput with batch sizes of 2^i, where i is more than 15, are especially well suited for this technique. Compared to comparable CUDA implementations, the Intel hardware SYCL version performs better, particularly for 64-width MLPs. A study has also indicated that this method requires less access to global memory than prior ones, which improves inference acceleration and theoretical peak performance. 

Benchmarks and applications, including Image Compression, Neural Radiance Fields (NeRFs), and Physics-Informed Machine Learning, have been tested in order to demonstrate performance improvements and possible applications. The provided approach performs significantly better than off-the-shelf implementations such as the CUDA PyTorch version on Nvidia’s H100 GPU and Intel Extension for PyTorch (IPEX) on the same Intel GPU in all circumstances.

The team has summarized their primary contributions as follows.

  1. The first SYCL implementation for fully-fused Multi-Layer Perceptrons designed for Intel GPUs using XMX instructions has been introduced. 
  1. The performance of the implementation has been assessed using a roofline model, which shows a rise in arithmetic intensity of up to 2.15 times when compared to a fully-fused implementation.
  1. Four sample applications have been used to validate the higher performance: the regression benchmark, image compression, neural radiation fields, and physics-informed neural networks. 
  1. The implementation is noteworthy because it can perform training 1.75 times quicker and inference 2.84 times faster than another fully-fused implementation. Its effectiveness across a variety of activities and datasets has been further demonstrated by the up to 30 times performance improvement it delivers over commercially available PyTorch versions.

Check out the Paper and Github. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. Join our Telegram Channel, Discord Channel, and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 39k+ ML SubReddit


YOU MAY ALSO LIKE

Salesforce Debuts Koa Reasoning Model for Agentforce, Trained on Nemotron – Unite.AI

2 Ways Android Users Can Take Advantage Of Apple’s MagSafe Accessories

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

Salesforce Debuts Koa Reasoning Model for Agentforce, Trained on Nemotron – Unite.AI
AI & Technology

Salesforce Debuts Koa Reasoning Model for Agentforce, Trained on Nemotron – Unite.AI

September 15, 2026
2 Ways Android Users Can Take Advantage Of Apple’s MagSafe Accessories
AI & Technology

2 Ways Android Users Can Take Advantage Of Apple’s MagSafe Accessories

September 15, 2026
Apple TV Cleaned Up At The Emmys With Eight Wins For Widow’s Bay And Pluribus
AI & Technology

Apple TV Cleaned Up At The Emmys With Eight Wins For Widow’s Bay And Pluribus

September 15, 2026
Elsevier Integrates LG AI Research’s Chemistry Vision Model Into Reaxys – Unite.AI
AI & Technology

Elsevier Integrates LG AI Research’s Chemistry Vision Model Into Reaxys – Unite.AI

September 15, 2026
Next Post
Verdict in the Proud Boys trial could have ‘significant’ implications for Trump

Verdict in the Proud Boys trial could have ‘significant’ implications for Trump

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Ex-Apple Researchers’ Nuance Labs Lands M Series A Backed by Nvidia – Unite.AI

Ex-Apple Researchers’ Nuance Labs Lands $50M Series A Backed by Nvidia – Unite.AI

September 14, 2026
Harvard fellow ‘deeply concerned’ over efforts to build bioweapons after Anthropic report

Harvard fellow ‘deeply concerned’ over efforts to build bioweapons after Anthropic report

September 12, 2026
Western Digital Corporation (WDC) Presents at Citi’s 2026 Global TMT Conference Transcript

Western Digital Corporation (WDC) Presents at Citi’s 2026 Global TMT Conference Transcript

September 9, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!