• bitcoinBitcoin(BTC)$84,878.000.59%
  • ethereumEthereum(ETH)$2,687.420.75%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$788.262.98%
  • rippleXRP(XRP)$1.491.06%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$120.011.62%
  • tronTRON(TRX)$0.3357340.32%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-2.03%
  • zcashZcash(ZEC)$1,328.313.18%
  • HyperliquidHyperliquid(HYPE)$89.262.95%
  • dogecoinDogecoin(DOGE)$0.0932561.75%
  • moneroMonero(XMR)$561.055.10%
  • chainlinkChainlink(LINK)$14.052.96%
  • whitebitWhiteBIT Coin(WBT)$84.400.51%
  • USDSUSDS(USDS)$1.000.02%
  • cardanoCardano(ADA)$0.2455602.06%
  • leo-tokenLEO Token(LEO)$8.990.43%
  • RainRain(RAIN)$0.0113762.04%
  • stellarStellar(XLM)$0.2162472.08%
  • bitcoin-cashBitcoin Cash(BCH)$315.723.31%
  • nearNEAR Protocol(NEAR)$4.793.16%
  • uniswapUniswap(UNI)$9.073.58%
  • litecoinLitecoin(LTC)$69.150.18%
  • avalanche-2Avalanche(AVAX)$11.124.97%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.1226703.51%
  • suiSui(SUI)$1.184.66%
  • Blockchain USDBlockchain USD(USDB)$0.871,000.00%
  • daiDai(DAI)$1.000.01%
  • hedera-hashgraphHedera(HBAR)$0.1026642.76%
  • USD1USD1(USD1)$1.000.01%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.511.62%
  • BitwayBitway(BTW)$1.524.22%
  • quant-networkQuant(QNT)$258.1112.85%
  • shiba-inuShiba Inu(SHIB)$0.0000062.64%
  • BittensorBittensor(TAO)$297.683.91%
  • tether-goldTether Gold(XAUT)$4,140.79-0.07%
  • crypto-com-chainCronos(CRO)$0.0662150.53%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • Pump.funPump.fun(PUMP)$0.00631317.71%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • aaveAave(AAVE)$180.18-0.96%
  • okbOKB(OKB)$120.440.29%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • EthenaEthena(ENA)$0.2387701.56%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • OndoOndo(ONDO)$0.4930821.21%
  • MemeCoreMemeCore(M)$1.04-3.30%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.00%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

MIT Researchers Introduce A Novel Lightweight Multi-Scale Attention For On-Device Semantic Segmentation

September 15, 2023
in AI & Technology
Reading Time: 5 mins read
A A
MIT Researchers Introduce A Novel Lightweight Multi-Scale Attention For On-Device Semantic Segmentation
ShareShareShareShareShare

The goal of semantic segmentation, a fundamental problem in computer vision, is to classify each pixel in the input image with a certain class. Autonomous driving, medical image processing, computational photography, etc., are just a few real-world contexts where semantic segmentation can be useful. Therefore, there is a high demand for installing SOTA semantic segmentation models on edge devices to benefit various consumers. However, SOTA semantic segmentation models have high processing requirements that edge devices cannot meet. This prevents these models from being used on edge devices. Semantic segmentation, in particular, is an example of a dense prediction task that necessitates high-resolution images and robust context information extraction capability. Therefore, transferring the effective model architecture used in image classification and applying it to semantic segmentation is inappropriate.

When asked to classify the millions of individual pixels in a high-resolution image, machine learning models face a formidable challenge. Recently, a highly effective use of a novel sort of model called a vision transformer has emerged.

The original intent of transformers was to improve the efficiency of NLP for languages. In such a setting, they tokenize the words in a sentence and create a network diagram that displays how those words are connected. The attention map enhances the model’s ability to comprehend context.

To generate an attention map, a vision transformer uses the same idea, slicing an image into patches of pixels and encoding each little patch into a token. The model employs a similarity function that learns the direct interaction between every pair of pixels to generate this attention map. By doing so, the model creates a “global receptive field,” allowing it to perceive all the important details in the image.

The attention map soon grows very large since a high-resolution image may include millions of pixels divided into thousands of patches. As a result, the computation required to process an image with increasing resolution climbs at a quadratic rate.

The MIT team replaced the nonlinear similarity function with a linear one to simplify the method used to construct the attention map in their new model series, dubbed EfficientViT. Because of this, the order in which operations are performed can be changed to reduce the number of calculations required without compromising functionality or the global receptive field, and with their approach, the amount of processing time needed to make a forecast scales linearly with the pixel count of the input image.

New models in the EfficientViT family do semantic segmentation locally on the device. EfficientViT is built around a novel lightweight multi-scale attention module for hardware-efficient global receptive field and multi-scale learning. Previous approaches for semantic segmentation in SOTA inspired this component.

The module was created to provide access to these two essential functionalities while minimizing the need for inefficient hardware operations. Specifically, we propose replacing the inefficient self-attention with lightweight ReLU-based global attention to achieve an international receptive field. The computational complexity of ReLU-based global attention can be reduced from quadratic to linear while keeping functionality by taking advantage of the associative property of matrix multiplication. And because it doesn’t use hardware-intensive algorithms like softmax, it’s better suited to on-device semantic segmentation.

Popular semantic segmentation benchmark datasets like Cityscapes and ADE20K have been used to conduct in-depth evaluations of EfficientViT. Compared to earlier SOTA semantic segmentation models, EfficientViT offers substantial performance improvements.

The following is a synopsis of the contributions:

  • Researchers have developed a revolutionary lightweight multi-scale attention to do semantic segmentation locally on the device. It performs well on edge devices while implementing a global receptive field and multi-scale learning.
  • Researchers developed a new family of models called EfficientViT based on the proposed lightweight multi-scale attention module.
  • The model shows a significant speedup on mobile over previous SOTA semantic segmentation models on prominent semantic segmentation benchmark datasets like ImageNet.

In conclusion, MIT researchers introduced a lightweight multi-scale attention module that achieves a global receptive field and multi-scale learning with light and hardware-efficient operations, thus providing significant speedup on edge devices without performance loss compared to SOTA semantic segmentation models. The EfficientViT models will be further scaled up, and their potential for use in other vision tasks will be investigated in further research.


Check out the Paper and Reference Article. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 30k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

How To Check The Temperature Of Your PC’s CPU

How To Customize Camera Control On Your iPhone

Dhanshree Shenwai is a Computer Science Engineer and has a good experience in FinTech companies covering Financial, Cards & Payments and Banking domain with keen interest in applications of AI. She is enthusiastic about exploring new technologies and advancements in today’s evolving world making everyone’s life easy.


🚀 The end of project management by humans (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Check The Temperature Of Your PC’s CPU
AI & Technology

How To Check The Temperature Of Your PC’s CPU

October 3, 2026
How To Customize Camera Control On Your iPhone
AI & Technology

How To Customize Camera Control On Your iPhone

October 3, 2026
What Does FDM Stand For In 3D Printing And How Does It Work?
AI & Technology

What Does FDM Stand For In 3D Printing And How Does It Work?

October 3, 2026
How To Adjust The Audio Quality In Apple Music
AI & Technology

How To Adjust The Audio Quality In Apple Music

October 3, 2026
Next Post
Digital Video Hits IPO Market

Digital Video Hits IPO Market

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
OpenAI halts training of latest models as reports mount of AI agents going rogue – The Guardian

OpenAI halts training of latest models as reports mount of AI agents going rogue – The Guardian

September 27, 2026
Tornadoes hit Kansas, severe storm floods multiple states

Tornadoes hit Kansas, severe storm floods multiple states

October 2, 2026
Jev – The New AI model that has people talking

Jev – The New AI model that has people talking

September 28, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!