• bitcoinBitcoin(BTC)$78,859.00-1.01%
  • ethereumEthereum(ETH)$2,474.11-0.25%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$737.42-1.21%
  • rippleXRP(XRP)$1.39-1.54%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$103.58-2.18%
  • tronTRON(TRX)$0.334027-0.39%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.060.00%
  • zcashZcash(ZEC)$1,149.91-2.03%
  • HyperliquidHyperliquid(HYPE)$85.44-3.16%
  • dogecoinDogecoin(DOGE)$0.0894290.55%
  • RainRain(RAIN)$0.016330-2.91%
  • moneroMonero(XMR)$533.471.38%
  • USDSUSDS(USDS)$1.000.00%
  • chainlinkChainlink(LINK)$12.844.84%
  • whitebitWhiteBIT Coin(WBT)$72.68-0.92%
  • leo-tokenLEO Token(LEO)$9.15-1.91%
  • cardanoCardano(ADA)$0.2190520.41%
  • stellarStellar(XLM)$0.1913433.88%
  • bitcoin-cashBitcoin Cash(BCH)$266.183.62%
  • daiDai(DAI)$1.000.01%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • litecoinLitecoin(LTC)$55.551.67%
  • USD1USD1(USD1)$1.000.01%
  • uniswapUniswap(UNI)$6.80-4.02%
  • CantonCanton(CC)$0.105486-2.94%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.40-1.60%
  • hedera-hashgraphHedera(HBAR)$0.0816211.35%
  • avalanche-2Avalanche(AVAX)$8.065.63%
  • suiSui(SUI)$0.811.73%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • shiba-inuShiba Inu(SHIB)$0.0000050.75%
  • nearNEAR Protocol(NEAR)$2.31-4.43%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.057225-0.07%
  • tether-goldTether Gold(XAUT)$4,410.52-0.25%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.12-0.51%
  • BittensorBittensor(TAO)$259.615.15%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$114.441.43%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.22%
  • AsterAster(ASTER)$0.770.40%
  • mantleMantle(MNT)$0.624.50%
  • aaveAave(AAVE)$131.45-1.80%
  • pax-goldPAX Gold(PAXG)$4,413.87-0.28%
  • OndoOndo(ONDO)$0.3831221.98%
  • polkadotPolkadot(DOT)$1.0912.88%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Meet Semantic-SAM: A Universal Image Segmentation Model Which Segments And Recognizes Objects At Any Desired Granularity Based On User Input

July 16, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Meet Semantic-SAM: A Universal Image Segmentation Model Which Segments And Recognizes Objects At Any Desired Granularity Based On User Input
ShareShareShareShareShare

Artificial Intelligence has greatly advanced in recent times. Its current development, i.e., the introduction of Large Language Models, has gained everyone’s attention due to its incredible human-imitating capabilities. Not only Language processing, these models have also gained success in the field of Computer vision. Though the success of AI systems in Natural Language Processing and controllable image generation is remarkable, the field of pixel-level image understanding, including universal image segmentation, still has certain limitations. 

Image segmentation, which is the technique of splitting an image into different sections, has shown great improvements, but creating a universal picture segmentation model that can handle a variety of images with different granularities is still in discussion. The two primary challenges to progress in this area are the availability of adequate training data and restrictions on the flexibility of model design. Existing methods frequently use a single-input, single-output pipeline that cannot forecast segmentation masks at various granularities and handle levels of detail. Also, it is expensive to scale up segmentation datasets with both semantic and granularity knowledge.

To address these limitations, a team of researchers has introduced Semantic-SAM, a universal image segmentation model which segments and recognizes objects at any desired granularity based on user input. The model is capable of providing semantic labels for both objects and pieces and predicts masks at various granularities in response to a user click. The decoder architecture of Semantic-SAM incorporates a multi-choice learning strategy to give the model the capacity to handle several granularities. Each click is represented by numerous queries, each of which has a distinct level of embedding. The queries are trained to learn from ground-truth masks with dissimilar granularities.

[Sponsored] 🔥 Build your personal brand with Taplio  🚀 The 1st all-in-one AI-powered tool to grow on LinkedIn. Create better LinkedIn content 10x faster, schedule, analyze your stats & engage. Try it for free!

The team has shared how Semantic-SAM tackles the problem of semantic awareness by using a decoupled categorization strategy for parts and objects. The model individually encodes objects and parts using a shared text encoder, enabling distinct segmentation procedures while changing the loss function according to the input type. This strategy guarantees that the model can handle data from the SAM dataset, which lacks some categorization labels, as well as data from general segmentation data.

The team has combined seven datasets that represent various granularities in order to enhance semantics and granularity, including the SA-1B dataset, part segmentation datasets like PASCAL Part, PACO, and PartImagenet, and generic segmentation datasets like MSCOCO and Objects365. The data formats have been rearranged to comply with Semantic-SAM’s training goals. 

Upon evaluation and testing, Semantic-SAM has demonstrated superior performance as compared to existing models. Performance is significantly improved when interactive segmentation techniques like SA-1B promptable segmentation and COCO panoptic segmentation are used in conjunction with training. A stunning 2.3 box AP gain and 1.2 mask AP gain are achieved by the model. It also performs better than SAM by more than 3.4 1-IoU in terms of granularity completeness.

Semantic-SAM is definitely an innovative advancement in the field of image segmentation. This model creates new opportunities for pixel-level image analysis by merging universal representation, semantic awareness, and granularity abundance.


Check out the Paper and GitHub link. Don’t forget to join our 26k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]

🚀 Check Out 800+ AI Tools in AI Tools Club


YOU MAY ALSO LIKE

Grupo Financiero Inbursa Adopts Harvey Across Its Legal Organization – Unite.AI

How To Find And Hide An App On Android Auto

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


🔥 StoryBird.ai just dropped some amazing features. Generate an illustrated story from a prompt. Check it out here. (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

Grupo Financiero Inbursa Adopts Harvey Across Its Legal Organization – Unite.AI
AI & Technology

Grupo Financiero Inbursa Adopts Harvey Across Its Legal Organization – Unite.AI

September 7, 2026
How To Find And Hide An App On Android Auto
AI & Technology

How To Find And Hide An App On Android Auto

September 7, 2026
How To Change Siri’s Voice
AI & Technology

How To Change Siri’s Voice

September 7, 2026
What Is a Foundation Model? How General-Purpose AI Is Built and Adapted – Unite.AI
AI & Technology

What Is a Foundation Model? How General-Purpose AI Is Built and Adapted – Unite.AI

September 7, 2026
Next Post
Wall Street Watches Earnings Reports and the Announcement From the FOMC on Wednesday

Wall Street Watches Earnings Reports and the Announcement From the FOMC on Wednesday

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Spanish firefighters save calf from smoke and flames

Spanish firefighters save calf from smoke and flames

September 5, 2026
Bodycam shows Tony Romo being arrested on suspicion of impaired driving

Bodycam shows Tony Romo being arrested on suspicion of impaired driving

August 31, 2026
Oil price surge as war with Iran expands

Oil price surge as war with Iran expands

September 6, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!