• bitcoinBitcoin(BTC)$78,224.00-0.29%
  • ethereumEthereum(ETH)$2,467.41-0.58%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$735.01-1.94%
  • rippleXRP(XRP)$1.40-0.89%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$102.38-0.74%
  • tronTRON(TRX)$0.3393860.45%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-1.63%
  • zcashZcash(ZEC)$1,242.328.08%
  • HyperliquidHyperliquid(HYPE)$84.810.78%
  • dogecoinDogecoin(DOGE)$0.086672-3.34%
  • RainRain(RAIN)$0.015952-3.32%
  • USDSUSDS(USDS)$1.00-0.01%
  • moneroMonero(XMR)$509.311.76%
  • whitebitWhiteBIT Coin(WBT)$80.75-0.60%
  • chainlinkChainlink(LINK)$11.77-6.05%
  • leo-tokenLEO Token(LEO)$9.19-0.21%
  • cardanoCardano(ADA)$0.213515-3.29%
  • stellarStellar(XLM)$0.182834-2.74%
  • bitcoin-cashBitcoin Cash(BCH)$254.73-0.60%
  • daiDai(DAI)$1.000.01%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.02%
  • litecoinLitecoin(LTC)$53.52-1.05%
  • CantonCanton(CC)$0.104826-2.64%
  • uniswapUniswap(UNI)$6.44-4.91%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.37-1.81%
  • avalanche-2Avalanche(AVAX)$7.86-1.56%
  • hedera-hashgraphHedera(HBAR)$0.077204-2.69%
  • Global DollarGlobal Dollar(USDG)$1.00-0.02%
  • nearNEAR Protocol(NEAR)$2.518.68%
  • suiSui(SUI)$0.78-3.32%
  • shiba-inuShiba Inu(SHIB)$0.000005-2.20%
  • crypto-com-chainCronos(CRO)$0.058869-0.61%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • MemeCoreMemeCore(M)$1.19-1.53%
  • tether-goldTether Gold(XAUT)$4,394.470.81%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$258.03-0.41%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • okbOKB(OKB)$112.82-0.94%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.06%
  • mantleMantle(MNT)$0.61-3.83%
  • AsterAster(ASTER)$0.74-1.29%
  • aaveAave(AAVE)$126.05-2.23%
  • pax-goldPAX Gold(PAXG)$4,397.700.87%
  • polkadotPolkadot(DOT)$1.12-10.73%
  • Pump.funPump.fun(PUMP)$0.0043340.11%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Revolutionizing Video Object Segmentation: Unveiling Cutie with Advanced Object-Level Memory Reading Techniques

October 28, 2023
in AI & Technology
Reading Time: 5 mins read
A A
Revolutionizing Video Object Segmentation: Unveiling Cutie with Advanced Object-Level Memory Reading Techniques
ShareShareShareShareShare

Tracking and segmenting objects from an open vocabulary defined in a first-frame annotation is necessary for Video Object Segmentation (VOS), more precisely, the “semisupervised” option. VOS techniques may be coupled with Segment Anything Models (SAMs) for all-purpose video segmentation (such as Tracking Anything) and for robotics, video editing, and cost-reduction in data annotation. Modern VOS methods use a memory-based paradigm. Any new query frame “reads” from this memory to extract features for segmentation. This memory representation is generated using previous segmented frames (either supplied as input or segmented by the model). 

Importantly, these methods create the segmentation bottom-up from the pixel memory readout and primarily employ pixel-level matching for memory reading, either with one or several matching layers. Pixel-level matching converts each memory pixel into a linear combination of query pixels (for example, using an attention layer). As a result, pixel-level matching has low-level consistency and is susceptible to matching noise, particularly when distractors are present. As a result, individuals perform worse in difficult situations, including occlusions and frequent distractions. Concretely, when assessing the recently suggested difficult MOSE dataset rather than the default DAVIS-2017 dataset, the performance of current techniques is more than 20 points in J & F worse. 

They believe the absence of object-level thinking is to blame for the disappointing outcomes in difficult cases. They suggest object-level memory reading to solve this problem, which effectively returns the object from memory to the query frame (Figure 1). They use an object transformer to achieve their object-level memory reading since current query-based object detection/segmentation methods that describe objects as “object queries” serve as inspiration. To 1) iteratively probe and calibrate a feature map (started by a pixel-level memory readout) and 2) encode object level information, this object transformer employs a limited collection of end-to-end trained object queries. This method allows for bidirectional top-down and bottom-up communication by maintaining a high-level/global object query representation and a low-level/high-resolution feature map. 

Figure 1 contrasts object-level memory reading with reading at the pixel level. The reference frame is on the left in each box, and the segmentable query frame is on the right. Wrong matches are shown with red arrows. When there are distractions, low-level pixel matching (like might become loud. For more reliable video object segmentation, we recommend object-level memory reading.

A series of attention layers, including a suggested foreground-background masked attention, are parameterized for this communication. Extended from foreground-only masked attention, masked attention allows some object queries to focus only on the foreground. In contrast, the remaining questions focus only on the background, enabling global feature interaction and clear foreground/background semantic distinction. Additionally, they incorporate a compact object memory (in addition to a pixel memory) to condense the characteristics of the target objects. With target-specific characteristics, this object memory improves end-to-end object searches and enables an effective long-term representation of target objects. 

In tests, the suggested method, Cutie, outperforms previous methods in difficult situations (such as +8.7 J & F in MOSE over XMem) while maintaining competitive accuracy and efficiency levels on common datasets like DAVIS and YouTubeVOS. In conclusion, researchers from the University of Illinois Urbana-Champaign and Adobe Research created Cutie, which has an object transformer for reading object-level memories. 

• It combines pixel-level bottom-up features with high-level top-down queries for effective video object segmentation in difficult situations with significant occlusions and distractions. 

• They extend the masked focus to the foreground and background to distinguish the target item from distractions while preserving the rich scene elements. 

• To store object characteristics in a compact form for later retrieval as target-specific object-level representations during querying, they build a compact object memory.


Check out the Paper, Project, and Github. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 32k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..

We are also on WhatsApp. Join our AI Channel on Whatsapp..


YOU MAY ALSO LIKE

OpenAI Names Paul Christiano to Foundation Board and Safety Committee – Unite.AI

Lightfield Raises $47M Series A Led by a16z to Accelerate Growth – Unite.AI

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


🔥 Meet Retouch4me: A Family of Artificial Intelligence-Powered Plug-Ins for Photography Retouching

Credit: Source link

ShareTweetSendSharePin

Related Posts

OpenAI Names Paul Christiano to Foundation Board and Safety Committee – Unite.AI
AI & Technology

OpenAI Names Paul Christiano to Foundation Board and Safety Committee – Unite.AI

September 9, 2026
Lightfield Raises M Series A Led by a16z to Accelerate Growth – Unite.AI
AI & Technology

Lightfield Raises $47M Series A Led by a16z to Accelerate Growth – Unite.AI

September 9, 2026
Everything Announced During Nintendo Direct
AI & Technology

Everything Announced During Nintendo Direct

September 9, 2026
Why It’s Time to Abandon the ‘Set It and Forget It’ Model – Unite.AI
AI & Technology

Why It’s Time to Abandon the ‘Set It and Forget It’ Model – Unite.AI

September 9, 2026
Next Post
Pennsylvania Man Dies After Pet Snake Wraps Itself Around His Neck

Pennsylvania Man Dies After Pet Snake Wraps Itself Around His Neck

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Semi-truck hit by train after getting stuck on the tracks

Semi-truck hit by train after getting stuck on the tracks

September 5, 2026
Great Americans: A conversation with travel writer Rick Steves

Great Americans: A conversation with travel writer Rick Steves

September 5, 2026
Carclo plc (CCEGF) Shareholder/Analyst Call Transcript

Carclo plc (CCEGF) Shareholder/Analyst Call Transcript

September 9, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!