• bitcoinBitcoin(BTC)$77,984.001.35%
  • ethereumEthereum(ETH)$2,517.411.17%
  • tetherTether(USDT)$1.000.02%
  • binancecoinBNB(BNB)$720.580.38%
  • rippleXRP(XRP)$1.425.72%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$102.322.58%
  • tronTRON(TRX)$0.337886-0.04%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.00%
  • zcashZcash(ZEC)$1,166.088.79%
  • HyperliquidHyperliquid(HYPE)$80.453.04%
  • dogecoinDogecoin(DOGE)$0.0838481.21%
  • RainRain(RAIN)$0.014255-6.12%
  • USDSUSDS(USDS)$1.000.01%
  • moneroMonero(XMR)$516.620.18%
  • whitebitWhiteBIT Coin(WBT)$80.701.24%
  • chainlinkChainlink(LINK)$11.602.84%
  • leo-tokenLEO Token(LEO)$8.96-0.46%
  • cardanoCardano(ADA)$0.2082171.71%
  • stellarStellar(XLM)$0.1944248.98%
  • Ethena USDeEthena USDe(USDE)$1.000.02%
  • daiDai(DAI)$1.000.01%
  • bitcoin-cashBitcoin Cash(BCH)$222.850.28%
  • USD1USD1(USD1)$1.000.01%
  • uniswapUniswap(UNI)$6.707.95%
  • litecoinLitecoin(LTC)$52.90-2.36%
  • CantonCanton(CC)$0.0962250.28%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.35-0.13%
  • hedera-hashgraphHedera(HBAR)$0.0781243.25%
  • avalanche-2Avalanche(AVAX)$7.552.85%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • nearNEAR Protocol(NEAR)$2.455.19%
  • shiba-inuShiba Inu(SHIB)$0.0000050.55%
  • suiSui(SUI)$0.722.12%
  • paypal-usdPayPal USD(PYUSD)$1.000.02%
  • crypto-com-chainCronos(CRO)$0.0590732.40%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,298.82-1.09%
  • BittensorBittensor(TAO)$232.84-0.78%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • MemeCoreMemeCore(M)$1.10-3.14%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$112.70-1.12%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.13%
  • aaveAave(AAVE)$128.552.94%
  • BitwayBitway(BTW)$0.728.03%
  • mantleMantle(MNT)$0.572.45%
  • AsterAster(ASTER)$0.700.59%
  • pax-goldPAX Gold(PAXG)$4,301.78-1.11%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0573290.67%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

A New AI Research Introduces Recognize Anything Model (RAM): A Robust Base Model For Image Tagging

June 11, 2023
in AI & Technology
Reading Time: 4 mins read
A A
A New AI Research Introduces Recognize Anything Model (RAM): A Robust Base Model For Image Tagging
ShareShareShareShareShare

When it comes to natural language processing (NLP) tasks, large language models (LLM) trained on massive online datasets perform exceptionally well. Segment Anything Model (SAM) has shown outstanding zero-shot localization abilities in computer vision (CV) by scaling up data. 

Unfortunately, SAM cannot produce semantic labels, a fundamental task on par with localization. Recognizing many labels for a single image is the goal of multi-label image recognition, also known as image tagging. Since images contain various labels, including objects, sceneries, properties, and activities, image tagging is an important and useful computer vision problem.

Two main factors hinder image labeling as follows:

🚀 JOIN the fastest ML Subreddit Community
  1. The extensive collection of high-quality data. An efficient data annotation engine that can semi-automatically or automatically annotate massive amounts of photos across various categories is still lacking, as is a standardized and comprehensive labeling system.
  2. There are not enough open-vocabulary and powerful models built using an efficient and flexible model design that takes advantage of large-scale weakly-supervised data.

The Recognize Anything Model (RAM) is a robust base model for image tagging, and it has just been introduced by researchers at the OPPO Research Institute, the International Digital Economy Academy (IDEA), and AI2 Robotics. When it comes to data, RAM can overcome problems such as inadequate labeling systems, insufficient datasets, inefficient data engines, and architectural constraints.

The researchers start by creating a standard, global naming convention. They use academic datasets (classification, detection, and segmentation) and commercial taggers (Google, Microsoft, and Apple) to enrich their tagging system. By combining all available public tags with common text-based tags, the labeling method yields 6,449 labels that collectively address the vast majority of use cases. The researchers state that it is possible to recognize the remaining open-vocabulary labels using open-set recognition.

Annotating large-scale photographs using the label system automatically is a challenging task. The proposed approach to image tagging is inspired by previous work in the field, which uses large-scale public image-text pairs to train robust visual models. To put these massive amounts of picture-text data to good use for tagging, the team employed automatic text semantic parsing to extract the image tags. With this method, they could obtain a large set of picture tags based on image-text pairs without relying on manual annotations.

Internet-sourced image-text combinations tend to be imprecise due to random noise. The team creates a data tagging engine to improve the accuracy of annotations. To solve the problem of missing labels, they adopt preexisting models to produce supplementary classifications. When dealing with mislabeled areas, they pinpoint certain sections within the image that correlate to distinct labels. Then, they use region clustering methods to find and eliminate anomalies within the same category. In addition, the labels that make inconsistent predictions are also removed to get a more precise annotation. 

RAM permits generalization to novel classes by adding semantic context to label searches. RAM’s identification abilities can be boosted by this model architecture for any visual dataset, demonstrating its versatility. By showing that a general model trained on noisy, annotation-free data may beat highly supervised models, RAM introduces a new paradigm to picture tagging. RAM necessitates a free and publicly available dataset with no annotations. The most powerful version of RAM must only be trained for three days on eight A100 GPUs. 

According to the team, improvements can yet be made to RAM. This includes running many iterations of the data engine, increasing the backbone parameters to boost the model’s capacity, and expanding the training dataset beyond 14 million photos to better cover varied areas.


Check Out The Paper, Project, and Github. Don’t forget to join our 23k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]

🚀 Check Out 100’s AI Tools in AI Tools Club


YOU MAY ALSO LIKE

How To Use Meta Display Glasses While Driving With The Audio Only Feature

The EPA Wants To Stop Regulating Power Plant Emissions

Tanushree Shenwai is a consulting intern at MarktechPost. She is currently pursuing her B.Tech from the Indian Institute of Technology(IIT), Bhubaneswar. She is a Data Science enthusiast and has a keen interest in the scope of application of artificial intelligence in various fields. She is passionate about exploring the new advancements in technologies and their real-life application.


Check out https://aitoolsclub.com to find 100’s of Cool AI Tools

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Use Meta Display Glasses While Driving With The Audio Only Feature
AI & Technology

How To Use Meta Display Glasses While Driving With The Audio Only Feature

September 15, 2026
The EPA Wants To Stop Regulating Power Plant Emissions
AI & Technology

The EPA Wants To Stop Regulating Power Plant Emissions

September 14, 2026
Agent Harness vs Agent Framework vs MCP: Which Layer Owns the Loop, State, Tools, Permissions, and Recovery
AI & Technology

Agent Harness vs Agent Framework vs MCP: Which Layer Owns the Loop, State, Tools, Permissions, and Recovery

September 14, 2026
Reward AI Releases OM-1: A Robot Policy Trained on Human Demonstrations Only, With No Teleoperation or On-Robot Data
AI & Technology

Reward AI Releases OM-1: A Robot Policy Trained on Human Demonstrations Only, With No Teleoperation or On-Robot Data

September 14, 2026
Next Post
I’m Concerned About Retirement

I'm Concerned About Retirement

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Bunching Deductions Into One Year Can Beat the Standard Deduction

Bunching Deductions Into One Year Can Beat the Standard Deduction

September 12, 2026
LIVE NOW: CPI DATA INFLATION REPORT!

LIVE NOW: CPI DATA INFLATION REPORT!

September 13, 2026
Moment of silence held when second plane struck the World Trade Center

Moment of silence held when second plane struck the World Trade Center

September 13, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!