• bitcoinBitcoin(BTC)$83,554.000.70%
  • ethereumEthereum(ETH)$2,700.001.22%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$760.680.26%
  • rippleXRP(XRP)$1.533.28%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$119.631.33%
  • tronTRON(TRX)$0.3348840.28%
  • zcashZcash(ZEC)$1,426.53-7.05%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-3.43%
  • HyperliquidHyperliquid(HYPE)$87.12-0.53%
  • dogecoinDogecoin(DOGE)$0.0948512.68%
  • chainlinkChainlink(LINK)$15.005.64%
  • moneroMonero(XMR)$544.543.33%
  • whitebitWhiteBIT Coin(WBT)$83.690.89%
  • USDSUSDS(USDS)$1.00-0.01%
  • cardanoCardano(ADA)$0.2500333.66%
  • RainRain(RAIN)$0.012268-2.10%
  • leo-tokenLEO Token(LEO)$9.000.02%
  • stellarStellar(XLM)$0.2291505.84%
  • nearNEAR Protocol(NEAR)$4.97-0.01%
  • bitcoin-cashBitcoin Cash(BCH)$309.650.34%
  • uniswapUniswap(UNI)$8.952.14%
  • litecoinLitecoin(LTC)$68.02-1.81%
  • CantonCanton(CC)$0.128564-0.76%
  • avalanche-2Avalanche(AVAX)$11.3610.80%
  • hedera-hashgraphHedera(HBAR)$0.113580-3.21%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • suiSui(SUI)$1.160.95%
  • daiDai(DAI)$1.00-0.02%
  • USD1USD1(USD1)$1.000.00%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.55-3.97%
  • BittensorBittensor(TAO)$313.244.97%
  • BitwayBitway(BTW)$1.30-7.18%
  • crypto-com-chainCronos(CRO)$0.0701574.16%
  • quant-networkQuant(QNT)$239.382.78%
  • shiba-inuShiba Inu(SHIB)$0.0000064.36%
  • tether-goldTether Gold(XAUT)$4,166.860.99%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • Pump.funPump.fun(PUMP)$0.00574714.14%
  • aaveAave(AAVE)$172.9118.82%
  • EthenaEthena(ENA)$0.253964-1.67%
  • okbOKB(OKB)$120.502.76%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • OndoOndo(ONDO)$0.510.27%
  • MemeCoreMemeCore(M)$1.06-7.80%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.02%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

How can Businesses Improve the Accuracy of Multilingual Product Classifiers? This AI Paper Proposes LAMM: An Active Learning Approach Aimed at Bolstering the Classification Accuracy in Languages with Limited Training Data

August 24, 2023
in AI & Technology
Reading Time: 4 mins read
A A
How can Businesses Improve the Accuracy of Multilingual Product Classifiers? This AI Paper Proposes LAMM: An Active Learning Approach Aimed at Bolstering the Classification Accuracy in Languages with Limited Training Data
ShareShareShareShareShare

By capitalizing on shared representations common to different languages, cross-lingual learning is known to enhance the accuracy of NLP models on Low-Resource Languages (LRLs) that have limited data for model training. However, there is a significant disparity in the accuracy of high-resource languages (HRLs) and low-resource languages (LRLs), and this connects to the relative scarcity of pre-training data from the LRLs, even for state-of-the-art (SOTA) models. Targets for language-level accuracy are frequently imposed in professional contexts. This is when techniques like neural machine translation, transliteration, and label propagation on similar data are useful since they may be used to enhance the existing training data synthetically.

These methods can be used to augment the quantity and quality of training data without resorting to prohibitively expensive manual annotation. As a result of the limitations of machine translation, it may need to catch up to the commercial goals even though translation usually improves LRL accuracy.

A team of researchers from Amazon offers an approach to improving low-resource language (LRL) accuracy by employing active learning to collect labeled data selectively. Active learning for multilingual data has previously been studied, although most focus has been on training a model for a single language. To that end, they are working to perfect a single model that can effectively translate between languages. The method, Language Aware Active Learning for Multilingual Models (LAMM), is analogous to the work, which showed that active learning can improve model performance across languages while utilizing a single model. This approach does not, unfortunately, offer a means of specifically targeting and enhancing an LRL’s accuracy. As a result of their insistence on getting labels for languages that have already exceeded their accuracy objectives, today’s state-of-the-art active learning algorithms waste manual annotations in situations where meeting language-level targets is essential. To improve LRL accuracy without negatively impacting HRL performance, they present an active-learning-based strategy for collecting labeled data strategically. The suggested strategy, LAMM, enhances the likelihood of achieving accuracy targets across all relevant languages.

Researchers frame LAMM as an MOP with multiple goals to achieve. The objective is to pick examples of unlabeled data that are:

  • Indeterminate (the model has little faith in its results)
  • From language families, the classifier’s performance could be better than the goals.  

Amazon researchers compare LAMM’s performance to two benchmarks on four multilingual classification datasets using the typical pool-based active learning setup. Two examples of public datasets are Amazon Reviews and MLDoc. Two multilingual product classification datasets are used internally by Amazon. These are the standard procedures:

  • Least Confidence (LC) gathers the most entropically uncertain samples possible.
  • Equal Allocation (EC), to fill the per-language annotation budget, samples with high entropy are gathered, and the annotation budget is divided equally across the languages.

They found that LAMM outperforms the competition on all LRLs while only slightly underperforming on HRL. The percentage of HRL labels is reduced by 62.1% when using LAMM, although the accuracy of AUC is reduced by just 1.2% when comparing LAMM to LC. Using four different datasets for product classification, two publicly available and two proprietary, they show that LAMM can increase LRL performance by 4–11% relative to robust baselines.


Check out the Paper. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 29k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, please follow us on Twitter


YOU MAY ALSO LIKE

Why Wi-Fi Extenders Simply Aren’t Worth Buying In 2026

Nothing’s Flagship $399 Headphone 1 Pro Actually Have Some Professional Features

Dhanshree Shenwai is a Computer Science Engineer and has a good experience in FinTech companies covering Financial, Cards & Payments and Banking domain with keen interest in applications of AI. She is enthusiastic about exploring new technologies and advancements in today’s evolving world making everyone’s life easy.


🚀 CodiumAI enables busy developers to generate meaningful tests (Sponsored)


Credit: Source link

ShareTweetSendSharePin

Related Posts

Why Wi-Fi Extenders Simply Aren’t Worth Buying In 2026
AI & Technology

Why Wi-Fi Extenders Simply Aren’t Worth Buying In 2026

September 29, 2026
Nothing’s Flagship 9 Headphone 1 Pro Actually Have Some Professional Features
AI & Technology

Nothing’s Flagship $399 Headphone 1 Pro Actually Have Some Professional Features

September 29, 2026
Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitting
AI & Technology

Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitting

September 29, 2026
OpenAI Reportedly Cancels GPT-6.1 Astra’s Release Over Deceptive Behavior
AI & Technology

OpenAI Reportedly Cancels GPT-6.1 Astra’s Release Over Deceptive Behavior

September 29, 2026
Next Post
Baldur’s Gate III is coming to Xbox this year after a Series S compromise

Baldur's Gate III is coming to Xbox this year after a Series S compromise

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
My Financial Advisor Tells Me I Can’t Afford What I Want

My Financial Advisor Tells Me I Can’t Afford What I Want

September 26, 2026
Meta Bets on AI, Devices for Its Next Chapter

Meta Bets on AI, Devices for Its Next Chapter

September 28, 2026
Cricut’s New DIY Machines Let You Print And Cut Your Own Stickers

Cricut’s New DIY Machines Let You Print And Cut Your Own Stickers

September 25, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!