• bitcoinBitcoin(BTC)$85,930.005.92%
  • ethereumEthereum(ETH)$2,747.354.61%
  • tetherTether(USDT)$1.000.03%
  • binancecoinBNB(BNB)$798.074.26%
  • rippleXRP(XRP)$1.507.11%
  • usd-coinUSDC(USDC)$1.000.02%
  • solanaSolana(SOL)$117.326.79%
  • tronTRON(TRX)$0.3443010.23%
  • zcashZcash(ZEC)$1,479.321.45%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.010.00%
  • HyperliquidHyperliquid(HYPE)$92.820.35%
  • dogecoinDogecoin(DOGE)$0.09749812.21%
  • moneroMonero(XMR)$567.913.19%
  • whitebitWhiteBIT Coin(WBT)$86.424.44%
  • RainRain(RAIN)$0.0140870.14%
  • chainlinkChainlink(LINK)$12.883.59%
  • USDSUSDS(USDS)$1.000.01%
  • cardanoCardano(ADA)$0.2423906.51%
  • leo-tokenLEO Token(LEO)$8.88-0.58%
  • stellarStellar(XLM)$0.2085087.03%
  • uniswapUniswap(UNI)$8.830.42%
  • bitcoin-cashBitcoin Cash(BCH)$264.005.22%
  • nearNEAR Protocol(NEAR)$3.98-1.76%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • avalanche-2Avalanche(AVAX)$10.94-1.27%
  • litecoinLitecoin(LTC)$61.705.14%
  • daiDai(DAI)$1.000.01%
  • CantonCanton(CC)$0.1153417.75%
  • USD1USD1(USD1)$1.000.01%
  • suiSui(SUI)$1.0112.57%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.443.33%
  • hedera-hashgraphHedera(HBAR)$0.0910915.58%
  • shiba-inuShiba Inu(SHIB)$0.0000067.57%
  • MemeCoreMemeCore(M)$1.49-0.88%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • BittensorBittensor(TAO)$282.747.81%
  • crypto-com-chainCronos(CRO)$0.0636667.77%
  • paypal-usdPayPal USD(PYUSD)$1.000.02%
  • tether-goldTether Gold(XAUT)$4,350.10-0.41%
  • okbOKB(OKB)$122.504.45%
  • BitwayBitway(BTW)$0.9329.14%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.34%
  • aaveAave(AAVE)$142.795.43%
  • OndoOndo(ONDO)$0.4413904.47%
  • EthenaEthena(ENA)$0.210456-7.72%
  • mantleMantle(MNT)$0.634.30%
  • pepePepe(PEPE)$0.00000522.90%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Path: A Machine Learning Method for Training Small-Scale (Under 100M Parameter) Neural Information Retrieval Models with as few as 10 Gold Relevance Labels

June 26, 2024
in AI & Technology
Reading Time: 4 mins read
A A
Path: A Machine Learning Method for Training Small-Scale (Under 100M Parameter) Neural Information Retrieval Models with as few as 10 Gold Relevance Labels
ShareShareShareShareShare

The creative applications and management of pretrained language models have led to some great improvements in the quality of information retrieval (IR). Existing IR models are usually trained on large datasets comprising hundreds of thousands or even millions of queries and relevance judgments, especially those that can generalize to new, uncommon topics. 

The usefulness and necessity of such large-scale data for language model optimization for information retrieval tasks are questioned, raising scientific and engineering issues. In particular, it is not apparent from a scientific standpoint whether this massive amount of data is necessary, and from an engineering standpoint, it is not evident how to train IR models for languages with little or no labeled IR data or for niche domains.

YOU MAY ALSO LIKE

Tesla Will Soon Roll Out FSD Supervised In The Czech Republic

Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing

In recent research, a team of researchers from the University of Waterloo, Stanford University, and IBM Research AI has presented a technique for training small-scale neural information retrieval models using as few as ten gold relevance labels, that is, models with less than 100 million parameters. This approach has been named PATH – Prompts as Auto-optimized Training Hyperparameters. 

The foundation of this method is the creation of fictitious document queries via a language model (LM). The key innovation is that the language model automatically optimizes the prompt it uses to create these fictitious queries, guaranteeing that the training quality is optimized.

The team has shared the procedure, which is as follows. A text corpus and a very small number of relevant labels are the starting points. Then potential search queries are created that might be pertinent to the documents in the corpus using an LM. In order to create training data, pairs of queries and passages must be created. Optimizing the LM prompt, which directs the creation of the inquiry, is a crucial step in raising the caliber of the synthetic data in response to input from the training procedure.

Using the BIRCO benchmark, which consists of difficult and unusual IR tasks, the team has conducted trials and discovered that this approach greatly improves the performance of the trained models. In particular, the small-scale models outperform RankZephyr and are competitive with RankLLama, having been trained with minimally labeled data and optimized prompts. These later models, which included 7 billion parameters and were trained on datasets with more than 100,000 labels, are significantly larger.

These outcomes demonstrate how well automatic rapid optimization produces artificial datasets of superior quality. This approach not only shows that effective IR models can be trained with fewer resources, but it also shows that, with the right adjustments to the data creation process, smaller models can outperform much bigger models. 


Check out the Paper. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. 

Join our Telegram Channel and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 45k+ ML SubReddit


🚀 Create, edit, and augment tabular data with the first compound AI system, Gretel Navigator, now generally available! [Advertisement]


Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.

[Announcing Gretel Navigator] Create, edit, and augment tabular data with the first compound AI system trusted by EY, Databricks, Google, and Microsoft


Credit: Source link

ShareTweetSendSharePin

Related Posts

Tesla Will Soon Roll Out FSD Supervised In The Czech Republic
AI & Technology

Tesla Will Soon Roll Out FSD Supervised In The Czech Republic

September 21, 2026
Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing
AI & Technology

Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing

September 21, 2026
Collaboration Must Sit At the Heart of Manufacturing’s Multi-Agentic AI Approach. Here’s How. – Unite.AI
AI & Technology

Collaboration Must Sit At the Heart of Manufacturing’s Multi-Agentic AI Approach. Here’s How. – Unite.AI

September 21, 2026
How To Choose The Right USB To USB-C Adapter
AI & Technology

How To Choose The Right USB To USB-C Adapter

September 21, 2026
Next Post
Four women killed in Memphis shooting spree

Four women killed in Memphis shooting spree

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
A Hague court convicts Kosovo’s former President Thaci of war crimes and sentences him to 25 years – AP News

A Hague court convicts Kosovo’s former President Thaci of war crimes and sentences him to 25 years – AP News

September 16, 2026
SpaceX soars 5% after Starship reusable rocket launch date revealed

SpaceX soars 5% after Starship reusable rocket launch date revealed

September 16, 2026
Hero waiter tackles man trying to shoot ex-wife

Hero waiter tackles man trying to shoot ex-wife

September 14, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!