• bitcoinBitcoin(BTC)$80,349.00-1.22%
  • ethereumEthereum(ETH)$2,578.90-2.37%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$750.44-1.94%
  • rippleXRP(XRP)$1.38-2.25%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$108.34-3.11%
  • tronTRON(TRX)$0.3415451.10%
  • zcashZcash(ZEC)$1,438.62-7.98%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.02-1.31%
  • HyperliquidHyperliquid(HYPE)$90.80-1.38%
  • dogecoinDogecoin(DOGE)$0.084978-2.56%
  • moneroMonero(XMR)$522.14-9.78%
  • whitebitWhiteBIT Coin(WBT)$81.77-1.86%
  • USDSUSDS(USDS)$1.00-0.01%
  • RainRain(RAIN)$0.013339-4.49%
  • chainlinkChainlink(LINK)$12.03-3.40%
  • cardanoCardano(ADA)$0.220443-1.43%
  • leo-tokenLEO Token(LEO)$8.920.23%
  • stellarStellar(XLM)$0.190375-1.08%
  • uniswapUniswap(UNI)$8.78-4.16%
  • bitcoin-cashBitcoin Cash(BCH)$246.53-0.85%
  • Ethena USDeEthena USDe(USDE)$1.00-0.02%
  • nearNEAR Protocol(NEAR)$3.54-2.91%
  • daiDai(DAI)$1.00-0.01%
  • litecoinLitecoin(LTC)$57.05-0.09%
  • USD1USD1(USD1)$1.00-0.02%
  • avalanche-2Avalanche(AVAX)$9.748.12%
  • CantonCanton(CC)$0.103982-5.43%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.370.21%
  • hedera-hashgraphHedera(HBAR)$0.0816012.81%
  • suiSui(SUI)$0.82-0.72%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.58%
  • MemeCoreMemeCore(M)$1.366.02%
  • crypto-com-chainCronos(CRO)$0.057905-2.24%
  • BittensorBittensor(TAO)$252.85-3.91%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • tether-goldTether Gold(XAUT)$4,372.410.04%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • okbOKB(OKB)$115.71-0.97%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.43%
  • AkedoAkedo(AKE)$0.09716058.95%
  • aaveAave(AAVE)$136.57-5.64%
  • EthenaEthena(ENA)$0.2067299.16%
  • OndoOndo(ONDO)$0.4098981.81%
  • AsterAster(ASTER)$0.73-3.06%
  • mantleMantle(MNT)$0.59-2.66%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers at the University of Wisconsin-Madison Propose a Finetuning Approach Utilizing a Carefully Designed Synthetic Dataset Comprising Numerical Key-Value Retrieval Tasks

July 3, 2024
in AI & Technology
Reading Time: 4 mins read
A A
Researchers at the University of Wisconsin-Madison Propose a Finetuning Approach Utilizing a Carefully Designed Synthetic Dataset Comprising Numerical Key-Value Retrieval Tasks
ShareShareShareShareShare

It is observed that LLMs often struggle to retrieve relevant information from the middle of long input contexts, exhibiting a “lost-in-the-middle” behavior. The research paper addresses the critical issue of the performance of large language models (LLMs) when handling longer-context inputs. Specifically, LLMs like GPT-3.5 Turbo and Mistral 7B often struggle with accurately retrieving information and maintaining reasoning capabilities across extensive textual data. This limitation hampers their effectiveness in tasks that require processing and reasoning over long passages, such as multi-document question answering (MDQA) and flexible length question answering (FLenQA). 

Current methods to enhance the performance of LLMs in long-context settings typically involve finetuning on real-world datasets. However, these datasets often include outdated or irrelevant information, which can lead to hallucinations and other inaccuracies. Traditional datasets such as MDQA and FLenQA have shown that LLMs tend to exhibit a “lost-in-the-middle” behavior, where their performance is optimal at the beginning or end of the input context but deteriorates for information in the middle.

YOU MAY ALSO LIKE

Crusoe CEO: Data Center Indusry Has a ‘Marketing Issue’

Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages

A team of researchers from the University of Wisconsin-Madison proposes a novel finetuning approach utilizing a carefully designed synthetic dataset to address these challenges. This dataset comprises numerical key-value retrieval tasks designed to enhance the LLMs’ ability to handle long contexts more effectively. By using synthetic data that avoids the pitfalls of outdated or irrelevant information, the researchers aim to improve LLMs’ information retrieval and reasoning capabilities without introducing hallucinations.

The proposed synthetic dataset consists of simple dictionary key-value retrieval tasks, where each task involves multiple dictionaries with a few keys each. For instance, the dataset for Mistral 7B includes 350 samples, each containing 85 dictionaries, resulting in prompts with roughly 3900 tokens. Finetuning is conducted on the answer part of these tasks, masking out other elements to focus the model’s learning process.

Experiments demonstrate that this approach significantly enhances the performance of LLMs in long-context tasks. For example, finetuning GPT-3.5 Turbo on the synthetic data resulted in a 10.5% improvement on the 20 documents MDQA benchmark at the tenth position. Moreover, this method mitigates the “lost-in-the-middle” phenomenon and reduces the primacy bias, leading to more accurate information retrieval across the entire input context. The performance of models finetuned on the synthetic data was compared against those finetuned on real-world datasets, with the synthetic approach showing superior results in maintaining consistent accuracy across different context positions. 

The study introduces an innovative approach to finetuning LLMs using synthetic data, significantly enhancing their performance in long-context settings. The proposed method demonstrates substantial improvements over traditional finetuning techniques by addressing the “lost-in-the-middle” phenomenon and reducing primacy bias. This research highlights the potential of synthetic datasets in overcoming the limitations of real-world data, paving the way for more effective and reliable LLMs in handling extensive textual information.


Check out the Paper. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. 

Join our Telegram Channel and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 45k+ ML SubReddit


Shreya Maji is a consulting intern at MarktechPost. She is pursued her B.Tech at the Indian Institute of Technology (IIT), Bhubaneswar. An AI enthusiast, she enjoys staying updated on the latest advancements. Shreya is particularly interested in the real-life applications of cutting-edge technology, especially in the field of data science.

🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

Crusoe CEO: Data Center Indusry Has a ‘Marketing Issue’
AI & Technology

Crusoe CEO: Data Center Indusry Has a ‘Marketing Issue’

September 20, 2026
Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages
AI & Technology

Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages

September 20, 2026
How Long Can You Expect Your Old Cassette Tapes To Last?
AI & Technology

How Long Can You Expect Your Old Cassette Tapes To Last?

September 20, 2026
How To Record Audio On Your iPhone
AI & Technology

How To Record Audio On Your iPhone

September 20, 2026
Next Post
Dave Ramsey has a stern message for Gen Z

Dave Ramsey has a stern message for Gen Z

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Indonesia cancels flights as Anak Krakatau volcano erupts

Indonesia cancels flights as Anak Krakatau volcano erupts

September 16, 2026
Baby goats spark rabies fears in North Carolina

Baby goats spark rabies fears in North Carolina

September 20, 2026
Tibet side of border shows flood’s devastating impact

Tibet side of border shows flood’s devastating impact

September 16, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!