• bitcoinBitcoin(BTC)$83,998.000.25%
  • ethereumEthereum(ETH)$2,687.92-0.08%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$773.26-0.15%
  • rippleXRP(XRP)$1.54-2.65%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$120.910.88%
  • tronTRON(TRX)$0.336321-0.12%
  • zcashZcash(ZEC)$1,538.40-2.57%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-0.14%
  • HyperliquidHyperliquid(HYPE)$92.410.47%
  • dogecoinDogecoin(DOGE)$0.0975070.37%
  • chainlinkChainlink(LINK)$14.404.19%
  • moneroMonero(XMR)$551.020.30%
  • whitebitWhiteBIT Coin(WBT)$83.790.11%
  • USDSUSDS(USDS)$1.000.00%
  • cardanoCardano(ADA)$0.2562551.31%
  • RainRain(RAIN)$0.0128978.89%
  • leo-tokenLEO Token(LEO)$8.961.58%
  • stellarStellar(XLM)$0.2181621.05%
  • bitcoin-cashBitcoin Cash(BCH)$334.541.38%
  • nearNEAR Protocol(NEAR)$4.84-4.24%
  • uniswapUniswap(UNI)$9.660.11%
  • litecoinLitecoin(LTC)$72.484.02%
  • CantonCanton(CC)$0.13722310.40%
  • Ethena USDeEthena USDe(USDE)$1.00-0.03%
  • avalanche-2Avalanche(AVAX)$10.996.67%
  • suiSui(SUI)$1.175.26%
  • daiDai(DAI)$1.00-0.01%
  • USD1USD1(USD1)$1.000.01%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.484.57%
  • hedera-hashgraphHedera(HBAR)$0.0940630.55%
  • BittensorBittensor(TAO)$326.918.01%
  • shiba-inuShiba Inu(SHIB)$0.0000062.33%
  • crypto-com-chainCronos(CRO)$0.0655970.05%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • EthenaEthena(ENA)$0.27962912.36%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • MemeCoreMemeCore(M)$1.213.01%
  • OndoOndo(ONDO)$0.551.56%
  • tether-goldTether Gold(XAUT)$4,279.910.27%
  • okbOKB(OKB)$121.150.89%
  • Ripple USDRipple USD(RLUSD)$1.00-0.02%
  • BitwayBitway(BTW)$0.92-22.70%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • aaveAave(AAVE)$154.604.77%
  • mantleMantle(MNT)$0.705.27%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.15-0.07%
  • polkadotPolkadot(DOT)$1.244.89%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Stanford study challenges assumptions about language models: Larger context doesn’t mean better understanding 

July 21, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Stanford study challenges assumptions about language models: Larger context doesn’t mean better understanding 
ShareShareShareShareShare

Head over to our on-demand library to view sessions from VB Transform 2023. Register Here


A study released this month by researchers from Stanford University, UC Berkeley and Samaya AI has found that large language models (LLMs) often fail to access and use relevant information given to them in longer context windows.

YOU MAY ALSO LIKE

These Xbox Players Got GTA 6 For Free The Hard Way

Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Building

In language models, a context window refers to the length of text a model can process and respond to in a given instance. It can be thought of as a working memory for a particular text analysis or chatbot conversation.

The study caught widespread attention last week after its release because many developers and other users experimenting with LLMs had assumed that the trend toward larger context windows would continue to improve LLM performance and their usefulness across various applications.

>>Don’t miss our special issue: The Future of the data center: Handling greater and greater demands.<<

Event

VB Transform 2023 On-Demand

Did you miss a session from VB Transform 2023? Register to access the on-demand library for all of our featured sessions.

 

Register Now

If an LLM could take an entire document or article as input for its context window, the conventional thinking went, the LLM could provide perfect comprehension of the full scope of that document when asked questions about it. 

Assumptions around context window flawed

LLM companies like Anthropic have fueled excitement around the idea of longer content windows, where users can provide ever more input to be analyzed or summarized. Anthropic just released a new model called Claude 2, which provides a huge 100k token context window, and said it can enable new use cases such as summarizing long conversations or drafting memos and op-eds.

But the study shows that some assumptions around the context window are flawed when it comes to the LLM’s ability to search and analyze it accurately. 

The study found that LLMs performed best “when relevant information occurs at the beginning or end of the input context, and significantly degrades when models must access relevant information in the middle of long contexts. Furthermore, performance substantially decreases as the input context grows longer, even for explicitly long-context models.”

Last week, industry insiders like Bob Wiederhold, COO of vector database company Pinecone, cited the study as evidence that stuffing entire documents into a document window for doing things like search and analysis won’t be the panacea many had hoped for. 

Semantic search preferable to document stuffing

Vector databases like Pinecone help developers increase LLM memory by searching for relevant information to pull into the context window. Wiederhold pointed to the study as evidence that vector databases will remain viable for the foreseeable future, since the study suggests semantic search provided by vector databases is better than document stuffing. 

Stanford University’s Nelson Liu, study lead author, agreed that if you try to inject an entire PDF into a language model context window and then ask questions about the document, a vector database search will generally be more efficient to use.

“If you’re searching over large amounts of documents, you want to be using something that’s built for search, at least for now,” said Liu. 

Liu cautioned, however, that the study isn’t necessarily claiming that sticking entire documents into a context window won’t work. Results will depend specifically on the sort of content contained in the documents the LLMs are analyzing. Language models are bad at differentiating between many things that are closely related or which seem relevant, Liu explained. But they are good at finding the one thing that is clearly relevant when most other things are not relevant.

“So I think it’s a bit more nuanced than ‘You should always use a vector database, or you should never use a vector database’,” he said.

Language models’ best use case: Generating content

Liu said his study assumed that most commercial applications are operating in a setting where they use some sort of vector database to help return multiple possible results into a context window. The study found that having more results in the context window didn’t always improve performance. 

As a specialist in language processing, Liu said he was surprised that people were thinking of using a context window to search for content, or to aggregate or synthesize it, although he said he could understand why people would want to. He said people should continue to think of language models as best used to generate content, and search engines as best to search content. 

“The hope that you can just throw everything into a language model and just sort of pray it works, I don’t think we’re there yet,” he said. “But maybe we’ll be there in a few years or even a few months. It’s not super clear to me how fast this space will move, but I think right now, language models aren’t going to replace vector databases and search engines.”

VentureBeat’s mission is to be a digital town square for technical decision-makers to gain knowledge about transformative enterprise technology and transact. Discover our Briefings.

Credit: Source link

ShareTweetSendSharePin

Related Posts

These Xbox Players Got GTA 6 For Free The Hard Way
AI & Technology

These Xbox Players Got GTA 6 For Free The Hard Way

September 26, 2026
Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Building
AI & Technology

Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Building

September 26, 2026
End-to-End Multimodal Data Augmentation and Adversarial Robustness Benchmark with AugLy for Images, Text, Audio, and PyTorch
AI & Technology

End-to-End Multimodal Data Augmentation and Adversarial Robustness Benchmark with AugLy for Images, Text, Audio, and PyTorch

September 26, 2026
Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding
AI & Technology

Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding

September 25, 2026
Next Post
Corona Light, Coors Light gain traction over Bud Light

Corona Light, Coors Light gain traction over Bud Light

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Oracle: OpenAI Just Blinked

Oracle: OpenAI Just Blinked

September 19, 2026
The Pros And Cons Of Using Wired Vs. Wireless Xbox Controllers

The Pros And Cons Of Using Wired Vs. Wireless Xbox Controllers

September 19, 2026
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB

Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB

September 25, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!