• bitcoinBitcoin(BTC)$80,977.004.21%
  • ethereumEthereum(ETH)$2,514.664.61%
  • tetherTether(USDT)$1.000.02%
  • binancecoinBNB(BNB)$724.784.60%
  • rippleXRP(XRP)$1.456.14%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$104.043.10%
  • tronTRON(TRX)$0.3296261.33%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.031.92%
  • HyperliquidHyperliquid(HYPE)$87.186.07%
  • zcashZcash(ZEC)$949.9815.78%
  • dogecoinDogecoin(DOGE)$0.0874335.38%
  • RainRain(RAIN)$0.0171542.59%
  • USDSUSDS(USDS)$1.000.00%
  • moneroMonero(XMR)$505.21-1.15%
  • chainlinkChainlink(LINK)$11.946.63%
  • whitebitWhiteBIT Coin(WBT)$73.943.83%
  • leo-tokenLEO Token(LEO)$9.32-0.09%
  • cardanoCardano(ADA)$0.2246539.04%
  • stellarStellar(XLM)$0.1839403.29%
  • bitcoin-cashBitcoin Cash(BCH)$256.203.38%
  • daiDai(DAI)$1.00-0.02%
  • CantonCanton(CC)$0.112035-1.13%
  • Ethena USDeEthena USDe(USDE)$1.000.03%
  • USD1USD1(USD1)$1.000.02%
  • uniswapUniswap(UNI)$6.4711.72%
  • litecoinLitecoin(LTC)$51.202.09%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.383.01%
  • hedera-hashgraphHedera(HBAR)$0.0783563.29%
  • avalanche-2Avalanche(AVAX)$7.523.71%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • suiSui(SUI)$0.780.57%
  • shiba-inuShiba Inu(SHIB)$0.0000052.89%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • crypto-com-chainCronos(CRO)$0.0579286.38%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,465.401.07%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • nearNEAR Protocol(NEAR)$1.962.80%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • MemeCoreMemeCore(M)$1.04-3.06%
  • okbOKB(OKB)$109.082.52%
  • BittensorBittensor(TAO)$228.884.84%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.00%
  • aaveAave(AAVE)$134.195.61%
  • AsterAster(ASTER)$0.720.03%
  • pax-goldPAX Gold(PAXG)$4,474.310.98%
  • mantleMantle(MNT)$0.571.20%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0579943.09%
  • OndoOndo(ONDO)$0.3650793.82%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Meet LLM-Blender: A Novel Ensembling Framework to Attain Consistently Superior Performance by Leveraging the Diverse Strengths of Multiple Open-Source Large Language Models (LLMs)

June 19, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Meet LLM-Blender: A Novel Ensembling Framework to Attain Consistently Superior Performance by Leveraging the Diverse Strengths of Multiple Open-Source Large Language Models (LLMs)
ShareShareShareShareShare

Large Language Models have shown remarkable performance in a massive range of tasks. From producing unique and creative content and questioning answers to translating languages and summarizing textual paragraphs, LLMs have been successful in imitating humans. Some well-known LLMs like GPT, BERT, and PaLM have been in the headlines for accurately following instructions and accessing vast amounts of high-quality data. Models like GPT4 and PaLM are not open-source, which prevents anyone from understanding their architectures and the training data. On the other hand, the open-source nature of LLMs like Pythia, LLaMA, and Flan-T5 provides an opportunity to researchers to fine-tune and improve the models on custom instruction datasets. This enables the development of smaller and more efficient LLMs like Alpaca, Vicuna, OpenAssistant, and MPT.

There is no single open-source LLM that leads the market, and the best LLMs for various examples can differ greatly from one another. Therefore, in order to continuously produce improved answers for each input, it is essential to dynamically ensemble these LLMs. Biases, errors, and uncertainties can be reduced by integrating the distinctive contributions of various LLMs, thus resulting in outcomes that more closely match human preferences. To address this, researchers from the Allen Institute for Artificial Intelligence, the University of Southern California, and Zhejiang University have proposed LLM-BLENDER, an ensembling framework that consistently obtains superior performance by utilizing the many advantages of several open-source large language models. 

YOU MAY ALSO LIKE

The Ternus Era At Apple Begins, But Cook Isn’t Leaving

Mobile Games Designed to Be Addictive Get More Kid-Friendly

LLM-BLENDER consists of two modules – PAIRRANKER and GENFUSER. These modules show that the optimal LLM for different examples can vary significantly. PAIRRANKER, the first module, has been developed to identify minute variations among potential outputs. It uses an advanced pairwise comparison technique in which the original text and two candidate outputs from various LLMs act as inputs. In order to jointly encode the input and the candidate pair, it makes use of cross-attention encoders like RoBERTa, where the quality of the two candidates can be determined by PAIRRANKER using this encoding. 

🚀 JOIN the fastest ML Subreddit Community

The second module, GENFUSER, focuses on merging the top-ranked candidates to generate an improved output. It makes the most of the advantages of the chosen candidates while minimizing their disadvantages. GENFUSER aims to develop an output that is superior to the output of any one LLM by merging the outputs of various LLMs.

For evaluation, the team has provided a benchmark dataset called MixInstruct, which incorporates Oracle pairwise comparisons and combines various instruction datasets. This dataset uses 11 popular open-source LLMs to generate multiple candidates for each input across various instruction-following tasks. It comprises training, validation, and test examples with Oracle comparisons for automatic evaluation. These oracle comparisons have been used to give candidate outputs a ground truth ranking, allowing the performance of LLM-BLENDER and other benchmark techniques to be assessed.

The experimental findings have shown that LLM-BLENDER performs much better across a range of evaluation parameters than individual LLMs and baseline techniques. It establishes a sizable performance gap and shows that employing the LLM-BLENDER ensembling methodology results in higher-quality output when compared to using a single LLM or baseline method. PAIRRANKER’s selections have outperformed individual LLM models because of their better performance in reference-based metrics and GPT-Rank. Through efficient fusion, GENFUSER significantly improves response quality by utilizing the top picks from PAIRRANKER. 

LLM-BLENDER has also outperformed individual LLMs, like Vicuna, and has thus shown great potential for improving LLM deployment and research through ensemble learning.


Check Out The Paper, Project, and Github. Don’t forget to join our 24k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]


Featured Tools From AI Tools Club

🚀 Check Out 100’s AI Tools in AI Tools Club


Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


➡️ Meet Notion: Your Wiki, Docs, & Projects Together

Credit: Source link

ShareTweetSendSharePin

Related Posts

The Ternus Era At Apple Begins, But Cook Isn’t Leaving
AI & Technology

The Ternus Era At Apple Begins, But Cook Isn’t Leaving

September 4, 2026
Mobile Games Designed to Be Addictive Get More Kid-Friendly
AI & Technology

Mobile Games Designed to Be Addictive Get More Kid-Friendly

September 4, 2026
Nvidia Makes .5 Billion Bet on MediaTek
AI & Technology

Nvidia Makes $3.5 Billion Bet on MediaTek

September 4, 2026
Nvidia Makes MediaTek Partnership Even Bigger, Huang Says
AI & Technology

Nvidia Makes MediaTek Partnership Even Bigger, Huang Says

September 4, 2026
Next Post
92% of US-based developers already using AI-powered coding tools at work: GitHub report 

92% of US-based developers already using AI-powered coding tools at work: GitHub report 

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Michigan reports two deaths linked to cyclosporiasis outbreak

Michigan reports two deaths linked to cyclosporiasis outbreak

August 30, 2026
Longtime New Yorker cultivates rooftop garden and community

Longtime New Yorker cultivates rooftop garden and community

August 29, 2026
He Definitely Needs a Prenup

He Definitely Needs a Prenup

September 2, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!