• bitcoinBitcoin(BTC)$76,741.00-0.76%
  • ethereumEthereum(ETH)$2,479.26-2.22%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$716.07-2.70%
  • rippleXRP(XRP)$1.34-2.17%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$99.72-2.25%
  • tronTRON(TRX)$0.3409780.04%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-1.59%
  • zcashZcash(ZEC)$1,095.61-4.71%
  • HyperliquidHyperliquid(HYPE)$77.40-3.61%
  • dogecoinDogecoin(DOGE)$0.083500-1.85%
  • RainRain(RAIN)$0.0153381.49%
  • moneroMonero(XMR)$538.110.88%
  • USDSUSDS(USDS)$1.00-0.01%
  • whitebitWhiteBIT Coin(WBT)$79.59-0.99%
  • chainlinkChainlink(LINK)$11.29-2.28%
  • leo-tokenLEO Token(LEO)$9.06-0.60%
  • cardanoCardano(ADA)$0.204574-2.23%
  • stellarStellar(XLM)$0.178346-2.15%
  • Ethena USDeEthena USDe(USDE)$1.00-0.02%
  • daiDai(DAI)$1.000.01%
  • bitcoin-cashBitcoin Cash(BCH)$223.38-3.29%
  • USD1USD1(USD1)$1.00-0.02%
  • litecoinLitecoin(LTC)$53.78-0.51%
  • uniswapUniswap(UNI)$6.27-1.50%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.35-1.89%
  • CantonCanton(CC)$0.094908-4.09%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.0752190.72%
  • avalanche-2Avalanche(AVAX)$7.32-1.69%
  • shiba-inuShiba Inu(SHIB)$0.000005-3.20%
  • nearNEAR Protocol(NEAR)$2.30-2.81%
  • suiSui(SUI)$0.71-2.52%
  • crypto-com-chainCronos(CRO)$0.0588291.36%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,346.77-0.04%
  • MemeCoreMemeCore(M)$1.15-2.38%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$113.49-0.68%
  • BittensorBittensor(TAO)$232.50-1.70%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.06%
  • aaveAave(AAVE)$124.15-2.92%
  • pax-goldPAX Gold(PAXG)$4,351.93-0.04%
  • AsterAster(ASTER)$0.690.70%
  • mantleMantle(MNT)$0.56-2.86%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0576301.03%
  • BitwayBitway(BTW)$0.6719.29%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Google Deepmind proposes ‘self-discover’ framework for LLMs, improves GPT-4 performance

February 8, 2024
in AI & Technology
Reading Time: 4 mins read
A A
Google Deepmind proposes ‘self-discover’ framework for LLMs, improves GPT-4 performance
ShareShareShareShareShare

In a bid to enhance the reasoning capabilities of large language models (LLMs), researchers from Google Deepmind and University of Southern California have proposed a new ‘self-discover’ prompting framework.

Published on arXiV and Hugging Face this morning, the approach goes beyond existing prompting techniques used by LLMs and has been found capable of improving the performance of known models out there, including OpenAI’s GPT-4 and Google’s PaLM 2. 

YOU MAY ALSO LIKE

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents

Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference

“Self-discover substantially improves GPT-4 and PaLM 2’s performance on challenging reasoning benchmarks such as BigBench-Hard, grounded agent reasoning and MATH by as much as 32% compared to Chain of Thought (CoT),” the researchers write in the paper.

The framework revolves around LLMs self-discovering task-intrinsic reasoning structures to solve a problem. The models look at multiple atomic reasoning modules, such as critical thinking and step-by-step thinking, and compose them into an explicit reasoning structure for LLMs to follow during decoding. 

VB Event

The AI Impact Tour – NYC

We’ll be in New York on February 29 in partnership with Microsoft to discuss how to balance risks and rewards of AI applications. Request an invite to the exclusive event below.

 

Request an invite

More interestingly, this approach works with 10 to 40 times less inference compute — something that can be great for enterprises.

Self-discovering unique structures

LLMs have evolved to handle numerous tasks, thanks to their ability to follow instructions, reason and generate coherent responses. To make this happen, the models, powered by transformer architecture, use various prompting techniques inspired by cognitive theories of how humans reason and solve problems. This includes few-shot and zero-shot chain-of-thought, inspired by how we solve a problem step-by-step, decomposition prompting of how we break a problem into multiple subproblems and step-back prompting of how we reflect on the nature of a task to establish general principles. 

While all these methods, most notably chain-of-thought, do the job, they all work by making an implicit prior assumption of how to tackle a given task. This approach, the researchers argue, may not be the best as each task has a unique intrinsic structure and one particular technique may be better at solving it than the other.

With the latest research, Deepmind and USC researchers have proposed a general prompting framework that self-discovers this unique underlying structure to pick the right reasoning technique for the task while also being efficient at the same time.

“Self-discover is inspired by how humans internally devise a reasoning program for problem-solving. From a set of atomic reasoning modules described in natural language such as ‘break down into sub-tasks’ and ‘critical thinking’, an LLM, and task examples without labels, it composes a coherent reasoning structure intrinsic to the task (Stage1) and then solves instances of the task using the discovered structure (Stage2). Stage 1 operates at the task level and uses three actions to guide the LLM to generate a reasoning structure for the task. At Stage 2, during the final decoding, the LLM simply follows the self-discovered structure to arrive at the final answer,” the researchers explain.

Notable performance improvements for known LLMs

To see how the new approach works, the researchers tested it with multiple models – including GPT-4 and PaLM 2-L, on 25 reasoning tasks, including Big-Bench Hard, Thinking for Doing and Math. In 21 out of 25 tasks, self-discover was found to outperform chain-of-thought reasoning and other techniques with performance gains of up to 32%. The researchers also found that it did better in terms of efficiency by requiring 10 to 40 times less inference compute.

According to the data shared in the paper, when working with GPT-4, the self-discover approach achieved results with an accuracy of 81%, 85% and 73% across Big-Bench Hard, Thinking for Doing and Math tasks, respectively. However, when working with chain-of-thought, the results dropped to 75%, 52% and 71%, respectively. A nearly similar gap was noted when it was compared with the plan-and-solve approach.

On the other hand, PaLM 2-L achieved results with an accuracy of 67%, 69% and 50.5% across the three tasks. This is lower than that of GPT-4 but still much better than what was achieved with chain-of-thought (60%, 40% and 42%) and plan-and-solve (61%, 42% and 49%) approaches.

Improved reasoning is key to AI success

While the idea of a self-discover prompting framework has just been proposed, it has the potential to push the boundary of problem-solving and give LLMs the ability to address challenging problems with ease – ultimately moving toward the goal of general intelligence. Notably, the transferability studies conducted by the researchers show that the composed reasoning structures are universally applicable across model families and share commonalities with human reasoning patterns.

“Forward looking, we are excited to explore more on LLM structured reasoning to push the boundary of problem-solving and discover potentials for Human-AI collaboration,” the team added.

VentureBeat’s mission is to be a digital town square for technical decision-makers to gain knowledge about transformative enterprise technology and transact. Discover our Briefings.

Credit: Source link

ShareTweetSendSharePin

Related Posts

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents
AI & Technology

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents

September 13, 2026
Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference
AI & Technology

Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference

September 13, 2026
Why Do Routers Have So Many Antennas?
AI & Technology

Why Do Routers Have So Many Antennas?

September 13, 2026
Hyundai Motor Group Puts Data Flywheel Into Full Operation – Unite.AI
AI & Technology

Hyundai Motor Group Puts Data Flywheel Into Full Operation – Unite.AI

September 13, 2026
Next Post
Why Marjorie Taylor Greene was ‘kicked out’ of the Freedom Caucus according to Rep. Buck

Why Marjorie Taylor Greene was 'kicked out' of the Freedom Caucus according to Rep. Buck

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Lease End Review – Online Lease Buyouts That Cost You Nothing to Arrange

Lease End Review – Online Lease Buyouts That Cost You Nothing to Arrange

September 11, 2026
Stay Tuned NOW Streaming Behind The Scenes! – Jul 21

Stay Tuned NOW Streaming Behind The Scenes! – Jul 21

September 7, 2026
NBC Nightly News with Tom Llamas Full Episode – July 22

NBC Nightly News with Tom Llamas Full Episode – July 22

September 6, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!