• bitcoinBitcoin(BTC)$78,458.00-0.70%
  • ethereumEthereum(ETH)$2,483.67-0.03%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$752.321.87%
  • rippleXRP(XRP)$1.421.68%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$103.35-0.30%
  • tronTRON(TRX)$0.3389031.35%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.00%
  • zcashZcash(ZEC)$1,182.224.08%
  • HyperliquidHyperliquid(HYPE)$84.83-0.30%
  • dogecoinDogecoin(DOGE)$0.089949-0.56%
  • RainRain(RAIN)$0.016223-0.40%
  • USDSUSDS(USDS)$1.000.02%
  • whitebitWhiteBIT Coin(WBT)$81.256.28%
  • moneroMonero(XMR)$504.76-2.05%
  • chainlinkChainlink(LINK)$12.50-1.73%
  • leo-tokenLEO Token(LEO)$9.20-0.12%
  • cardanoCardano(ADA)$0.2198020.21%
  • stellarStellar(XLM)$0.187945-2.51%
  • bitcoin-cashBitcoin Cash(BCH)$258.05-0.05%
  • daiDai(DAI)$1.00-0.01%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • USD1USD1(USD1)$1.000.00%
  • CantonCanton(CC)$0.1075822.54%
  • litecoinLitecoin(LTC)$54.35-1.61%
  • uniswapUniswap(UNI)$6.75-1.52%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.400.86%
  • hedera-hashgraphHedera(HBAR)$0.079245-3.14%
  • avalanche-2Avalanche(AVAX)$7.99-1.08%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • suiSui(SUI)$0.81-0.55%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.27%
  • nearNEAR Protocol(NEAR)$2.310.17%
  • crypto-com-chainCronos(CRO)$0.0588774.12%
  • paypal-usdPayPal USD(PYUSD)$1.000.02%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • MemeCoreMemeCore(M)$1.236.89%
  • tether-goldTether Gold(XAUT)$4,354.04-1.21%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$260.580.98%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • okbOKB(OKB)$113.80-1.78%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.11%
  • polkadotPolkadot(DOT)$1.2417.80%
  • mantleMantle(MNT)$0.631.96%
  • AsterAster(ASTER)$0.75-2.74%
  • aaveAave(AAVE)$128.90-2.08%
  • pax-goldPAX Gold(PAXG)$4,356.92-1.22%
  • OndoOndo(ONDO)$0.374420-2.03%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from ETH Zurich and Microsoft Introduce SCREWS: An Artificial Intelligence Framework for Enhancing the Reasoning in Large Language Models

October 5, 2023
in AI & Technology
Reading Time: 5 mins read
A A
Researchers from ETH Zurich and Microsoft Introduce SCREWS: An Artificial Intelligence Framework for Enhancing the Reasoning in Large Language Models
ShareShareShareShareShare

Large Language Models (LLMs) have succeeded in several different reasoning tasks. To guarantee that the intended aim is met, it is sometimes required to iteratively adjust the LLM results because the output is only occasionally accurate on the first try. These refinement techniques assume that consecutive results (from the same model, an external model, or some tool) result in improved performance. However, there is no assurance that later versions will always be better as Figure 1 shows, refining might result in a false positive. This encourages the model to choose an earlier outcome using the selection technique. Furthermore, prior research on iterative refining frequently uses a single, fixed reasoning technique. But humans are more adaptable. 

Figure 1: A case study illustrative of how Conditional Resampling (also known as “refinement”) may result in improper modification of the initial response. The original response, which in this case is the right one, may be chosen by a selection module in place of the alteration.

A product manager may use a brainstorming technique to generate several ideas before switching to a prioritization technique to rank them according to their viability or effect. Similarly, a student preparing for an exam might use deductive reasoning to answer issues and inductive reasoning to confirm the results. They thus suggest a modular strategy for answering refinements, enabling us to try various tactics. In this paper, researchers from  ETH Zurich and Microsoft Semantic Machines present SCREWS, a modular framework for reasoning about changes. Sampling, Conditional Resampling, and Selection are the three core components of the architecture that are introduced in detail in Figure 2. They instantiate SCREWS by fixing the submodules for each module (for example, they may choose “Chain of Thought” for Sampling). This is done for a specific job and input sequence. 

Figure 2 presents a high-level picture of the modular SCREWS system for reasoning about revisions. The three substantial boxes (or “modules”) each contain a number of choices (or “submodules”). Many previous efforts, including Self-Refine, Least to Most, LLMs Know (Mostly), Self-Consistency, Self-Improve, PHP CoT, Self-Correct, Socratic CoT, Programme of Thoughts, and many more, may be seen as examples of the framework. (…) denotes additional sub-components that may be added to each module, including, but not limited to, cached memory or online search for the Sampling module, a fine-tuned model or an external verifier for Conditional Resampling, and selection based on humans or an oracle for the Selection module.

Sampling’s first outputs are handed on to Conditional Resampling, which determines whether to create a revision based on the original sample and does so if necessary. The Selection module then chooses the best from all the samples and revisions. Given the modular design of their framework, additional framework elements can be used to enhance several newly suggested self-refining approaches. One example is the combination of their model-based selection technique and self-refinement method, which can improve overall performance. They use ChatGPT or GPT-4 to assess SCREWS on various reasoning tasks, including multi-hop question answering, arithmetic reasoning, and code debugging. 

Compared to the standard sample and resampling procedures, their suggested solutions produce significant improvements (10–15%). They show the value of heterogeneous resampling, showing how it may influence the model’s logic and substantially improve the baselines at a very low total cost. They also explain the significance of a model-based selection approach, a crucial element of contemporary LLMs that enables the model to revert to earlier, more certain outputs.


Check out the Paper. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 31k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction

SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


Credit: Source link

ShareTweetSendSharePin

Related Posts

Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction
AI & Technology

Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction

September 8, 2026
SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas
AI & Technology

SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas

September 8, 2026
What Is Roku’s Secret Menu And How Do You Unlock It?
AI & Technology

What Is Roku’s Secret Menu And How Do You Unlock It?

September 8, 2026
NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels
AI & Technology

NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels

September 8, 2026
Next Post
Medicare Fraud Crack Down

Medicare Fraud Crack Down

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Extended Interview: Gov. Kathy Hochul on N.Y.’s Data Center Moratorium

Extended Interview: Gov. Kathy Hochul on N.Y.’s Data Center Moratorium

September 6, 2026
These 2 Stocks Are About To Report Earnings This Week!

These 2 Stocks Are About To Report Earnings This Week!

September 8, 2026
Peak Coal, Postponed Again | Seeking Alpha

Peak Coal, Postponed Again | Seeking Alpha

September 7, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!