• bitcoinBitcoin(BTC)$76,712.00-0.66%
  • ethereumEthereum(ETH)$2,477.39-1.78%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$715.94-1.38%
  • rippleXRP(XRP)$1.34-1.79%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$99.94-1.67%
  • tronTRON(TRX)$0.339560-0.12%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.000.00%
  • zcashZcash(ZEC)$1,076.18-4.23%
  • HyperliquidHyperliquid(HYPE)$77.53-2.67%
  • dogecoinDogecoin(DOGE)$0.082299-2.77%
  • RainRain(RAIN)$0.015169-3.71%
  • USDSUSDS(USDS)$1.00-0.02%
  • moneroMonero(XMR)$521.58-3.14%
  • whitebitWhiteBIT Coin(WBT)$79.58-0.86%
  • chainlinkChainlink(LINK)$11.18-2.69%
  • leo-tokenLEO Token(LEO)$9.04-1.16%
  • cardanoCardano(ADA)$0.202657-1.97%
  • stellarStellar(XLM)$0.176823-1.67%
  • Ethena USDeEthena USDe(USDE)$1.00-0.03%
  • daiDai(DAI)$1.000.02%
  • bitcoin-cashBitcoin Cash(BCH)$220.65-2.16%
  • USD1USD1(USD1)$1.00-0.03%
  • litecoinLitecoin(LTC)$53.610.16%
  • uniswapUniswap(UNI)$6.14-3.14%
  • CantonCanton(CC)$0.094708-2.68%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.34-2.82%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.0748560.33%
  • avalanche-2Avalanche(AVAX)$7.30-1.14%
  • shiba-inuShiba Inu(SHIB)$0.000005-3.03%
  • nearNEAR Protocol(NEAR)$2.31-1.79%
  • suiSui(SUI)$0.70-2.99%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • crypto-com-chainCronos(CRO)$0.057119-4.33%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,334.14-0.36%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.14-3.05%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$112.04-1.73%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.02%
  • BittensorBittensor(TAO)$231.63-0.42%
  • BitwayBitway(BTW)$0.7231.45%
  • aaveAave(AAVE)$124.54-0.65%
  • pax-goldPAX Gold(PAXG)$4,338.19-0.37%
  • AsterAster(ASTER)$0.690.26%
  • mantleMantle(MNT)$0.56-1.48%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056648-1.39%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from AI2 and the University of Washington Uncover the Superficial Nature of Alignment in LLMs and Introduce URIAL: A Novel Tuning-Free Method

December 9, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Researchers from AI2 and the University of Washington Uncover the Superficial Nature of Alignment in LLMs and Introduce URIAL: A Novel Tuning-Free Method
ShareShareShareShareShare

Large Language Models (LLMs) are recent innovations in the field of Artificial Intelligence (AI) and Deep Learning. Some of the well-known LLMs, like GPT, PaLM, LLaMa, etc, have demonstrated incredible potential in generating content. From question answering and text summarization to language translation and code completion, these models can do a lot. These models, including ChatGPT, have gone through extensive pre-training on vast unsupervised text corpora. However, recent studies have suggested that the commonly adopted practice of fine-tuning may not be as essential as previously thought.

Alignment tuning, which is the process of improving base LLMs for usage as open-domain AI assistants, has been accepted as the industry standard. This includes Reinforcement Learning from Human Feedback (RLHF) and Supervised Fine-Tuning (SFT). This standard was questioned by a study called LIMA, which showed that as few as 1,000 samples for SFT may be sufficient to achieve meaningful alignment performance.

The Superficial Alignment Hypothesis, put forth by LIMA, proposed that alignment tuning, as opposed to radically changing basic LLMs’ behavior, may instead train them to choose particular data formats for user engagement. This showed that a few examples can produce high-quality, aligned models under supervised fine-tuning.

Since not enough research has been done to find solid support for the superficial alignment theory, a team of researchers from the Allen Institute for Artificial Intelligence and the University of Washington has addressed the widely used technique of alignment tuning in a recent paper to make basic LLMs into useful AI assistants for the open domain. Preference tuning has been accomplished through reinforcement learning from human feedback, and instruction learning has been accomplished through supervised fine-tuning.

The team has examined the shift in token distribution between base LLMs and their aligned counterparts, like Llama-2 and Llama-2-chat, in order to study the impact of alignment adjustment. They have found out that base LLMs and their aligned versions share the top-ranked tokens and perform nearly identically in decoding on most token positions. Discourse markers and safety disclaimers are examples of style tokens that experience the most distribution fluctuations. This study has provided compelling evidence for the hypothesis that alignment adjustment mostly concentrates on assimilating the linguistic style of AI assistants, with the base LLMs supplying the information required to respond to user inquiries.

The team has also presented a research topic in response to these findings: to what extent may base LLMs be aligned without SFT or RLHF? They have suggested URIAL (Untuned LLMs with Restyled In-context Alignment), an alignment technique that does not require tuning. With just three continual style examples and a system prompt, URIAL accomplishes effective alignment solely through in-context learning (ICL) with base LLMs. 

In a series of instances dubbed just-eval-instruct, the team has provided a detailed and comprehensible analysis that shows how base LLMs with URIAL can perform on par with or better than LLMs aligned with SFT (Mistral-7b-Instruct) or SFT+RLHF (Llama-2-70b-chat). The results have demonstrated that deliberate prompting and in-context learning can dramatically close the gap between tuning-free and tuning-based alignment strategies.

In conclusion, the evaluation results have highlighted shallow alignment tuning and have shown that it mostly entails adopting linguistic styles and depends on the preexisting knowledge of the basic LLMs.


Check out the Paper and Project. All credit for this research goes to the researchers of this project. Also, don’t forget to join our 33k+ ML SubReddit, 41k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

How To Fix iMessage “Not Delivered” Error On iPhones

How To Adjust The Liquid Glass Effect On Your iPhone With iOS 27

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


🐝 [FREE AI WEBINAR] ‘Beginners Guide to LangChain: Chat with Your Multi-Model Data’ Dec 11, 2023 10 am PST

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Fix iMessage “Not Delivered” Error On iPhones
AI & Technology

How To Fix iMessage “Not Delivered” Error On iPhones

September 13, 2026
How To Adjust The Liquid Glass Effect On Your iPhone With iOS 27
AI & Technology

How To Adjust The Liquid Glass Effect On Your iPhone With iOS 27

September 13, 2026
Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstruction
AI & Technology

Hierarchical NeRF with JAX3D for Volumetric Rendering, Novel-View Synthesis, and 3D Reconstruction

September 13, 2026
Car Manufacturers Are Ditching CarPlay In 2026: Here’s Why
AI & Technology

Car Manufacturers Are Ditching CarPlay In 2026: Here’s Why

September 13, 2026
Next Post
Rep. Henry Cuellar carjacked outside of his home in D.C.

Rep. Henry Cuellar carjacked outside of his home in D.C.

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas

SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas

September 8, 2026
CPI REPORT IS COMING: Do This Before Market Open!

CPI REPORT IS COMING: Do This Before Market Open!

September 11, 2026
Pros And Cons Of Using A Chromebook As Your Everyday PC

Pros And Cons Of Using A Chromebook As Your Everyday PC

September 8, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!