• bitcoinBitcoin(BTC)$77,362.00-0.57%
  • ethereumEthereum(ETH)$2,532.07-1.75%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$731.910.29%
  • rippleXRP(XRP)$1.37-0.33%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$102.12-0.15%
  • tronTRON(TRX)$0.3403081.22%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.01-2.26%
  • zcashZcash(ZEC)$1,141.13-3.71%
  • HyperliquidHyperliquid(HYPE)$80.30-1.91%
  • dogecoinDogecoin(DOGE)$0.085050-0.79%
  • RainRain(RAIN)$0.015046-4.75%
  • moneroMonero(XMR)$526.362.28%
  • USDSUSDS(USDS)$1.000.00%
  • whitebitWhiteBIT Coin(WBT)$80.41-0.84%
  • chainlinkChainlink(LINK)$11.54-2.08%
  • leo-tokenLEO Token(LEO)$9.11-0.43%
  • cardanoCardano(ADA)$0.208443-0.60%
  • stellarStellar(XLM)$0.181797-0.24%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • bitcoin-cashBitcoin Cash(BCH)$228.70-2.21%
  • daiDai(DAI)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$53.97-0.10%
  • uniswapUniswap(UNI)$6.383.56%
  • CantonCanton(CC)$0.097936-0.42%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.380.99%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.074780-0.96%
  • avalanche-2Avalanche(AVAX)$7.40-2.54%
  • nearNEAR Protocol(NEAR)$2.40-7.89%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.14%
  • suiSui(SUI)$0.73-1.97%
  • crypto-com-chainCronos(CRO)$0.0589833.21%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • MemeCoreMemeCore(M)$1.18-0.56%
  • tether-goldTether Gold(XAUT)$4,349.81-0.27%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • okbOKB(OKB)$113.29-0.79%
  • BittensorBittensor(TAO)$235.39-0.74%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.04%
  • aaveAave(AAVE)$125.92-0.47%
  • mantleMantle(MNT)$0.57-4.61%
  • pax-goldPAX Gold(PAXG)$4,355.37-0.23%
  • AsterAster(ASTER)$0.700.14%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.05946113.90%
  • polkadotPolkadot(DOT)$1.04-2.38%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Decoding the Impact of Feedback Protocols on Large Language Model Alignment: Insights from Ratings vs. Rankings

January 22, 2024
in AI & Technology
Reading Time: 7 mins read
A A
Decoding the Impact of Feedback Protocols on Large Language Model Alignment: Insights from Ratings vs. Rankings
ShareShareShareShareShare

Alignment has become a pivotal concern for the development of next-generation text-based assistants, particularly in ensuring that large language models (LLMs) align with human values. This alignment aims to enhance LLM-generated content’s accuracy, coherence, and harmlessness in response to user queries. The alignment process comprises three key elements: feedback acquisition, alignment algorithms, and model evaluation. While previous efforts focused on alignment algorithms, this study delves into the nuances of feedback acquisition, specifically comparing ratings and rankings protocols, shedding light on a significant consistency challenge.

In existing literature, alignment algorithms such as PPO, DPO, and PRO have been extensively explored under specific feedback protocols and evaluation setups. Meanwhile, feedback acquisition strategies have concentrated on developing fine-grained and dense protocols, which can be challenging and costly. This study analyzes the impact of two feedback protocols, ratings and rankings, on LLM alignment. Figure 1 provides an illustration of their pipeline. 

Understanding Feedback Protocols: Ratings vs. Rankings

Ratings involve assigning an absolute value to a response using a predefined scale, while rankings require annotators to select their preferred response from a pair. Ratings quantify response goodness but can be challenging for complex instructions, whereas rankings are easier for such instructions but lack quantification of the gap between responses (Listed in Table 1).

Now we will delve deeper into the initially announced feedback inconsistency problem. The authors make use of the observation that the ratings on a pair of responses for a given instruction can be compared to convert the ratings feedback data into its rankings form. This conversion of the ratings data DA to the rankings data DRA allows us a unique opportunity to study the interplay between the absolute feedback DA and relative feedback DR collected from the annotators, independently. Here, they define the term consistency as the agreement between the ratings (converted to its rankings form) and the rankings received by a pair of responses to a given instruction independent of the ratings data.

We can clearly observe consistency issues from Table 3 and 4 in both human and AI feedback data. Interestingly, the consistency score falls within a similar range of 40% − 42% for both humans and AI, suggesting that a substantial portion of the feedback data can yield contradictory preferences depending on the feedback protocol employed. This consistency problem underscores several critical points: (a) it indicates variations in the perceived quality of responses based on the choice of the feedback acquisition protocols, (b) it underscores that the alignment pipeline can vary significantly depending on whether ratings or rankings are used as sparse forms of feedback, and (c) it emphasizes the necessity of meticulous data curation when working with multiple feedback protocols for aligning LLMs. 

Exploring Feedback Inconsistency:

The study delves into the identified feedback inconsistency problem, leveraging an insightful observation. By comparing ratings on a pair of responses, the authors convert rating feedback data (DA) into rankings data (DRA). This conversion offers a unique opportunity to independently study the interplay between absolute feedback (DA) and relative feedback (DR) from annotators. Consistency, defined as the agreement between converted ratings and original rankings, is assessed. Notably, Tables 3 and 4 reveal consistent issues in both human and AI feedback, with a noteworthy consistency score range of 40%−42%. This underscores variations in perceived response quality based on feedback acquisition protocols, highlighting the significant impact on the alignment pipeline and emphasizing the need for meticulous data curation when handling diverse feedback protocols in aligning LLMs.

Feedback Data Acquisition

The study uses diverse instructions from sources like Dolly, Self-Instruct, and Super-NI to collect feedback. Alpaca-7B serves as the base LLM, generating candidate responses for evaluation. The authors leverage GPT-3.5-Turbo for large-scale ratings and rankings feedback data collection. They also collect feedback data under the ratings and rankings protocols. 

Analysis of rating distribution (shown in Figure 2) indicates human annotators tend to give higher scores, while AI feedback is more balanced. The study also ensures feedback data is unbiased towards longer or unique responses. Agreement analysis (shown in Table 2) between human-human and human-AI feedback shows reasonable alignment rates. In summary, the agreement results indicate that GPT-3.5-Turbo can provide ratings and rankings feedback close to the human’s gold label for the responses to the instructions in our dataset.

Impact on Alignment and Model Evaluation

The study trains reward models based on ratings and rankings feedback and assesses Best-of-n policies. Evaluation on unseen instructions reveals Best-of-n policies, especially with rankings feedback, outperform the base LLM (SFT) and demonstrate improvement in alignment (shown in Figure 3). 

A surprising revelation in the study unveils an evaluation inconsistency phenomenon, where the feedback protocol choice during evaluation seems to favor the alignment algorithm that aligns with the same feedback protocol. Notably, the gap in win rates between the Best-of-n (rankings) policy and the SFT is more pronounced (11.2%) than the gap observed between the Best-of-n (ratings) policy and SFT (5.3%) under the rankings protocol. Conversely, under the ratings protocol, the gap between the Best-of-n (ratings) policy and SFT (5%) slightly outweighs the gap between the Best-of-n (rankings) policy and SFT (4.3%). This inconsistency extends to evaluations involving GPT-3.5-Turbo, indicating a nuanced perception of policy response quality by annotators (both human and AI) under distinct feedback protocols. These findings underscore the substantial implications for practitioners, highlighting that the feedback acquisition protocol significantly influences each stage of the alignment pipeline.

In conclusion, The study underscores the paramount importance of meticulous data curation within sparse feedback protocols, shedding light on the potential repercussions of feedback protocol choices on evaluation outcomes. In the pursuit of model alignment, future research avenues may delve into the cognitive aspects of the identified consistency problem, aiming to enhance alignment strategies. Exploring richer forms of feedback beyond the scope of absolute and relative preferences is crucial for a more comprehensive understanding and improved alignment in diverse application domains. Despite its valuable insights, the study acknowledges limitations, including its focus on specific types of feedback, potential subjectivity in human annotations, and the necessity to explore the impact on different demographic groups and specialized domains. Addressing these limitations will contribute to developing more robust and universally applicable alignment methodologies in the evolving landscape of artificial intelligence.


Check out the Paper and Github. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. Join our 36k+ ML SubReddit, 41k+ Facebook Community, Discord Channel, and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our Telegram Channel


YOU MAY ALSO LIKE

Amodei Calls for Slowing the Pace of AI Capability Improvement – Unite.AI

Are You Using The Right Ethernet Port On Your Router? Here’s How To Know

Vineet Kumar is a consulting intern at MarktechPost. He is currently pursuing his BS from the Indian Institute of Technology(IIT), Kanpur. He is a Machine Learning enthusiast. He is passionate about research and the latest advancements in Deep Learning, Computer Vision, and related fields.


🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

Amodei Calls for Slowing the Pace of AI Capability Improvement – Unite.AI
AI & Technology

Amodei Calls for Slowing the Pace of AI Capability Improvement – Unite.AI

September 12, 2026
Are You Using The Right Ethernet Port On Your Router? Here’s How To Know
AI & Technology

Are You Using The Right Ethernet Port On Your Router? Here’s How To Know

September 12, 2026
Is A 256GB SSD Better Than A 1TB Hard Drive? It Depends How You’re Using It
AI & Technology

Is A 256GB SSD Better Than A 1TB Hard Drive? It Depends How You’re Using It

September 12, 2026
What Is Benchmark Saturation? Why Yesterday’s AI Tests Stop Working – Unite.AI
AI & Technology

What Is Benchmark Saturation? Why Yesterday’s AI Tests Stop Working – Unite.AI

September 12, 2026
Next Post
‘Flu Shot Cheerleader’ speaks out years after stoking anti-vaccine movement

‘Flu Shot Cheerleader’ speaks out years after stoking anti-vaccine movement

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Cat is caught smuggling drugs into a Russian prison

Cat is caught smuggling drugs into a Russian prison

September 6, 2026
Australian mother welcomes identical quadruplet girls

Australian mother welcomes identical quadruplet girls

September 7, 2026
Trial date for Maduro and his wife set for next June

Trial date for Maduro and his wife set for next June

September 7, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!