• bitcoinBitcoin(BTC)$80,352.00-1.12%
  • ethereumEthereum(ETH)$2,574.61-2.48%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$749.38-2.26%
  • rippleXRP(XRP)$1.38-2.47%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$108.03-3.42%
  • tronTRON(TRX)$0.3421371.38%
  • zcashZcash(ZEC)$1,437.43-8.48%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.02-1.32%
  • HyperliquidHyperliquid(HYPE)$90.75-1.26%
  • dogecoinDogecoin(DOGE)$0.084641-3.03%
  • moneroMonero(XMR)$522.78-9.78%
  • whitebitWhiteBIT Coin(WBT)$81.74-1.82%
  • USDSUSDS(USDS)$1.000.00%
  • RainRain(RAIN)$0.013220-5.18%
  • chainlinkChainlink(LINK)$11.96-4.40%
  • cardanoCardano(ADA)$0.219148-2.04%
  • leo-tokenLEO Token(LEO)$8.920.28%
  • stellarStellar(XLM)$0.189307-1.46%
  • uniswapUniswap(UNI)$8.78-4.60%
  • bitcoin-cashBitcoin Cash(BCH)$245.52-1.43%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • nearNEAR Protocol(NEAR)$3.53-3.75%
  • daiDai(DAI)$1.000.02%
  • litecoinLitecoin(LTC)$56.85-0.63%
  • USD1USD1(USD1)$1.00-0.02%
  • avalanche-2Avalanche(AVAX)$9.768.80%
  • CantonCanton(CC)$0.103569-6.13%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.370.07%
  • hedera-hashgraphHedera(HBAR)$0.0809371.75%
  • suiSui(SUI)$0.81-2.40%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-1.36%
  • MemeCoreMemeCore(M)$1.353.83%
  • crypto-com-chainCronos(CRO)$0.057700-2.92%
  • BittensorBittensor(TAO)$250.87-5.03%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • tether-goldTether Gold(XAUT)$4,372.06-0.02%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • okbOKB(OKB)$115.56-1.00%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.26%
  • AkedoAkedo(AKE)$0.09672054.22%
  • aaveAave(AAVE)$135.38-6.33%
  • EthenaEthena(ENA)$0.2014685.41%
  • AsterAster(ASTER)$0.73-3.49%
  • OndoOndo(ONDO)$0.405587-0.14%
  • mantleMantle(MNT)$0.59-3.01%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Trust-Align: An AI Framework for Improving the Trustworthiness of Retrieval-Augmented Generation in Large Language Models

September 23, 2024
in AI & Technology
Reading Time: 5 mins read
A A
Trust-Align: An AI Framework for Improving the Trustworthiness of Retrieval-Augmented Generation in Large Language Models
ShareShareShareShareShare

Large language models (LLMs) have gained significant attention due to their potential to enhance various artificial intelligence applications, particularly in natural language processing. When integrated into frameworks like Retrieval-Augmented Generation (RAG), these models aim to refine AI systems’ output by drawing information from external documents rather than relying solely on their internal knowledge base. This approach is crucial in ensuring that AI-generated content remains factually accurate, which is a persistent issue in models not tied to external sources.

A key problem faced in this area is the occurrence of hallucinations in LLMs—where models generate seemingly plausible but factually incorrect information. This becomes especially problematic in tasks requiring high accuracy, such as answering factual questions or assisting in legal and educational fields. Many state-of-the-art LLMs rely heavily on parametric knowledge information learned during training, making them unsuitable for tasks where responses must strictly come from specific documents. To tackle this issue, new methods must be introduced to evaluate and improve the trustworthiness of these models.

YOU MAY ALSO LIKE

Anthropic’s Existential Risk Warnings Hijack Larger AI Debate

Anthropic Investor Franklin: AI Safety Concerns Won’t Slow Spending

Traditional methods focus on evaluating the end results of LLMs within the RAG framework, but few explore the intrinsic trustworthiness of the models themselves. Currently, approaches like prompting techniques align the models’ responses with document-grounded information. However, these methods often fall short, either failing to adapt the models or resulting in overly sensitive outputs that respond inappropriately. Researchers identified the need for a new metric to measure LLM performance and ensure that the models provide grounded, trustworthy responses based solely on retrieved documents.

Researchers from the Singapore University of Technology and Design, in collaboration with DSO National Laboratories, introduced a novel framework called “TRUST-ALIGN.” This method focuses on enhancing the trustworthiness of LLMs in RAG tasks by aligning their outputs to provide more accurate, document-supported answers. The researchers also developed a new evaluation metric, TRUST-SCORE, which assesses models based on multiple dimensions, such as their ability to determine whether a question can be answered using the provided documents and their precision in citing relevant sources.

TRUST-ALIGN works by fine-tuning LLMs using a dataset containing 19,000 question-document pairs, each labeled with preferred and unpreferred responses. This dataset was created by synthesizing natural responses from LLMs like GPT-4 and negative responses derived from common hallucinations. The key advantage of this method lies in its ability to directly optimize LLM behavior toward providing grounded refusals when necessary, ensuring that models only answer questions when sufficient information is available. It improves the models’ citation accuracy by guiding them to reference the most relevant portions of the documents, thus preventing over-citation or improper attribution.

Regarding performance, the introduction of TRUST-ALIGN showed substantial improvements across several benchmark datasets. For example, when evaluated on the ASQA dataset, LLaMA-3-8b, aligned with TRUST-ALIGN, achieved a 10.73% increase in the TRUST-SCORE, surpassing models like GPT-4 and Claude-3.5 Sonnet. On the QAMPARI dataset, the method outperformed the baseline models by 29.24%, while the ELI5 dataset showed a performance boost of 14.88%. These figures demonstrate the effectiveness of the TRUST-ALIGN framework in generating more accurate and reliable responses compared to other methods.

One of the significant improvements brought by TRUST-ALIGN was in the models’ ability to refuse to answer when the available documents were insufficient correctly. On ASQA, the refusal metric improved by 9.87%, while on QAMPARI, it showed an even higher increase of 22.53%. The ability to refuse was further highlighted in ELI5, where the improvement reached 5.32%. These results indicate that the framework enhanced the models’ accuracy and significantly reduced their tendency to over-answer questions without proper justification from the provided documents.

Another noteworthy achievement of TRUST-ALIGN was in improving citation quality. On ASQA, the citation precision scores rose by 26.67%, while on QAMPARI, citation recall increased by 31.96%. The ELI5 dataset also showed an improvement of 29.30%. This improvement in citation groundedness ensures that the models provide well-supported answers, making them more trustworthy for users who rely on fact-based systems.

In conclusion, this research addresses a critical issue in deploying large language models in real-world applications. By developing TRUST-SCORE and the TRUST-ALIGN framework, researchers have created a reliable method to align LLMs toward generating document-grounded responses, minimizing hallucinations, and improving overall trustworthiness. This advancement is particularly significant in fields where accuracy and the ability to provide well-cited information are paramount, paving the way for more reliable AI systems in the future.


Check out the Paper and GitHub page. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter and join our Telegram Channel and LinkedIn Group. If you like our work, you will love our newsletter..

Don’t Forget to join our 50k+ ML SubReddit

⏩ ⏩ FREE AI WEBINAR: ‘SAM 2 for Video: How to Fine-tune On Your Data’ (Wed, Sep 25, 4:00 AM – 4:45 AM EST)


Nikhil is an intern consultant at Marktechpost. He is pursuing an integrated dual degree in Materials at the Indian Institute of Technology, Kharagpur. Nikhil is an AI/ML enthusiast who is always researching applications in fields like biomaterials and biomedical science. With a strong background in Material Science, he is exploring new advancements and creating opportunities to contribute.

⏩ ⏩ FREE AI WEBINAR: ‘SAM 2 for Video: How to Fine-tune On Your Data’ (Wed, Sep 25, 4:00 AM – 4:45 AM EST)


Credit: Source link

ShareTweetSendSharePin

Related Posts

Anthropic’s Existential Risk Warnings Hijack Larger AI Debate
AI & Technology

Anthropic’s Existential Risk Warnings Hijack Larger AI Debate

September 20, 2026
Anthropic Investor Franklin: AI Safety Concerns Won’t Slow Spending
AI & Technology

Anthropic Investor Franklin: AI Safety Concerns Won’t Slow Spending

September 20, 2026
Former FTC Technologist Warns Against an AI ‘Cartel’
AI & Technology

Former FTC Technologist Warns Against an AI ‘Cartel’

September 20, 2026
Crusoe CEO: Data Center Indusry Has a ‘Marketing Issue’
AI & Technology

Crusoe CEO: Data Center Indusry Has a ‘Marketing Issue’

September 20, 2026
Next Post
Ukrainian F-16 crashes during Russian missile attack

Ukrainian F-16 crashes during Russian missile attack

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Inside the aftermath of the Amazon plane crash in Miami

Inside the aftermath of the Amazon plane crash in Miami

September 16, 2026
Sanlam Limited 2026 Q2 – Results – Earnings Call Presentation (OTCMKTS:SLLDY) 2026-09-14

Sanlam Limited 2026 Q2 – Results – Earnings Call Presentation (OTCMKTS:SLLDY) 2026-09-14

September 14, 2026
Dire warning that A.I. could end humanity

Dire warning that A.I. could end humanity

September 14, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!