• bitcoinBitcoin(BTC)$76,695.00-0.80%
  • ethereumEthereum(ETH)$2,475.77-2.34%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$715.14-2.87%
  • rippleXRP(XRP)$1.34-2.38%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$99.56-2.59%
  • tronTRON(TRX)$0.340691-0.09%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-1.59%
  • zcashZcash(ZEC)$1,088.90-6.20%
  • HyperliquidHyperliquid(HYPE)$77.41-3.60%
  • dogecoinDogecoin(DOGE)$0.083300-2.08%
  • RainRain(RAIN)$0.0152610.83%
  • moneroMonero(XMR)$530.510.38%
  • USDSUSDS(USDS)$1.000.00%
  • whitebitWhiteBIT Coin(WBT)$79.53-1.05%
  • chainlinkChainlink(LINK)$11.25-2.65%
  • leo-tokenLEO Token(LEO)$9.05-0.64%
  • cardanoCardano(ADA)$0.204309-2.10%
  • stellarStellar(XLM)$0.178157-1.96%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • daiDai(DAI)$1.000.01%
  • bitcoin-cashBitcoin Cash(BCH)$223.48-3.34%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$53.49-0.94%
  • uniswapUniswap(UNI)$6.26-2.96%
  • CantonCanton(CC)$0.095115-3.40%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.35-2.37%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.0751820.64%
  • avalanche-2Avalanche(AVAX)$7.30-1.97%
  • shiba-inuShiba Inu(SHIB)$0.000005-2.60%
  • nearNEAR Protocol(NEAR)$2.31-2.53%
  • suiSui(SUI)$0.71-2.70%
  • crypto-com-chainCronos(CRO)$0.0585840.45%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,346.32-0.02%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.14-2.65%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$112.72-1.24%
  • BittensorBittensor(TAO)$233.78-1.13%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.00%
  • aaveAave(AAVE)$125.36-1.35%
  • pax-goldPAX Gold(PAXG)$4,350.50-0.04%
  • AsterAster(ASTER)$0.691.00%
  • mantleMantle(MNT)$0.56-1.67%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0575070.97%
  • polkadotPolkadot(DOT)$1.01-3.51%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Unlearning Copyrighted Data From a Trained LLM – Is It Possible?

January 23, 2024
in AI & Technology
Reading Time: 5 mins read
A A
Unlearning Copyrighted Data From a Trained LLM – Is It Possible?
ShareShareShareShareShare

In the domains of artificial intelligence (AI) and machine learning (ML), large language models (LLMs) showcase both achievements and challenges. Trained on vast textual datasets, LLM models encapsulate human language and knowledge.

Yet their ability to absorb and mimic human understanding presents legal, ethical, and technological challenges. Moreover, the massive datasets powering LLMs may harbor toxic material, copyrighted texts, inaccuracies, or personal data.

YOU MAY ALSO LIKE

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents

Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference

Making LLMs forget selected data has become a pressing issue to ensure legal compliance and ethical responsibility.

Let’s explore the concept of making LLMs unlearn copyrighted data to address a fundamental question: Is it possible?

Why is LLM Unlearning Needed?

LLMs often contain disputed data, including copyrighted data. Having such data in LLMs poses legal challenges related to private information, biased information, copyright data, and false or harmful elements.

Hence, unlearning is essential to guarantee that LLMs adhere to privacy regulations and comply with copyright laws, promoting responsible and ethical LLMs.

However, extracting copyrighted content from the vast knowledge these models have acquired is challenging. Here are some unlearning techniques that can help address this problem:

  • Data filtering: It involves systematically identifying and removing copyrighted elements, noisy or biased data, from the model’s training data. However, filtering can lead to the potential loss of valuable non-copyrighted information during the filtering process.
  • Gradient methods: These methods adjust the model’s parameters based on the loss function’s gradient, addressing the copyrighted data issue in ML models. However, adjustments may adversely affect the model’s overall performance on non-copyrighted data.
  • In-context unlearning: This technique efficiently eliminates the impact of specific training points on the model by updating its parameters without affecting unrelated knowledge. However, the method faces limitations in achieving precise unlearning, especially with large models, and its effectiveness requires further evaluation.

These techniques are resource-intensive and time-consuming, making them difficult to implement.

Case Studies

To understand the significance of LLM unlearning, these real-world cases highlight how companies are swarming with legal challenges concerning large language models (LLMs) and copyrighted data.

OpenAI Lawsuits: OpenAI, a prominent AI company, has been hit by numerous lawsuits over LLMs’ training data. These legal actions question the utilization of copyrighted material in LLM training. Also, they have triggered inquiries into the mechanisms models employ to secure permission for each copyrighted work integrated into their training process.

Sarah Silverman Lawsuit: The Sarah Silverman case involves an allegation that the ChatGPT model generated summaries of her books without authorization. This legal action underscores the important issues regarding the future of AI and copyrighted data.

Updating legal frameworks to align with technological progress ensures responsible and legal utilization of AI models. Moreover, the research community must address these challenges comprehensively to make LLMs ethical and fair.

Traditional LLM Unlearning Techniques

LLM unlearning is like separating specific ingredients from a complex recipe, ensuring that only the desired components contribute to the final dish. Traditional LLM unlearning techniques, like fine-tuning with curated data and re-training, lack straightforward mechanisms for removing copyrighted data.

Their broad-brush approach often proves inefficient and resource-intensive for the sophisticated task of selective unlearning as they require extensive retraining.

While these traditional methods can adjust the model’s parameters, they struggle to precisely target copyrighted content, risking unintentional data loss and suboptimal compliance.

Consequently, the limitations of traditional techniques and robust solutions require experimentation with alternative unlearning techniques.

Novel Technique: Unlearning a Subset of Training Data

The Microsoft research paper introduces a groundbreaking technique for unlearning copyrighted data in LLMs. Focusing on the example of the Llama2-7b model and Harry Potter books, the method involves three core components to make LLM forget the world of Harry Potter. These components include:

  • Reinforced model identification: Creating a reinforced model involves fine-tuning target data (e.g., Harry Potter) to strengthen its knowledge of the content to be unlearned.
  • Replacing idiosyncratic expressions: Unique Harry Potter expressions in the target data are replaced with generic ones, facilitating a more generalized understanding.
  • Fine-tuning on alternative predictions: The baseline model undergoes fine-tuning based on these alternative predictions. Basically, it effectively deletes the original text from its memory when confronted with relevant context.

Although the Microsoft technique is in the early stage and may have limitations, it represents a promising advancement toward more powerful, ethical, and adaptable LLMs.

The Outcome of The Novel Technique

The innovative method to make LLMs forget copyrighted data presented in the Microsoft research paper is a step toward responsible and ethical models.

The novel technique involves erasing Harry Potter-related content from Meta’s Llama2-7b model, known to have been trained on the “books3” dataset containing copyrighted works. Notably, the model’s original responses demonstrated an intricate understanding of J.K. Rowling’s universe, even with generic prompts.

However, Microsoft’s proposed technique significantly transformed its responses. Here are examples of prompts showcasing the notable differences between the original Llama2-7b model and the fine-tuned version.

Fine-tuned Prompt Comparison with Baseline

Image source 

This table illustrates that the fine-tuned unlearning models maintain their performance across different benchmarks (such as Hellaswag, Winogrande, piqa, boolq, and arc).

Novel technique benchmark evaluation

Image source

The evaluation method, relying on model prompts and subsequent response analysis, proves effective but may overlook more intricate, adversarial information extraction methods.

While the technique is promising, further research is required for refinement and expansion, particularly in addressing broader unlearning tasks within LLMs.

Novel Unlearning Technique Challenges

While Microsoft’s unlearning technique shows promise, several AI copyright challenges and constraints exist.

Key limitations and areas for enhancement encompass:

  • Leaks of copyright information: The method may not entirely mitigate the risk of copyright information leaks, as the model might retain some knowledge of the target content during the fine-tuning process.
  • Evaluation of various datasets: To gauge effectiveness, the technique must undergo additional evaluation across diverse datasets, as the initial experiment focused solely on the Harry Potter books.
  • Scalability: Testing on larger datasets and more intricate language models is imperative to assess the technique’s applicability and adaptability in real-world scenarios.

The rise in AI-related legal cases, particularly copyright lawsuits targeting LLMs, highlights the need for clear guidelines. Promising developments, like the unlearning method proposed by Microsoft, pave a path toward ethical, legal, and responsible AI.

Don’t miss out on the latest news and analysis in AI and ML – visit unite.ai today.

Credit: Source link

ShareTweetSendSharePin

Related Posts

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents
AI & Technology

AWS Introduces Pizza Bot: An Open Source Inbox for Background AI Agents

September 13, 2026
Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference
AI & Technology

Implementation of Machine Learning Workflows with NVIDIA cuML, RAPIDS, GPU Benchmarking, Explainability, Clustering, and Model Inference

September 13, 2026
Why Do Routers Have So Many Antennas?
AI & Technology

Why Do Routers Have So Many Antennas?

September 13, 2026
Hyundai Motor Group Puts Data Flywheel Into Full Operation – Unite.AI
AI & Technology

Hyundai Motor Group Puts Data Flywheel Into Full Operation – Unite.AI

September 13, 2026
Next Post
Trump pleads not guilty to four counts in 2020 election interference case

Trump pleads not guilty to four counts in 2020 election interference case

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
For years, they warned AI could kill all humans. Now people are listening. – The Washington Post

For years, they warned AI could kill all humans. Now people are listening. – The Washington Post

September 10, 2026
Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI

Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI

September 10, 2026
Watches, warnings discontinued as Hurricane Lowell pulls away from Hawaii – Hawaii News Now

Watches, warnings discontinued as Hurricane Lowell pulls away from Hawaii – Hawaii News Now

September 8, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!