• bitcoinBitcoin(BTC)$78,778.00-0.40%
  • ethereumEthereum(ETH)$2,495.960.22%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$751.731.48%
  • rippleXRP(XRP)$1.421.55%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$103.66-0.02%
  • tronTRON(TRX)$0.3389571.16%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.00%
  • zcashZcash(ZEC)$1,181.794.50%
  • HyperliquidHyperliquid(HYPE)$85.871.59%
  • dogecoinDogecoin(DOGE)$0.090039-0.53%
  • RainRain(RAIN)$0.016073-1.25%
  • USDSUSDS(USDS)$1.000.00%
  • whitebitWhiteBIT Coin(WBT)$81.626.62%
  • moneroMonero(XMR)$498.80-2.83%
  • chainlinkChainlink(LINK)$12.46-1.71%
  • leo-tokenLEO Token(LEO)$9.230.29%
  • cardanoCardano(ADA)$0.217698-1.01%
  • stellarStellar(XLM)$0.187929-1.50%
  • bitcoin-cashBitcoin Cash(BCH)$258.51-0.36%
  • daiDai(DAI)$1.000.01%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.1092483.24%
  • USD1USD1(USD1)$1.00-0.01%
  • uniswapUniswap(UNI)$6.79-2.87%
  • litecoinLitecoin(LTC)$54.17-2.74%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.390.46%
  • hedera-hashgraphHedera(HBAR)$0.078707-4.32%
  • avalanche-2Avalanche(AVAX)$7.98-0.82%
  • suiSui(SUI)$0.81-1.65%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • shiba-inuShiba Inu(SHIB)$0.000005-1.34%
  • nearNEAR Protocol(NEAR)$2.29-1.22%
  • crypto-com-chainCronos(CRO)$0.0602935.75%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • MemeCoreMemeCore(M)$1.224.79%
  • tether-goldTether Gold(XAUT)$4,373.10-1.10%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$255.86-1.16%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$114.59-2.07%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.40%
  • mantleMantle(MNT)$0.631.02%
  • polkadotPolkadot(DOT)$1.2013.10%
  • AsterAster(ASTER)$0.75-2.01%
  • aaveAave(AAVE)$129.03-1.78%
  • pax-goldPAX Gold(PAXG)$4,375.51-1.10%
  • Pump.funPump.fun(PUMP)$0.0044440.24%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers use Harry Potter to make AI forget material

October 6, 2023
in AI & Technology
Reading Time: 3 mins read
A A
Researchers use Harry Potter to make AI forget material
ShareShareShareShareShare

VentureBeat presents: AI Unleashed – An exclusive executive event for enterprise data leaders. Network and learn with industry peers. Learn More


As the debate heats up around the use of copyrighted works to train large language models (LLMs) such as OpenAI’s ChatGPT, Meta’s Llama 2, Anthropic’s Claude 2, one obvious question arises: can these models even be altered or edited to remove their knowledge of such works, without totally retraining them or rearchitecting them? 

YOU MAY ALSO LIKE

How To Change And Customize Your Apple CarPlay Display

Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction

In a new paper published on the open access and non-peer reviewed site arXiv.org, co-authors Ronen Eldan of Microsoft Research and Mark Russinovich of Microsoft Azure propose a new way of doing exactly this by erasing specific information from a sample LLM — namely, all knowledge of the existence of the Harry Potter books (including characters and plots) from Meta’s open source Llama 2-7B. 

As the Microsoft researchers write: “While the model took over 184K GPU-hours to pretrain, we show that in about 1 GPU hour of finetuning, we effectively erase the model’s ability to generate or recall Harry Potter-related content.”

This work provides an important step toward adaptable language models. The ability to refine AI over time according to shifting organizational needs is key to long-term, enterprise-safe deployments.

Event

AI Unleashed

An exclusive invite-only evening of insights and networking, designed for senior enterprise executives overseeing data stacks and strategies.

 

Learn More

The magic formula

“Traditional models of [machine] learning predominantly focus on adding or reinforcing knowledge through basic fine-tuning but do not provide straightforward mechanisms to ‘forget’ or ‘unlearn’ knowledge,” the authors write.

How did they overcome this? They developed a three-part technique to approximate unlearning specific information in LLMs. 

First, they trained a model on the target data (Harry Potter books) to identify tokens most related to it by comparing predictions to a baseline model. 

Second, they replaced unique Harry Potter expressions with generic counterparts and generated alternative predictions approximating a model without that training. 

Third, they fine-tuned the baseline model on these alternative predictions, effectively erasing the original text from its memory when prompted with the context.

To evaluate, they tested the model’s ability to generate or discuss Harry Potter content using 300 automatically generated prompts, as well as by inspecting token probabilities. As Eldan and Russinovich state, “to the best of our knowledge, this is the first paper to present an effective technique for unlearning in generative language models.”

They found that while the original model could easily discuss intricate Harry Potter plot details, after only an hour of finetuning their technique, “it’s possible for the model to essentially ‘forget’ the intricate narratives of the Harry Potter series.” Performance on standard benchmarks like ARC, BoolQ and Winogrande “remains almost unaffected.”

Expelliarmus-ing expectations

As the authors note, more testing is still needed given limitations of their evaluation approach. Their technique may also be more effective for fictional texts than non-fiction, since fictional worlds contain more unique references. 

Nonetheless, this proof-of-concept provides “a foundational step towards creating more responsible, adaptable, and legally compliant LLMs in the future.” As the authors conclude, further refinement could help address “ethical guidelines, societal values, or specific user requirements.”

In summarizing their findings, the authors state: “Our technique offers a promising start, but its applicability across various content types remains to be thoroughly tested. The presented approach offers a foundation, but further research is needed to refine and extend the methodology for broader unlearning tasks in LLMs.” 

Moving forward, more general and robust techniques for selective forgetting could help ensure AI systems remain dynamically aligned with priorities, business or societal, as needs change over time.

VentureBeat’s mission is to be a digital town square for technical decision-makers to gain knowledge about transformative enterprise technology and transact. Discover our Briefings.

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Change And Customize Your Apple CarPlay Display
AI & Technology

How To Change And Customize Your Apple CarPlay Display

September 8, 2026
Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction
AI & Technology

Sierra Open-Sources Hyper-τ-Bench, a Benchmark for Agent Construction

September 8, 2026
SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas
AI & Technology

SpaceX’s Recovered Starship 40 Will Take Months To Get Back To Texas

September 8, 2026
What Is Roku’s Secret Menu And How Do You Unlock It?
AI & Technology

What Is Roku’s Secret Menu And How Do You Unlock It?

September 8, 2026
Next Post
What the Loss of the Note 7 Means for Samsung

What the Loss of the Note 7 Means for Samsung

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
The AI Arms Race | Seeking Alpha

The AI Arms Race | Seeking Alpha

September 3, 2026
Staying With Your Employer Is a Financial Choice

Staying With Your Employer Is a Financial Choice

September 5, 2026
Elon Musk launches steering-wheel-free Tesla Cybercabs

Elon Musk launches steering-wheel-free Tesla Cybercabs

September 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!