• bitcoinBitcoin(BTC)$76,674.000.37%
  • ethereumEthereum(ETH)$2,456.931.11%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$743.222.41%
  • rippleXRP(XRP)$1.300.27%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$102.653.49%
  • tronTRON(TRX)$0.335017-0.20%
  • zcashZcash(ZEC)$1,487.598.97%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.08%
  • HyperliquidHyperliquid(HYPE)$86.709.70%
  • dogecoinDogecoin(DOGE)$0.0823831.77%
  • moneroMonero(XMR)$517.773.46%
  • USDSUSDS(USDS)$1.000.01%
  • whitebitWhiteBIT Coin(WBT)$79.000.61%
  • RainRain(RAIN)$0.012701-1.81%
  • chainlinkChainlink(LINK)$11.533.70%
  • leo-tokenLEO Token(LEO)$8.89-1.04%
  • cardanoCardano(ADA)$0.2092656.50%
  • stellarStellar(XLM)$0.184450-0.08%
  • uniswapUniswap(UNI)$7.8715.21%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • bitcoin-cashBitcoin Cash(BCH)$235.796.49%
  • daiDai(DAI)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$54.444.62%
  • nearNEAR Protocol(NEAR)$3.2322.60%
  • CantonCanton(CC)$0.1058107.50%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.341.99%
  • avalanche-2Avalanche(AVAX)$7.702.13%
  • hedera-hashgraphHedera(HBAR)$0.0751771.54%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • suiSui(SUI)$0.766.30%
  • shiba-inuShiba Inu(SHIB)$0.0000055.33%
  • crypto-com-chainCronos(CRO)$0.0582842.78%
  • MemeCoreMemeCore(M)$1.2613.25%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,346.730.95%
  • BittensorBittensor(TAO)$236.305.59%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • okbOKB(OKB)$112.321.25%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.18%
  • AsterAster(ASTER)$0.754.49%
  • aaveAave(AAVE)$130.777.94%
  • BitwayBitway(BTW)$0.71-0.15%
  • Pump.funPump.fun(PUMP)$0.0040548.05%
  • pax-goldPAX Gold(PAXG)$4,344.190.85%
  • OndoOndo(ONDO)$0.3862679.39%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Here’s what’s really going on inside an LLM’s neural network

May 22, 2024
in Market & News
Reading Time: 4 mins read
A A
Here’s what’s really going on inside an LLM’s neural network
ShareShareShareShareShare

Aurich Lawson | Getty Images


With most computer programs—even complex ones—you can meticulously trace through the code and memory usage to figure out why that program generates any specific behavior or output. That’s generally not true in the field of generative AI, where the non-interpretable neural networks underlying these models make it hard for even experts to figure out precisely why they often confabulate information, for instance.

YOU MAY ALSO LIKE

Parents of nonverbal child found dead extradited to S.C.

Fox News parts ways with longtime host Maria Bartiromo

Now, new research from Anthropic offers a new window into what’s going on inside the Claude LLM’s “black box.” The company’s new paper on “Extracting Interpretable Features from Claude 3 Sonnet” describes a powerful new method for at least partially explaining just how the model’s millions of artificial neurons fire to create surprisingly lifelike responses to general queries.

Opening the hood

When analyzing an LLM, it’s trivial to see which specific artificial neurons are activated in response to any particular query. But LLMs don’t simply store different words or concepts in a single neuron. Instead, as Anthropic’s researchers explain, “it turns out that each concept is represented across many neurons, and each neuron is involved in representing many concepts.”

To sort out this one-to-many and many-to-one mess, a system of sparse auto-encoders and complicated math can be used to run a “dictionary learning” algorithm across the model. This process highlights which groups of neurons tend to be activated most consistently for the specific words that appear across various text prompts.

The same internal LLM
Enlarge / The same internal LLM “feature” describes the Golden Gate Bridge in multiple languages and modes.

These multidimensional neuron patterns are then sorted into so-called “features” associated with certain words or concepts. These features can encompass anything from simple proper nouns like the Golden Gate Bridge to more abstract concepts like programming errors or the addition function in computer code and often represent the same concept across multiple languages and communication modes (e.g., text and images).

Advertisement

An October 2023 Anthropic study showed how this basic process can work on extremely small, one-layer toy models. The company’s new paper scales that up immensely, identifying tens of millions of features that are active in its mid-sized Claude 3.0 Sonnet model. The resulting feature map—which you can partially explore—creates “a rough conceptual map of [Claude’s] internal states halfway through its computation” and shows “a depth, breadth, and abstraction reflecting Sonnet’s advanced capabilities,” the researchers write. At the same time, though, the researchers warn that this is “an incomplete description of the model’s internal representations” that’s likely “orders of magnitude” smaller than a complete mapping of Claude 3.

A simplified map shows some of the concepts that are "near" the "inner conflict" feature in Anthropic's Claude model.
Enlarge / A simplified map shows some of the concepts that are “near” the “inner conflict” feature in Anthropic’s Claude model.

Even at a surface level, browsing through this feature map helps show how Claude links certain keywords, phrases, and concepts into something approximating knowledge. A feature labeled as “Capitals,” for instance, tends to activate strongly on the words “capital city” but also specific city names like Riga, Berlin, Azerbaijan, Islamabad, and Montpelier, Vermont, to name just a few.

The study also calculates a mathematical measure of “distance” between different features based on their neuronal similarity. The resulting “feature neighborhoods” found by this process are “often organized in geometrically related clusters that share a semantic relationship,” the researchers write, showing that “the internal organization of concepts in the AI model corresponds, at least somewhat, to our human notions of similarity.” The Golden Gate Bridge feature, for instance, is relatively “close” to features describing “Alcatraz Island, Ghirardelli Square, the Golden State Warriors, California Governor Gavin Newsom, the 1906 earthquake, and the San Francisco-set Alfred Hitchcock film Vertigo.”

Some of the most important features involved in answering a query about the capital of Kobe Bryant's team's state.
Enlarge / Some of the most important features involved in answering a query about the capital of Kobe Bryant’s team’s state.

Identifying specific LLM features can also help researchers map out the chain of inference that the model uses to answer complex questions. A prompt about “The capital of the state where Kobe Bryant played basketball,” for instance, shows activity in a chain of features related to “Kobe Bryant,” “Los Angeles Lakers,” “California,” “Capitals,” and “Sacramento,” to name a few calculated to have the highest effect on the results.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Parents of nonverbal child found dead extradited to S.C.
Market & News

Parents of nonverbal child found dead extradited to S.C.

September 18, 2026
Fox News parts ways with longtime host Maria Bartiromo
Market & News

Fox News parts ways with longtime host Maria Bartiromo

September 18, 2026
Parents extradited after death of their 5-year-old daughter
Market & News

Parents extradited after death of their 5-year-old daughter

September 18, 2026
Police ID suspect in deadly Minneapolis shooting as man facing eviction
Market & News

Police ID suspect in deadly Minneapolis shooting as man facing eviction

September 18, 2026
Next Post
Fed officials worried inflation too stubborn to justify rate cut

Fed officials worried inflation too stubborn to justify rate cut

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Landmark 9/11 anniversary marked in New York, Pentagon and Shanksville ceremonies

Landmark 9/11 anniversary marked in New York, Pentagon and Shanksville ceremonies

September 13, 2026
Three Steps to Hedging a Portfolio With Futures

Three Steps to Hedging a Portfolio With Futures

September 16, 2026
Wells Fargo finance chief sees stronger 2026 loan growth, healthy US economy

Wells Fargo finance chief sees stronger 2026 loan growth, healthy US economy

September 15, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!