• bitcoinBitcoin(BTC)$76,543.000.46%
  • ethereumEthereum(ETH)$2,446.301.31%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$737.291.91%
  • rippleXRP(XRP)$1.300.64%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$101.392.93%
  • tronTRON(TRX)$0.334745-0.29%
  • zcashZcash(ZEC)$1,462.218.73%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.07%
  • HyperliquidHyperliquid(HYPE)$86.5910.81%
  • dogecoinDogecoin(DOGE)$0.0817701.54%
  • moneroMonero(XMR)$518.113.45%
  • USDSUSDS(USDS)$1.000.01%
  • whitebitWhiteBIT Coin(WBT)$78.800.74%
  • RainRain(RAIN)$0.012696-1.44%
  • chainlinkChainlink(LINK)$11.433.97%
  • leo-tokenLEO Token(LEO)$8.90-0.57%
  • cardanoCardano(ADA)$0.2048835.03%
  • stellarStellar(XLM)$0.183575-0.31%
  • uniswapUniswap(UNI)$7.7215.62%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • bitcoin-cashBitcoin Cash(BCH)$235.426.90%
  • daiDai(DAI)$1.000.01%
  • USD1USD1(USD1)$1.000.00%
  • litecoinLitecoin(LTC)$54.315.13%
  • nearNEAR Protocol(NEAR)$3.1820.19%
  • CantonCanton(CC)$0.1028374.33%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.342.25%
  • avalanche-2Avalanche(AVAX)$7.631.92%
  • hedera-hashgraphHedera(HBAR)$0.0750812.32%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.0000055.16%
  • suiSui(SUI)$0.755.57%
  • crypto-com-chainCronos(CRO)$0.0579182.75%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • MemeCoreMemeCore(M)$1.218.70%
  • tether-goldTether Gold(XAUT)$4,357.841.44%
  • BittensorBittensor(TAO)$234.264.45%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$111.691.48%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.23%
  • AsterAster(ASTER)$0.755.24%
  • aaveAave(AAVE)$128.767.33%
  • BitwayBitway(BTW)$0.71-2.85%
  • pax-goldPAX Gold(PAXG)$4,355.541.38%
  • Pump.funPump.fun(PUMP)$0.0040399.81%
  • mantleMantle(MNT)$0.573.27%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Here’s what’s really going on inside an LLM’s neural network

May 22, 2024
in Market & News
Reading Time: 4 mins read
A A
Here’s what’s really going on inside an LLM’s neural network
ShareShareShareShareShare

Aurich Lawson | Getty Images


With most computer programs—even complex ones—you can meticulously trace through the code and memory usage to figure out why that program generates any specific behavior or output. That’s generally not true in the field of generative AI, where the non-interpretable neural networks underlying these models make it hard for even experts to figure out precisely why they often confabulate information, for instance.

YOU MAY ALSO LIKE

Police ID suspect in deadly Minneapolis shooting as man facing eviction

Stay Tuned NOW Streaming Behind The Scenes! – Sept 03

Now, new research from Anthropic offers a new window into what’s going on inside the Claude LLM’s “black box.” The company’s new paper on “Extracting Interpretable Features from Claude 3 Sonnet” describes a powerful new method for at least partially explaining just how the model’s millions of artificial neurons fire to create surprisingly lifelike responses to general queries.

Opening the hood

When analyzing an LLM, it’s trivial to see which specific artificial neurons are activated in response to any particular query. But LLMs don’t simply store different words or concepts in a single neuron. Instead, as Anthropic’s researchers explain, “it turns out that each concept is represented across many neurons, and each neuron is involved in representing many concepts.”

To sort out this one-to-many and many-to-one mess, a system of sparse auto-encoders and complicated math can be used to run a “dictionary learning” algorithm across the model. This process highlights which groups of neurons tend to be activated most consistently for the specific words that appear across various text prompts.

The same internal LLM
Enlarge / The same internal LLM “feature” describes the Golden Gate Bridge in multiple languages and modes.

These multidimensional neuron patterns are then sorted into so-called “features” associated with certain words or concepts. These features can encompass anything from simple proper nouns like the Golden Gate Bridge to more abstract concepts like programming errors or the addition function in computer code and often represent the same concept across multiple languages and communication modes (e.g., text and images).

Advertisement

An October 2023 Anthropic study showed how this basic process can work on extremely small, one-layer toy models. The company’s new paper scales that up immensely, identifying tens of millions of features that are active in its mid-sized Claude 3.0 Sonnet model. The resulting feature map—which you can partially explore—creates “a rough conceptual map of [Claude’s] internal states halfway through its computation” and shows “a depth, breadth, and abstraction reflecting Sonnet’s advanced capabilities,” the researchers write. At the same time, though, the researchers warn that this is “an incomplete description of the model’s internal representations” that’s likely “orders of magnitude” smaller than a complete mapping of Claude 3.

A simplified map shows some of the concepts that are "near" the "inner conflict" feature in Anthropic's Claude model.
Enlarge / A simplified map shows some of the concepts that are “near” the “inner conflict” feature in Anthropic’s Claude model.

Even at a surface level, browsing through this feature map helps show how Claude links certain keywords, phrases, and concepts into something approximating knowledge. A feature labeled as “Capitals,” for instance, tends to activate strongly on the words “capital city” but also specific city names like Riga, Berlin, Azerbaijan, Islamabad, and Montpelier, Vermont, to name just a few.

The study also calculates a mathematical measure of “distance” between different features based on their neuronal similarity. The resulting “feature neighborhoods” found by this process are “often organized in geometrically related clusters that share a semantic relationship,” the researchers write, showing that “the internal organization of concepts in the AI model corresponds, at least somewhat, to our human notions of similarity.” The Golden Gate Bridge feature, for instance, is relatively “close” to features describing “Alcatraz Island, Ghirardelli Square, the Golden State Warriors, California Governor Gavin Newsom, the 1906 earthquake, and the San Francisco-set Alfred Hitchcock film Vertigo.”

Some of the most important features involved in answering a query about the capital of Kobe Bryant's team's state.
Enlarge / Some of the most important features involved in answering a query about the capital of Kobe Bryant’s team’s state.

Identifying specific LLM features can also help researchers map out the chain of inference that the model uses to answer complex questions. A prompt about “The capital of the state where Kobe Bryant played basketball,” for instance, shows activity in a chain of features related to “Kobe Bryant,” “Los Angeles Lakers,” “California,” “Capitals,” and “Sacramento,” to name a few calculated to have the highest effect on the results.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Police ID suspect in deadly Minneapolis shooting as man facing eviction
Market & News

Police ID suspect in deadly Minneapolis shooting as man facing eviction

September 18, 2026
Stay Tuned NOW Streaming Behind The Scenes! – Sept 03
Market & News

Stay Tuned NOW Streaming Behind The Scenes! – Sept 03

September 18, 2026
Wild animal attacks three people in New Hampshire
Market & News

Wild animal attacks three people in New Hampshire

September 18, 2026
Gas prices are about to take a big jump, analysts say, with the worst still to come – washingtonpost.com
Market & News

Gas prices are about to take a big jump, analysts say, with the worst still to come – washingtonpost.com

September 18, 2026
Next Post
Fed officials worried inflation too stubborn to justify rate cut

Fed officials worried inflation too stubborn to justify rate cut

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Widow shares story of FBI agent who died of cancer after 9/11

Widow shares story of FBI agent who died of cancer after 9/11

September 13, 2026
ATI Inc. (ATI) Presents at Morgan Stanley’s 14th Annual Laguna Conference Transcript

ATI Inc. (ATI) Presents at Morgan Stanley’s 14th Annual Laguna Conference Transcript

September 17, 2026
AI Safety Can’t Rely on an Honor Code – Unite.AI

AI Safety Can’t Rely on an Honor Code – Unite.AI

September 16, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!