• bitcoinBitcoin(BTC)$85,787.001.86%
  • ethereumEthereum(ETH)$2,737.870.89%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$786.720.47%
  • rippleXRP(XRP)$1.522.38%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$116.480.76%
  • tronTRON(TRX)$0.3470880.82%
  • zcashZcash(ZEC)$1,502.38-1.38%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.010.00%
  • HyperliquidHyperliquid(HYPE)$95.16-0.15%
  • dogecoinDogecoin(DOGE)$0.0976145.38%
  • moneroMonero(XMR)$567.73-1.61%
  • whitebitWhiteBIT Coin(WBT)$86.250.90%
  • chainlinkChainlink(LINK)$12.90-0.83%
  • USDSUSDS(USDS)$1.00-0.01%
  • RainRain(RAIN)$0.013464-4.77%
  • cardanoCardano(ADA)$0.2447602.11%
  • leo-tokenLEO Token(LEO)$8.980.37%
  • stellarStellar(XLM)$0.209662-0.76%
  • nearNEAR Protocol(NEAR)$4.433.78%
  • bitcoin-cashBitcoin Cash(BCH)$271.692.00%
  • uniswapUniswap(UNI)$8.74-1.59%
  • Ethena USDeEthena USDe(USDE)$1.00-0.03%
  • avalanche-2Avalanche(AVAX)$10.85-4.22%
  • CantonCanton(CC)$0.1177300.64%
  • litecoinLitecoin(LTC)$60.08-0.21%
  • daiDai(DAI)$1.000.02%
  • USD1USD1(USD1)$1.00-0.03%
  • suiSui(SUI)$1.010.62%
  • hedera-hashgraphHedera(HBAR)$0.0935564.35%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.42-0.60%
  • BittensorBittensor(TAO)$313.8412.26%
  • shiba-inuShiba Inu(SHIB)$0.0000064.71%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0652522.27%
  • MemeCoreMemeCore(M)$1.33-11.98%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • tether-goldTether Gold(XAUT)$4,330.51-0.48%
  • okbOKB(OKB)$122.310.43%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • BitwayBitway(BTW)$0.84-0.41%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.57%
  • aaveAave(AAVE)$140.74-3.75%
  • mantleMantle(MNT)$0.642.30%
  • Pump.funPump.fun(PUMP)$0.0045372.48%
  • OndoOndo(ONDO)$0.428403-3.61%
  • EthenaEthena(ENA)$0.205305-7.77%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

OpenAI experiment finds that sparse models could give AI builders the tools to debug neural networks

November 14, 2025
in AI & Technology
Reading Time: 3 mins read
A A
OpenAI experiment finds that sparse models could give AI builders the tools to debug neural networks
ShareShareShareShareShare

OpenAI researchers are experimenting with a new approach to designing neural networks, with the aim of making AI models easier to understand, debug, and govern. Sparse models can provide enterprises with a better understanding of how these models make decisions. 

YOU MAY ALSO LIKE

OpenAI Faces Lawsuit From British Columbia Over Tumbler Ridge Shooting

NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 49%

Understanding how models choose to respond, a big selling point of reasoning models for enterprises, can provide a level of trust for organizations when they turn to AI models for insights. 

The method called for OpenAI scientists and researchers to look at and evaluate models not by analyzing post-training performance, but by adding interpretability or understanding through sparse circuits.

OpenAI notes that much of the opacity of AI models stems from how most models are designed, so to gain a better understanding of model behavior, they must create workarounds. 

“Neural networks power today’s most capable AI systems, but they remain difficult to understand,” OpenAI wrote in a blog post. “We don’t write these models with explicit step-by-step instructions. Instead, they learn by adjusting billions of internal connections or weights until they master a task. We design the rules of training, but not the specific behaviors that emerge, and the result is a dense web of connections that no human can easily decipher.”

To enhance the interpretability of the mix, OpenAI examined an architecture that trains untangled neural networks, making them simpler to understand. The team trained language models with a similar architecture to existing models, such as GPT-2, using the same training schema. 

The result: improved interpretability. 

The path toward interpretability

Understanding how models work, giving us insight into how they're making their determinations, is important because these have a real-world impact, OpenAI says.  

The company defines interpretability as “methods that help us understand why a model produced a given output.” There are several ways to achieve interpretability: chain-of-thought interpretability, which reasoning models often leverage, and mechanistic interpretability, which involves reverse-engineering a model’s mathematical structure.

OpenAI focused on improving mechanistic interpretability, which it said “has so far been less immediately useful, but in principle, could offer a more complete explanation of the model’s behavior.”

“By seeking to explain model behavior at the most granular level, mechanistic interpretability can make fewer assumptions and give us more confidence. But the path from low-level details to explanations of complex behaviors is much longer and more difficult,” according to OpenAI. 

Better interpretability allows for better oversight and gives early warning signs if the model’s behavior no longer aligns with policy. 

OpenAI noted that improving mechanistic interpretability “is a very ambitious bet,” but research on sparse networks has improved this. 

How to untangle a model 

To untangle the mess of connections a model makes, OpenAI first cut most of these connections. Since transformer models like GPT-2 have thousands of connections, the team had to “zero out” these circuits. Each will only talk to a select number, so the connections become more orderly.

Next, the team ran “circuit tracing” on tasks to create groupings of interpretable circuits. The last task involved pruning the model “to obtain the smallest circuit which achieves a target loss on the target distribution,” according to OpenAI. It targeted a loss of 0.15 to isolate the exact nodes and weights responsible for behaviors. 

“We show that pruning our weight-sparse models yields roughly 16-fold smaller circuits on our tasks than pruning dense models of comparable pretraining loss. We are also able to construct arbitrarily accurate circuits at the cost of more edges. This shows that circuits for simple behaviors are substantially more disentangled and localizable in weight-sparse models than dense models,” the report said. 

Small models become easier to train

Although OpenAI managed to create sparse models that are easier to understand, these remain significantly smaller than most foundation models used by enterprises. Enterprises increasingly use small models, but frontier models, such as its flagship GPT-5.1, will still benefit from improved interpretability down the line. 

Other model developers also aim to understand how their AI models think. Anthropic, which has been researching interpretability for some time, recently revealed that it had “hacked” Claude’s brain — and Claude noticed. Meta also is working to find out how reasoning models make their decisions. 

As more enterprises turn to AI models to help make consequential decisions for their business, and eventually customers, research into understanding how models think would give the clarity many organizations need to trust models more. 

Credit: Source link

ShareTweetSendSharePin

Related Posts

OpenAI Faces Lawsuit From British Columbia Over Tumbler Ridge Shooting
AI & Technology

OpenAI Faces Lawsuit From British Columbia Over Tumbler Ridge Shooting

September 22, 2026
NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 49%
AI & Technology

NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 49%

September 22, 2026
SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same / Price as Grok 4.6
AI & Technology

SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6

September 22, 2026
Why It’s Important To Unplug Your PC During A Power Outage
AI & Technology

Why It’s Important To Unplug Your PC During A Power Outage

September 22, 2026
Next Post
Cardi B Gives Birth to Baby Boy, Her 1st Child with Patriots' Stefon Diggs – Bleacher Report

Cardi B Gives Birth to Baby Boy, Her 1st Child with Patriots' Stefon Diggs - Bleacher Report

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Google’s Latest Pixel Drop Will Keep You More Connected To Your VIPs

Google’s Latest Pixel Drop Will Keep You More Connected To Your VIPs

September 15, 2026
Nutanix: Profit Taking Is Appropriate Here (Downgrade)

Nutanix: Profit Taking Is Appropriate Here (Downgrade)

September 21, 2026
Katie Stein, CEO of ASAPP – Interview Series – Unite.AI

Katie Stein, CEO of ASAPP – Interview Series – Unite.AI

September 15, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!