• bitcoinBitcoin(BTC)$63,846.00-2.00%
  • ethereumEthereum(ETH)$1,870.17-2.60%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$599.43-1.50%
  • usd-coinUSDC(USDC)$1.000.00%
  • rippleXRP(XRP)$1.02-2.30%
  • solanaSolana(SOL)$75.74-1.70%
  • tronTRON(TRX)$0.3306330.30%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.010.60%
  • HyperliquidHyperliquid(HYPE)$54.820.20%
  • dogecoinDogecoin(DOGE)$0.069559-1.50%
  • USDSUSDS(USDS)$1.000.00%
  • RainRain(RAIN)$0.0128662.00%
  • leo-tokenLEO Token(LEO)$9.65-0.60%
  • zcashZcash(ZEC)$496.33-4.10%
  • moneroMonero(XMR)$390.97-1.50%
  • cardanoCardano(ADA)$0.193277-1.30%
  • whitebitWhiteBIT Coin(WBT)$55.09-2.20%
  • chainlinkChainlink(LINK)$8.22-1.20%
  • stellarStellar(XLM)$0.162194-0.40%
  • daiDai(DAI)$1.000.00%
  • bitcoin-cashBitcoin Cash(BCH)$213.38-1.70%
  • USD1USD1(USD1)$1.00-0.10%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.096011-2.40%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.32-0.80%
  • litecoinLitecoin(LTC)$45.00-2.70%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • Circle USYCCircle USYC(USYC)$1.130.00%
  • hedera-hashgraphHedera(HBAR)$0.067879-1.90%
  • suiSui(SUI)$0.69-2.00%
  • avalanche-2Avalanche(AVAX)$6.49-0.90%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.30%
  • tether-goldTether Gold(XAUT)$4,350.93-0.10%
  • uniswapUniswap(UNI)$3.92-2.50%
  • crypto-com-chainCronos(CRO)$0.046932-2.60%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.00%
  • nearNEAR Protocol(NEAR)$1.60-1.40%
  • okbOKB(OKB)$93.78-1.30%
  • BittensorBittensor(TAO)$200.41-2.30%
  • pax-goldPAX Gold(PAXG)$4,366.23-0.10%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0537861.40%
  • OndoOndo(ONDO)$0.344290-2.30%
  • HTX DAOHTX DAO(HTX)$0.0000020.70%
  • AsterAster(ASTER)$0.600.10%
  • usddUSDD(USDD)$1.000.00%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • mantleMantle(MNT)$0.4386503.30%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Meta AI Releases Muse Glimmer: A 30B Open-Weights Agentic Model That Runs on One Consumer GPU

August 10, 2026
in AI & Technology
Reading Time: 18 mins read
A A
Meta AI Releases Muse Glimmer: A 30B Open-Weights Agentic Model That Runs on One Consumer GPU
ShareShareShareShareShare

Meta has released Muse Glimmer, a 30-billion-parameter multimodal model distilled from Muse Spark. It is tuned for always-on local agent workflows, and ships under Apache 2.0. A 30B model normally needs over 55 GB of memory at full precision. Meta compresses it to roughly 4-bit, then adds block-level speculative decoding so it answers fast enough to sit inside a real agent loop. The result runs on one consumer GPU or a Mac, with no network call.

Is it deployable?

Yes, the weights are open under Apache 2.0. The Hugging Face collection carries BF16 weights, GGUF k-quants, ExecuTorch builds, and the DFlash drafter. Self-hosting is the day-one path.

  • Which companies: Solo developers and startups can run it on one 24 GB GPU or an M4/M5 Max Mac. Mid-market teams get on-prem inference without a per-token bill. Regulated enterprises get an air-gappable agent. Meta advises adding system-level guardrails rather than shipping the model as a bare endpoint.
  • Industries: Healthcare, legal, financial services, defense and public sector, manufacturing, and field service. These are the settings where data residency, offline operation, or latency rule out a cloud call.
  • Applications: Desktop agents that read screenshots, coding agents, and schema-based function calling. Also document and chart understanding, synthetic data generation, and LLM-as-a-judge evaluation.

Model and training

Muse Glimmer is a dense causal transformer with a dedicated perception encoder. Total parameters are roughly 30B, including the vision tower. Grouped-query attention uses 32 query heads and 2 KV heads. Attention repeats a [Local, Local, Local, Global] pattern with a 2,048 sliding window. RoPE is applied to local layers only, with theta 500,000. The vision side is a ~1.8B ViT-G/14 perception encoder accepting up to 4,096 visual tokens per image. Context length is 131,072+, vocabulary is 202,048 tokens, and the knowledge cutoff is January 4, 2026. Input is text and image; output is text. Audio is not supported, and video is processed as individual frames.

Training ran in three phases:

YOU MAY ALSO LIKE

What’s The Difference Between MagSafe And Qi Wireless Charging?

Most Enterprise AI Isn’t Secure. Here’s What Businesses Can Do – Unite.AI

  • Pre-training used logit distillation on Muse Spark’s outputs.
  • Mid-training added longer-context, agent-heavy data with richer reasoning traces.
  • Post-training combined supervised fine-tuning with on-policy distillation and reinforcement learning across general, reasoning, coding, and agentic domains.

Fitting 30B onto consumer hardware

At full precision the model needs over 55 GB of memory. Meta compresses weights to approximately 4-bit precision, which brings the language model under 20 GB. That leaves headroom inside a 24 GB or 32 GB envelope. The KV cache, perception encoder, and drafter share it. Two quantized builds ship. K-Quant-Dynamic targets 32 GB VRAM at 0.2% average degradation. K-Quant-17GB targets 24 GB VRAM at 1.0%. Degradation is averaged over accuracy metrics across 15 common benchmarks.

Generation speed comes from DFlash, a block-diffusion drafter that predicts 16 tokens in one forward pass. The main model verifies the block in parallel. The drafter uses 5 layers, sliding-window attention at 2,048, and 32 query / 8 KV heads. Meta measured K-Quant-17GB at batch size 1 with greedy decoding. On an RTX 5090, throughput rises from 74.9 to 233.4 tok/s, a 3.1x speedup. Apple M5 Max moves from 26.6 to 50.2 tok/s, and M4 Max from 23.7 to 37.8 tok/s.

Benchmarks

Meta compares Muse Glimmer against Gemma4-31B and Qwen3.6-27B in thinking mode. It leads on MCP Atlas at 75.5, against 54.2 and 62.5. It also leads on DeepSearch QA at 74.6, Gaia2 at 43.3, and SWE-Bench Pro at 51.2. Reasoning scores follow: AIME 2026 at 94.7, IFBench at 77.0, AA-LCR at 80.0. Qwen3.6-27B stays ahead on OSWorld-Verified, 75.6 versus 65.9. It also leads TerminalBench 2.1 at 60.7 and SWE-Bench Verified at 77.2. The pattern is consistent. Muse Glimmer wins on agentic orchestration and reasoning. It trails on computer-use and terminal work.

On safety, Siren AgentDojo attack success rate is 28.4 with utility 94.2. Meta states the model does not meet the Frontier AI definition in its Advanced AI Scaling Framework. It rates chem/bio, cyber, and loss-of-control risk at moderate or lower.

Key Takeaways

  • 30B open-weights agentic model, Apache 2.0, distilled from Muse Spark.
  • 4-bit quantization fits it in 24 GB VRAM at 1.0% degradation.
  • DFlash 16-token block speculation gives 3.1x decode speedup on RTX 5090.
  • Beats both comparators on MCP Atlas, DeepSearch QA, and SWE-Bench Pro.
  • Trails Qwen3.6-27B on OSWorld-Verified and TerminalBench 2.1.

Check out the Model weights on HF, Details and Meta Blog. Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.? Connect with us


Asif Razzaq is the CEO of Marktechpost Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.

Credit: Source link

ShareTweetSendSharePin

Related Posts

What’s The Difference Between MagSafe And Qi Wireless Charging?
AI & Technology

What’s The Difference Between MagSafe And Qi Wireless Charging?

August 10, 2026
Most Enterprise AI Isn’t Secure. Here’s What Businesses Can Do – Unite.AI
AI & Technology

Most Enterprise AI Isn’t Secure. Here’s What Businesses Can Do – Unite.AI

August 10, 2026
A MacBook Neo Alternative That Needs More RAM
AI & Technology

A MacBook Neo Alternative That Needs More RAM

August 10, 2026
Your agent didn’t hallucinate; it exceeded its authority
AI & Technology

Your agent didn’t hallucinate; it exceeded its authority

August 10, 2026
Next Post
Lightning strikes the Eiffel Tower as a thunderstorm rages over Paris

Lightning strikes the Eiffel Tower as a thunderstorm rages over Paris

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
San Francisco man goes viral for embracing Chinatown’s culture, language and community

San Francisco man goes viral for embracing Chinatown’s culture, language and community

August 8, 2026
Iran makes dramatic new demands around Strait of Hormuz – Politico

Iran makes dramatic new demands around Strait of Hormuz – Politico

August 8, 2026
Ichor Holdings, Ltd. (ICHR) Q2 2026 Earnings Call Transcript

Ichor Holdings, Ltd. (ICHR) Q2 2026 Earnings Call Transcript

August 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!