• bitcoinBitcoin(BTC)$76,087.000.58%
  • ethereumEthereum(ETH)$2,407.340.35%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$718.720.82%
  • rippleXRP(XRP)$1.312.64%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$98.401.68%
  • tronTRON(TRX)$0.3355071.05%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-3.18%
  • zcashZcash(ZEC)$1,306.2817.28%
  • HyperliquidHyperliquid(HYPE)$78.342.23%
  • dogecoinDogecoin(DOGE)$0.0803960.61%
  • USDSUSDS(USDS)$1.000.01%
  • RainRain(RAIN)$0.013254-6.21%
  • moneroMonero(XMR)$497.830.03%
  • whitebitWhiteBIT Coin(WBT)$78.050.30%
  • chainlinkChainlink(LINK)$10.94-0.08%
  • leo-tokenLEO Token(LEO)$8.85-0.60%
  • cardanoCardano(ADA)$0.194818-0.37%
  • stellarStellar(XLM)$0.1814523.67%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • daiDai(DAI)$1.000.03%
  • bitcoin-cashBitcoin Cash(BCH)$217.891.13%
  • USD1USD1(USD1)$1.00-0.01%
  • uniswapUniswap(UNI)$6.391.01%
  • litecoinLitecoin(LTC)$51.20-0.15%
  • CantonCanton(CC)$0.0943333.43%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.31-0.73%
  • nearNEAR Protocol(NEAR)$2.5711.25%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • avalanche-2Avalanche(AVAX)$7.341.04%
  • hedera-hashgraphHedera(HBAR)$0.073455-1.81%
  • suiSui(SUI)$0.714.20%
  • shiba-inuShiba Inu(SHIB)$0.000005-1.38%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • crypto-com-chainCronos(CRO)$0.0560731.43%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,261.23-0.68%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • MemeCoreMemeCore(M)$1.121.35%
  • BittensorBittensor(TAO)$220.340.82%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • okbOKB(OKB)$110.370.94%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.16%
  • BitwayBitway(BTW)$0.746.28%
  • AsterAster(ASTER)$0.692.48%
  • pax-goldPAX Gold(PAXG)$4,263.40-0.70%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0570750.32%
  • mantleMantle(MNT)$0.551.41%
  • aaveAave(AAVE)$116.72-4.54%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

MixedBread AI Introduces Binary MRL: A Novel Embeddings Compression Method, Making Vector Search Scalable and Enable Embeddings-based Applications

April 14, 2024
in AI & Technology
Reading Time: 4 mins read
A A
MixedBread AI Introduces Binary MRL: A Novel Embeddings Compression Method, Making Vector Search Scalable and Enable Embeddings-based Applications
ShareShareShareShareShare

Mixedbread.ai recently introduced Binary MRL, a 64-byte embedding to address the challenge of scaling embeddings in natural language processing (NLP) applications due to their memory-intensive nature. In natural language processing (NLP), embeddings play a vital role in various tasks, such as recommendation systems, retrieval, and similarity search. However, the memory requirements of embeddings pose a significant challenge, particularly when dealing with massive datasets. The method aims to find a way to decrease the memory use for embeddings while maintaining their utility and effectiveness in NLP applications.

Currently, state-of-the-art models produce embeddings with high dimensions (e.g., 1024 dimensions), encoded in float32 format, requiring large memory for storage and retrieval. To address these limitations, researchers at mixedbread.ai have found two main approaches: Matryoshka Representation Learning (MRL) and Vector Quantization. MRL focuses on reducing the number of output dimensions of an embedding model while preserving accuracy. This is done by putting more important data in the earlier dimensions of the embedding, which lets the less important dimensions be cut off. On the other hand, Vector Quantization aims to reduce the size of each dimension by representing them as binary values instead of floating-point numbers. 

The proposed approach, Binary MRL, combines both methods to achieve simultaneous dimensionality reduction and compression of embeddings. By integrating MRL and Vector Quantization, Binary MRL aims to retain the semantic information encoded in embeddings while significantly reducing their memory footprint.

Binary MRL achieves compression by first reducing the number of output dimensions of the embedding model using MRL techniques. This involves training the model to preserve important information in fewer dimensions, thereby allowing for the truncation of less relevant dimensions. Then, Vector Quantization is used to show each dimension of the reduced-dimensional embedding as a binary value. This binary representation significantly reduces the memory usage of embeddings while retaining semantic information. The evaluation of Binary MRL on various datasets demonstrates that the method can achieve over 90% of the performance of the original model while using significantly smaller embeddings.

In conclusion, Binary MRL represents a novel approach to addressing the scalability challenges of embeddings in NLP applications. By combining techniques from MRL and Vector Quantization, Binary MRL achieves significant compression of embeddings while preserving their utility and effectiveness. Not only does this method reduce the costs of large-scale retrieval, but it also makes new tasks possible that were not possible before because of memory limits.

Follow-up on binary embeddings: 64 bytes per embedding, yee-haw 🤠

Reduces memory usage of our embedding model by more than 98% (64x) while retaining over 90% of model performance with binary 🪆

Model: https://t.co/ZlbEJf3DKi
Blog: https://t.co/ZaalEm0U92

— mixedbreadai (@mixedbreadai) April 12, 2024


YOU MAY ALSO LIKE

AI Safety Can’t Rely on an Honor Code – Unite.AI

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

Pragati Jhunjhunwala is a consulting intern at MarktechPost. She is currently pursuing her B.Tech from the Indian Institute of Technology(IIT), Kharagpur. She is a tech enthusiast and has a keen interest in the scope of software and data science applications. She is always reading about the developments in different field of AI and ML.


🐝 Join the Fastest Growing AI Research Newsletter Read by Researchers from Google + NVIDIA + Meta + Stanford + MIT + Microsoft and many others…


Credit: Source link

ShareTweetSendSharePin

Related Posts

AI Safety Can’t Rely on an Honor Code – Unite.AI
AI & Technology

AI Safety Can’t Rely on an Honor Code – Unite.AI

September 16, 2026
Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI
AI & Technology

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

September 16, 2026
MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down
AI & Technology

MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down

September 16, 2026
The Boox Note Air6C E Ink Tablet Flips Pages Nearly 40 Percent Faster
AI & Technology

The Boox Note Air6C E Ink Tablet Flips Pages Nearly 40 Percent Faster

September 16, 2026
Next Post
1st Amendment lawyer says TikTok ban would be unconstitutional

1st Amendment lawyer says TikTok ban would be unconstitutional

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Microsoft’s Data Center Plans Face Big Costs

Microsoft’s Data Center Plans Face Big Costs

September 12, 2026
Emmy awards 2026 live: the red carpet, the winners, the losers, the speeches – The Guardian

Emmy awards 2026 live: the red carpet, the winners, the losers, the speeches – The Guardian

September 15, 2026
Should You Buy Chip Stocks or Software Stocks Ross Gerber Plays This or That

Should You Buy Chip Stocks or Software Stocks Ross Gerber Plays This or That

September 10, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!