• bitcoinBitcoin(BTC)$86,929.007.15%
  • ethereumEthereum(ETH)$2,792.195.95%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$803.924.66%
  • rippleXRP(XRP)$1.527.54%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$118.937.87%
  • tronTRON(TRX)$0.3443460.38%
  • zcashZcash(ZEC)$1,475.26-2.73%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.010.00%
  • HyperliquidHyperliquid(HYPE)$93.210.50%
  • dogecoinDogecoin(DOGE)$0.10012514.63%
  • moneroMonero(XMR)$586.034.11%
  • whitebitWhiteBIT Coin(WBT)$87.555.69%
  • RainRain(RAIN)$0.014007-0.70%
  • chainlinkChainlink(LINK)$13.175.04%
  • USDSUSDS(USDS)$1.000.00%
  • cardanoCardano(ADA)$0.2458868.10%
  • leo-tokenLEO Token(LEO)$8.93-0.09%
  • stellarStellar(XLM)$0.2148679.51%
  • uniswapUniswap(UNI)$8.831.62%
  • bitcoin-cashBitcoin Cash(BCH)$267.416.43%
  • nearNEAR Protocol(NEAR)$4.09-0.91%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • avalanche-2Avalanche(AVAX)$11.04-2.47%
  • litecoinLitecoin(LTC)$62.536.30%
  • CantonCanton(CC)$0.1174718.34%
  • daiDai(DAI)$1.000.03%
  • USD1USD1(USD1)$1.000.00%
  • suiSui(SUI)$1.0317.24%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.476.41%
  • hedera-hashgraphHedera(HBAR)$0.0916807.11%
  • BittensorBittensor(TAO)$310.6517.56%
  • shiba-inuShiba Inu(SHIB)$0.0000068.57%
  • MemeCoreMemeCore(M)$1.491.36%
  • crypto-com-chainCronos(CRO)$0.06587711.71%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,343.51-0.62%
  • BitwayBitway(BTW)$0.9730.10%
  • okbOKB(OKB)$123.945.03%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • aaveAave(AAVE)$147.287.37%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.48%
  • OndoOndo(ONDO)$0.4473044.48%
  • mantleMantle(MNT)$0.657.00%
  • EthenaEthena(ENA)$0.212890-2.34%
  • pepePepe(PEPE)$0.00000523.90%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

MIPRO: A Novel Optimizer that Outperforms Baselines on Five of Six Diverse Language Model LM Programs Using a Best-in-Class Open-Source Model (Llama-3-8B) by 12.9% accuracy

June 24, 2024
in AI & Technology
Reading Time: 5 mins read
A A
MIPRO: A Novel Optimizer that Outperforms Baselines on Five of Six Diverse Language Model LM Programs Using a Best-in-Class Open-Source Model (Llama-3-8B) by 12.9% accuracy
ShareShareShareShareShare

Language Models (LMs) have significantly advanced complex NLP tasks through sophisticated prompting techniques and multi-stage pipelines. However, designing these LM Programs relies heavily on manual “prompt engineering,” a time-consuming process of crafting lengthy prompts through trial and error. This approach faces challenges, particularly in multi-stage LM programs where gold labels or evaluation metrics for individual LM calls are often lacking. The absence of these metrics makes it difficult to assess and optimize each stage independently, hindering the overall efficiency and effectiveness of LM programs. As a result, there’s a pressing need for more systematic and automated approaches to optimize multi-stage LM pipelines.

Various approaches have been introduced to optimize LM programs, including gradient-guided search, reranking brute force search, evolutionary algorithms, and prompting other LMs. Some studies explored reinforcement learning for prompt optimization, focusing on word-level or phrase-level edits. Notable attempts include DSPy, which introduced a programming model for expressing and optimizing LM programs, and an approach modeling joint prompt optimization for stacked LLM calls as variational inference. However, these methods often fall short in addressing the complexities of multi-stage LM programs, particularly when dealing with arbitrary numbers of modules and diverse LM architectures. Existing approaches are limited by their focus on specific types of edits, reliance on log probabilities, or inability to optimize free-form instructions for sophisticated multi-prompt pipelines. This leaves a gap for a more flexible and comprehensive optimization approach that can handle complex, multi-stage LM pipelines without restrictive assumptions.

YOU MAY ALSO LIKE

Here’s Why Apple’s Mac Studio Has Become So Expensive

Tesla Will Soon Roll Out FSD Supervised In The Czech Republic

The researchers propose a robust approach to optimize prompts for LM programs, focusing on maximizing downstream metrics without requiring module-level labels or gradients. Their method, called MIPRO, factorizes the optimization problem into refining free-form instructions and few-shot demonstrations for each module in the LM program. MIPRO employs several innovative strategies to overccome the challenges of prompt optimization in multi-stage pipelines. These include program- and data-aware techniques for generating effective instructions, a stochastic mini-batch evaluation function to learn a surrogate model of the objective and a meta-optimization procedure that improves the LM’s proposal construction over time. This comprehensive approach enables MIPRO to navigate the complexities of credit assignment across modules and craft task-grounded instructions.

The researchers present a detailed architecture for optimizing multi-stage LM programs, MIPRO. This method focuses on optimizing free-form instructions and few-shot demonstrations for each module in the program. It addresses key challenges through several innovative strategies. For the proposal problem, it employs bootstrapping demonstrations, grounding techniques, and learning to propose. These approaches help generate task-relevant instructions and demonstrations. For credit assignment across modules, MIPRO explores greedy, surrogate, and history-based methods. The surrogate model uses a Bayesian approach to predict the quality of variable combinations, while the history-based method utilizes past evaluations to inform future proposals. It also incorporates a stochastic mini-batch evaluation function and a meta-optimization procedure to refine proposal generation over time. This comprehensive architecture enables MIPRO to efficiently navigate the complex optimization landscape of multi-stage LM programs.

The results of the MIPRO optimization approach reveal several key insights. Optimizing bootstrapped demonstrations as few-shot examples proved crucial for achieving the best performance in most tasks. MIPRO, which optimizes both instructions and few-shot examples, generally yielded the best overall performance across tasks. Instruction optimization was found to be particularly important for tasks with conditional rules that are not immediately obvious to the LM and are not easily expressed through a limited number of few-shot examples. Grounding techniques were generally helpful for instruction proposals, although the best proposal strategy varied by task. 

This study formalizes LM program optimization as a prompt search problem, addressing the challenges of proposal generation and credit assignment. By exploring various strategies for diverse tasks, the research demonstrates that optimizing few-shot demonstrations is highly effective, while instruction optimization is crucial for complex tasks. The study ultimately finds that jointly optimizing both demonstrations and instructions yields the best results, paving the way for more efficient and powerful multi-stage LM programs.


Check out the Paper. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. 

Join our Telegram Channel and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 45k+ ML SubReddit

🚀 Create, edit, and augment tabular data with the first compound AI system, Gretel Navigator, now generally available! [Advertisement]


Asjad is an intern consultant at Marktechpost. He is persuing B.Tech in mechanical engineering at the Indian Institute of Technology, Kharagpur. Asjad is a Machine learning and deep learning enthusiast who is always researching the applications of machine learning in healthcare.

[Announcing Gretel Navigator] Create, edit, and augment tabular data with the first compound AI system trusted by EY, Databricks, Google, and Microsoft


Credit: Source link

ShareTweetSendSharePin

Related Posts

Here’s Why Apple’s Mac Studio Has Become So Expensive
AI & Technology

Here’s Why Apple’s Mac Studio Has Become So Expensive

September 21, 2026
Tesla Will Soon Roll Out FSD Supervised In The Czech Republic
AI & Technology

Tesla Will Soon Roll Out FSD Supervised In The Czech Republic

September 21, 2026
Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing
AI & Technology

Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing

September 21, 2026
Collaboration Must Sit At the Heart of Manufacturing’s Multi-Agentic AI Approach. Here’s How. – Unite.AI
AI & Technology

Collaboration Must Sit At the Heart of Manufacturing’s Multi-Agentic AI Approach. Here’s How. – Unite.AI

September 21, 2026
Next Post
Recursion Pharmaceuticals: More Of A TechBio Industrializing Drug Discovery (NASDAQ:RXRX)

Recursion Pharmaceuticals: More Of A TechBio Industrializing Drug Discovery (NASDAQ:RXRX)

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Edouard weakens to a tropical depression after landfall

Edouard weakens to a tropical depression after landfall

September 19, 2026
Family of 9/11 hero honored by President Trump

Family of 9/11 hero honored by President Trump

September 15, 2026
Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model

Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model

September 19, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!