• bitcoinBitcoin(BTC)$76,033.000.17%
  • ethereumEthereum(ETH)$2,406.870.10%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$718.870.16%
  • rippleXRP(XRP)$1.30-0.09%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$98.270.63%
  • tronTRON(TRX)$0.3354080.80%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-3.31%
  • zcashZcash(ZEC)$1,291.1815.19%
  • HyperliquidHyperliquid(HYPE)$78.742.55%
  • dogecoinDogecoin(DOGE)$0.080367-0.45%
  • USDSUSDS(USDS)$1.000.02%
  • moneroMonero(XMR)$496.64-0.85%
  • whitebitWhiteBIT Coin(WBT)$78.060.03%
  • RainRain(RAIN)$0.012744-10.17%
  • chainlinkChainlink(LINK)$10.93-1.09%
  • leo-tokenLEO Token(LEO)$8.85-0.33%
  • cardanoCardano(ADA)$0.194102-1.62%
  • stellarStellar(XLM)$0.1807620.60%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • daiDai(DAI)$1.000.02%
  • bitcoin-cashBitcoin Cash(BCH)$218.590.52%
  • USD1USD1(USD1)$1.00-0.02%
  • uniswapUniswap(UNI)$6.380.16%
  • litecoinLitecoin(LTC)$50.89-1.21%
  • CantonCanton(CC)$0.0936041.60%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.30-1.39%
  • nearNEAR Protocol(NEAR)$2.5711.04%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • avalanche-2Avalanche(AVAX)$7.330.14%
  • hedera-hashgraphHedera(HBAR)$0.073372-2.53%
  • suiSui(SUI)$0.712.59%
  • shiba-inuShiba Inu(SHIB)$0.000005-2.98%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.055922-0.06%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,270.67-0.56%
  • MemeCoreMemeCore(M)$1.131.31%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$218.90-0.74%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$110.460.81%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.07%
  • BitwayBitway(BTW)$0.745.95%
  • AsterAster(ASTER)$0.691.70%
  • pax-goldPAX Gold(PAXG)$4,272.62-0.59%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0570040.04%
  • aaveAave(AAVE)$116.70-5.37%
  • mantleMantle(MNT)$0.540.89%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

FAMO: A Fast Optimization Method for Multitask Learning (MTL) that Mitigates the Conflicting Gradients using O(1) Space and Time

May 5, 2024
in AI & Technology
Reading Time: 5 mins read
A A
FAMO: A Fast Optimization Method for Multitask Learning (MTL) that Mitigates the Conflicting Gradients using O(1) Space and Time
ShareShareShareShareShare

Multitask learning (MLT) involves training a single model to perform multiple tasks simultaneously, leveraging shared information to enhance performance. While beneficial, MLT poses challenges in managing large models and optimizing across tasks. Optimizing the average loss may lead to suboptimal performance if tasks progress unevenly. Balancing task performance and optimization strategies is critical for effective MLT.

Existing solutions for mitigating the under-optimization problem in multitask learning involve gradient manipulation technics. These methods compute a new update vector to the average loss, ensuring that all task losses decrease more evenly. However, while these approaches show improved performance, they can become computationally expensive with many tasks and model size. This is due to the need to compute and store all task gradients at each iteration, resulting in significant space and time complexities. In contrast, computing the average gradient is more efficient, requiring less computational overhead per iteration.

To overcome these limitations, a research team from The University of Texas at Austin, Salesforce AI Research, and Sony AI recently published a new paper. In their work, they introduced Fast Adaptive Multitask Optimization (FAMO), a method designed to address the under-optimization issue in multitask learning without the computational burden associated with existing gradient manipulation techniques. 

FAMO dynamically adjusts task weights to ensure a balanced loss decrease across tasks, leveraging loss history instead of computing all task gradients. Key contributions include introducing FAMO, an MTL optimizer with O(1) space and time complexity per iteration, and demonstrating its comparable or superior performance to existing methods across various MTL benchmarks, with significant computational efficiency improvements. 

The proposed approach comprises two main ideas: achieving a balanced loss decrease across tasks and amortizing computation over time.

  1. Balanced Rate of Loss Improvement:
  • FAMO aims to decrease all task losses at an equal rate as much as possible. It defines the rate of improvement for each task based on the change in loss over time.
  • By formulating an optimization problem, FAMO seeks an update direction that maximizes the worst-case improvement rate across all tasks.
  1. Fast Approximation by Amortizing over Time:
    • Instead of solving the optimization problem at each step, FAMO performs a single-step gradient descent on a parameter representing task weights, amortizing computation over the optimization trajectory.
    • This is achieved by updating the task weights based on the change in log losses and approximating the gradient.

Practically, FAMO reparameterizes the task weights to ensure they stay within a valid range and introduces regularization to focus more on recent updates. The algorithm iteratively updates task weights and parameters based on the observed losses to find a balance between task performance and computational efficiency.

Overall, FAMO offers a computationally efficient approach to multitask optimization by dynamically adjusting task weights and amortizing computation over time. This leads to improved performance without the need for extensive gradient computations.

To evaluate Famo, the authors conducted empirical experiments across various experiment settings. They started with a toy 2-task problem, demonstrating Famo’s ability to efficiently mitigate conflicting gradients (CG). Compared to state-of-the-art methods in MLT supervised and reinforcement learning benchmarks, Famo consistently performed well. It showcased significant efficiency improvements, particularly in training time, compared to methods like NASHMTL. Additionally, an ablation study on the regularization coefficient γ highlighted Famo’s robustness across different settings, except for specific cases like CityScapes, where tuning γ could stabilize performance. The evaluation emphasized Famo’s effectiveness and efficiency across diverse multitask learning scenarios.

In conclusion, FAMO presents a promising solution to the challenges of MLT by dynamically adjusting task weights and amortizing computation over time. The method effectively mitigates under-optimization issues without the computational burden associated with existing gradient manipulation techniques. Through empirical experiments, FAMO demonstrated consistent performance improvements across various MLT scenarios, showcasing its effectiveness and efficiency. With its balanced loss decrease approach and efficient optimization strategy, FAMO offers a valuable contribution to the field of multitask learning, paving the way for more scalable and effective machine learning models.


Check out the Paper. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter. Join our Telegram Channel, Discord Channel, and LinkedIn Group.

If you like our work, you will love our newsletter..

Don’t Forget to join our 41k+ ML SubReddit


YOU MAY ALSO LIKE

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down

Mahmoud is a PhD researcher in machine learning. He also holds a
bachelor’s degree in physical science and a master’s degree in
telecommunications and networking systems. His current areas of
research concern computer vision, stock market prediction and deep
learning. He produced several scientific articles about person re-
identification and the study of the robustness and stability of deep
networks.


✅ [FREE AI WEBINAR Alert] Live RAG Comparison Test: Pinecone vs Mongo vs Postgres vs SingleStore: May 9, 2024 10:00am – 11:00am PDT


Credit: Source link

ShareTweetSendSharePin

Related Posts

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI
AI & Technology

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

September 16, 2026
MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down
AI & Technology

MindsEye Developer Build A Rocket Boy Is Reportedly Shutting Down

September 16, 2026
NVIDIA Vera Rubin NVL72 Posts First MLPerf Inference Preview Results – Unite.AI
AI & Technology

NVIDIA Vera Rubin NVL72 Posts First MLPerf Inference Preview Results – Unite.AI

September 16, 2026
Samsung Brings One UI 9 To The Rest Of The Galaxy S26 Series
AI & Technology

Samsung Brings One UI 9 To The Rest Of The Galaxy S26 Series

September 16, 2026
Next Post
Family friend charged with capital murder in Audrii Cunningham case

Family friend charged with capital murder in Audrii Cunningham case

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Kornacki: ‘A very tall order’ for Republicans to win New Hampshire Senate seat

Kornacki: ‘A very tall order’ for Republicans to win New Hampshire Senate seat

September 15, 2026
Powerful waves from Hurricane Marie slam Southern California

Powerful waves from Hurricane Marie slam Southern California

September 16, 2026
Here’s the investing playbook if the Federal Reserve hikes rates on Wednesday

Here’s the investing playbook if the Federal Reserve hikes rates on Wednesday

September 14, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!