• bitcoinBitcoin(BTC)$85,635.00-0.36%
  • ethereumEthereum(ETH)$2,726.97-0.48%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$782.52-0.50%
  • rippleXRP(XRP)$1.572.36%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$117.08-0.16%
  • tronTRON(TRX)$0.342985-0.86%
  • zcashZcash(ZEC)$1,620.166.01%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.031.75%
  • HyperliquidHyperliquid(HYPE)$95.27-0.11%
  • dogecoinDogecoin(DOGE)$0.0994011.57%
  • moneroMonero(XMR)$561.09-1.80%
  • whitebitWhiteBIT Coin(WBT)$85.98-0.49%
  • USDSUSDS(USDS)$1.000.01%
  • chainlinkChainlink(LINK)$12.76-1.48%
  • cardanoCardano(ADA)$0.2510012.01%
  • RainRain(RAIN)$0.012855-4.43%
  • leo-tokenLEO Token(LEO)$8.96-0.25%
  • stellarStellar(XLM)$0.2144101.64%
  • bitcoin-cashBitcoin Cash(BCH)$348.1628.49%
  • nearNEAR Protocol(NEAR)$4.712.93%
  • uniswapUniswap(UNI)$9.7411.99%
  • avalanche-2Avalanche(AVAX)$11.142.62%
  • Ethena USDeEthena USDe(USDE)$1.000.01%
  • litecoinLitecoin(LTC)$62.473.44%
  • daiDai(DAI)$1.00-0.01%
  • CantonCanton(CC)$0.112340-5.03%
  • USD1USD1(USD1)$1.000.00%
  • hedera-hashgraphHedera(HBAR)$0.0954001.53%
  • suiSui(SUI)$1.01-0.30%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.441.78%
  • shiba-inuShiba Inu(SHIB)$0.0000061.00%
  • BittensorBittensor(TAO)$304.53-5.55%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0654250.39%
  • MemeCoreMemeCore(M)$1.27-3.41%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,312.17-0.29%
  • okbOKB(OKB)$121.88-0.26%
  • BitwayBitway(BTW)$0.9410.00%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • aaveAave(AAVE)$148.264.97%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.33%
  • mantleMantle(MNT)$0.682.70%
  • EthenaEthena(ENA)$0.2113881.07%
  • OndoOndo(ONDO)$0.4346791.14%
  • pepePepe(PEPE)$0.000005-3.98%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

SaRA: A Memory-Efficient Fine-Tuning Method for Enhancing Pre-Trained Diffusion Models

September 15, 2024
in AI & Technology
Reading Time: 4 mins read
A A
SaRA: A Memory-Efficient Fine-Tuning Method for Enhancing Pre-Trained Diffusion Models
ShareShareShareShareShare

Recent advancements in diffusion models have significantly improved tasks like image, video, and 3D generation, with pre-trained models like Stable Diffusion being pivotal. However, adapting these models to new tasks efficiently remains a challenge. Existing fine-tuning approaches—Additive, Reparameterized, and Selective-based—have limitations, such as added latency, overfitting, or complex parameter selection. A proposed solution involves leveraging “temporarily ineffective” parameters—those with minimal current impact but the potential to learn new information—by reactivating them to enhance the model’s generative capabilities without the drawbacks of existing methods.

Researchers from Shanghai Jiao Tong University and Youtu Lab, Tencent, propose SaRA, a fine-tuning method for pre-trained diffusion models. Inspired by model pruning, SaRA reuses “temporarily ineffective” parameters with small absolute values by optimizing them using sparse matrices while preserving prior knowledge. They employ a nuclear-norm-based low-rank training scheme and a progressive parameter adjustment strategy to prevent overfitting. SaRA’s memory-efficient nonstructural backpropagation reduces memory costs by 40% compared to LoRA. Experiments on Stable Diffusion models show SaRA’s superior performance across various tasks, requiring only a single line of code modification for implementation.

YOU MAY ALSO LIKE

Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model

Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning

Diffusion models, such as Stable Diffusion, excel in image generation tasks but are limited by their large parameter sizes, making full fine-tuning challenging. Methods like ControlNet, LoRA, and DreamBooth address this by adding external networks or fine-tuning to enable controlled generation or adaptation to new tasks. Parameter-efficient fine-tuning approaches like Addictive Fine-Tuning (AFT) and Reparameterized Fine-Tuning (RFT) introduce low-rank matrices or adapters. At the same time, Selective Fine-Tuning (SFT) focuses on modifying specific parameters. SaRA improves on these methods by reusing ineffective parameters, maintaining model architecture, reducing memory costs, and enhancing fine-tuning efficiency without additional inference latency.

In diffusion models, “ineffective” parameters, identified by their small absolute values, show minimal impact on performance when pruned. Experiments on Stable Diffusion models (v1.4, v1.5, v2.0, v3.0) revealed that setting parameters below a certain threshold to zero sometimes even improves generative tasks. The ineffectiveness is due to optimization randomness, not model structure. Fine-tuning can make these parameters effective again. SaRA, a method, leverages these temporarily ineffective parameters for fine-tuning, using low-rank constraints and progressive adjustment to prevent overfitting and enhance efficiency, significantly reducing memory and computation costs compared to existing methods like LoRA.

The proposed method was evaluated on tasks like backbone fine-tuning, image customization, and video generation using FID, CLIP score, and VLHI metrics. It outperformed existing fine-tuning approaches (LoRA, AdaptFormer, LT-SFT) across datasets, showing superior task-specific learning and prior preservation. Image and video generation achieved better consistency and avoided artifacts. The method also reduced memory usage and training time by over 45%. Ablation studies highlighted the importance of progressive parameter adjustment and low-rank constraints. Correlation analysis revealed more effective knowledge acquisition than other methods, enhancing task performance.

SaRA is a parameter-efficient fine-tuning method that leverages the least impactful parameters in pre-trained models. By utilizing a nuclear norm-based low-rank loss, SaRA prevents overfitting, while its progressive parameter adjustment enhances fine-tuning effectiveness. The unstructured backpropagation reduces memory costs, benefiting other selective fine-tuning methods. SaRA significantly improves generative capabilities in tasks like domain transfer and image editing, outperforming methods like LoRA. It requires only a one-line code modification for easy integration, demonstrating superior performance on models such as Stable Diffusion 1.5, 2.0, and 3.0 across multiple applications.


Check out the Model. All credit for this research goes to the researchers of this project. Also, don’t forget to follow us on Twitter and join our Telegram Channel and LinkedIn Group. If you like our work, you will love our newsletter..

Don’t Forget to join our 50k+ ML SubReddit

⏩ ⏩ FREE AI WEBINAR: ‘SAM 2 for Video: How to Fine-tune On Your Data’ (Wed, Sep 25, 4:00 AM – 4:45 AM EST)


Sana Hassan, a consulting intern at Marktechpost and dual-degree student at IIT Madras, is passionate about applying technology and AI to address real-world challenges. With a keen interest in solving practical problems, he brings a fresh perspective to the intersection of AI and real-life solutions.

⏩ ⏩ FREE AI WEBINAR: ‘SAM 2 for Video: How to Fine-tune On Your Data’ (Wed, Sep 25, 4:00 AM – 4:45 AM EST)


Credit: Source link

ShareTweetSendSharePin

Related Posts

Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model
AI & Technology

Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model

September 23, 2026
Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning
AI & Technology

Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning

September 23, 2026
OpenAI Releases GPT-6 Sol and Luna: 50% Cheaper API Pricing and Benchmarks
AI & Technology

OpenAI Releases GPT-6 Sol and Luna: 50% Cheaper API Pricing and Benchmarks

September 23, 2026
The Pros And Cons Of Using A Password Manager Over An Authenticator App
AI & Technology

The Pros And Cons Of Using A Password Manager Over An Authenticator App

September 23, 2026
Next Post
LLaMA-Omni: A Novel AI Model Architecture Designed for Low-Latency and High-Quality Speech Interaction with LLMs

LLaMA-Omni: A Novel AI Model Architecture Designed for Low-Latency and High-Quality Speech Interaction with LLMs

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
NBC Nightly News with Tom Llamas Full Episode – Sept. 3

NBC Nightly News with Tom Llamas Full Episode – Sept. 3

September 17, 2026
Fighting Over The Cost Of Freezer Meals?

Fighting Over The Cost Of Freezer Meals?

September 22, 2026
Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

Denise Ruffner, VP Business Development and Commercial Operations Worldwide, Haiqu – Interview Series – Unite.AI

September 16, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!