• bitcoinBitcoin(BTC)$77,351.000.16%
  • ethereumEthereum(ETH)$2,539.053.18%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$726.601.69%
  • rippleXRP(XRP)$1.360.71%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$102.812.78%
  • tronTRON(TRX)$0.338410-0.19%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-0.02%
  • zcashZcash(ZEC)$1,186.854.77%
  • HyperliquidHyperliquid(HYPE)$80.790.44%
  • dogecoinDogecoin(DOGE)$0.0845750.51%
  • RainRain(RAIN)$0.015594-1.62%
  • moneroMonero(XMR)$520.552.03%
  • USDSUSDS(USDS)$1.000.01%
  • whitebitWhiteBIT Coin(WBT)$80.420.62%
  • chainlinkChainlink(LINK)$11.630.34%
  • leo-tokenLEO Token(LEO)$9.15-0.40%
  • cardanoCardano(ADA)$0.206740-1.11%
  • stellarStellar(XLM)$0.1792590.99%
  • Ethena USDeEthena USDe(USDE)$1.000.02%
  • bitcoin-cashBitcoin Cash(BCH)$229.571.49%
  • daiDai(DAI)$1.000.02%
  • USD1USD1(USD1)$1.000.04%
  • litecoinLitecoin(LTC)$53.792.70%
  • CantonCanton(CC)$0.097921-1.12%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.370.74%
  • uniswapUniswap(UNI)$6.080.62%
  • avalanche-2Avalanche(AVAX)$7.48-1.55%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • hedera-hashgraphHedera(HBAR)$0.074575-1.59%
  • nearNEAR Protocol(NEAR)$2.50-1.13%
  • shiba-inuShiba Inu(SHIB)$0.0000051.58%
  • suiSui(SUI)$0.73-1.39%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.056343-0.35%
  • MemeCoreMemeCore(M)$1.192.54%
  • tether-goldTether Gold(XAUT)$4,346.550.45%
  • Circle USYCCircle USYC(USYC)$1.140.03%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • okbOKB(OKB)$113.451.94%
  • BittensorBittensor(TAO)$236.18-1.79%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.14%
  • aaveAave(AAVE)$125.602.38%
  • mantleMantle(MNT)$0.581.85%
  • pax-goldPAX Gold(PAXG)$4,353.730.59%
  • AsterAster(ASTER)$0.68-3.36%
  • polkadotPolkadot(DOT)$1.05-5.93%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.054531-3.36%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Adobe Researchers Propose DMV3D: A Novel 3D Generation Approach that Uses a Transformer-based 3D Large Reconstruction Model to Denoise Multi-View Diffusion

December 7, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Adobe Researchers Propose DMV3D: A Novel 3D Generation Approach that Uses a Transformer-based 3D Large Reconstruction Model to Denoise Multi-View Diffusion
ShareShareShareShareShare

A common challenge in 3D asset creation for Augmented Reality (AR), Virtual Reality (VR), robotics, and gaming has emerged. The surge in the popularity of 3D diffusion models, which simplify the complex 3D asset creation process, comes with a hitch. These models require access to ground-truth 3D models or point clouds for training, which can be challenging for real images. Moreover, the latent 3D diffusion approach often results in a complex and challenging-to-denoise latent space on diverse 3D datasets, making high-quality rendering a hurdle.

Some existing solutions tackle this challenge but often demand a lot of manual work and optimization processes. A team of researchers from Adobe Research and Stanford have been working to make the 3D generation process faster, more realistic, and more generic. A recent paper introduces a new approach called DMV3D, a single-stage category-agnostic diffusion model. This model can generate 3D Neural Radiance Fields (NeRFs) from either text or a single-image input condition through direct model inference, significantly cutting down the time needed to create 3D objects.

The critical contributions of DMV3D include a pioneering single-stage diffusion framework using a multi-view 2D image diffusion model for 3D generation. They also introduced a Large Reconstruction Model (LRM), a multi-view denoiser that reconstructs noise-free triplane NeRFs from noisy multi-view images. The model provides a general probabilistic approach for high-quality text-to-3D generation and single-image reconstruction, achieving fast direct model inference, taking only about 30 seconds on a single A100 GPU.

DMV3D integrates 3D NeRF reconstruction and rendering into its denoiser, creating a 2D multi-view image diffusion model trained without direct 3D supervision. This eliminates the need for separately training 3D NeRF encoders for latent-space diffusion and streamlines the per-asset optimization process. The researchers strategically use a sparse set of four multi-view images surrounding an object, effectively describing a 3D object without significant self-occlusions.

Leveraging large transformer models, the researchers address the challenging task of sparse-view 3D reconstruction. Built upon the recent 3D Large Reconstruction Model (LRM), they introduce a novel joint reconstruction and denoising model capable of handling various noise levels in the diffusion process. This model integrates as the multi-view image denoiser in a multi-view image diffusion framework.

Trained on large-scale datasets comprising synthetic renderings and real captures, DMV3D demonstrates the ability to generate single-stage 3D in approximately 30 seconds on a single A100 GPU. It achieves state-of-the-art results in single-image 3D reconstruction. This work provides a fresh perspective on addressing 3D generation tasks by bridging the realms of 2D and 3D generative models, unifying 3D reconstruction and generation. The implications extend beyond immediate applications, opening doors for developing foundational models to tackle various challenges in 3D vision and graphics.


Check out the Paper and Project. All credit for this research goes to the researchers of this project. Also, don’t forget to join our 33k+ ML SubReddit, 41k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

Lenovo’s Googlebook 15 Seems Decidedly Premium Based On A New Leak

New Images Show A Detailed View Of Meta’s Upcoming Mixed Reality Headset

Niharika is a Technical consulting intern at Marktechpost. She is a third year undergraduate, currently pursuing her B.Tech from Indian Institute of Technology(IIT), Kharagpur. She is a highly enthusiastic individual with a keen interest in Machine learning, Data science and AI and an avid reader of the latest developments in these fields.


✅ [Featured AI Model] Check out LLMWare and It’s RAG- specialized 7B Parameter LLMs

Credit: Source link

ShareTweetSendSharePin

Related Posts

Lenovo’s Googlebook 15 Seems Decidedly Premium Based On A New Leak
AI & Technology

Lenovo’s Googlebook 15 Seems Decidedly Premium Based On A New Leak

September 11, 2026
New Images Show A Detailed View Of Meta’s Upcoming Mixed Reality Headset
AI & Technology

New Images Show A Detailed View Of Meta’s Upcoming Mixed Reality Headset

September 11, 2026
Dzmitry Lazerka, Co-Founder of VictoriaMetrics – Interview Series – Unite.AI
AI & Technology

Dzmitry Lazerka, Co-Founder of VictoriaMetrics – Interview Series – Unite.AI

September 11, 2026
Where Should Apple Go After The iPhone Duo? Bring On Smaller And Larger Foldables
AI & Technology

Where Should Apple Go After The iPhone Duo? Bring On Smaller And Larger Foldables

September 11, 2026
Next Post
GTA 6 Trailer, ByteDance Deal | Bloomberg Technology

GTA 6 Trailer, ByteDance Deal | Bloomberg Technology

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Rep. Jim Jordan formally asks DOJ to prosecute Jack Smith

Rep. Jim Jordan formally asks DOJ to prosecute Jack Smith

September 6, 2026
Federal program cuts spark concern over FDA and CDC response to cyclospora

Federal program cuts spark concern over FDA and CDC response to cyclospora

September 5, 2026
DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

September 10, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!