• bitcoinBitcoin(BTC)$77,669.00-1.82%
  • ethereumEthereum(ETH)$2,419.64-2.43%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$688.57-0.68%
  • rippleXRP(XRP)$1.35-3.04%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$100.19-3.87%
  • tronTRON(TRX)$0.322710-2.74%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.054.38%
  • HyperliquidHyperliquid(HYPE)$83.46-0.81%
  • zcashZcash(ZEC)$837.24-3.19%
  • dogecoinDogecoin(DOGE)$0.081875-2.26%
  • RainRain(RAIN)$0.0172352.89%
  • USDSUSDS(USDS)$1.00-0.01%
  • moneroMonero(XMR)$519.710.87%
  • leo-tokenLEO Token(LEO)$9.25-1.27%
  • whitebitWhiteBIT Coin(WBT)$71.40-2.00%
  • chainlinkChainlink(LINK)$11.26-2.14%
  • cardanoCardano(ADA)$0.198521-2.05%
  • stellarStellar(XLM)$0.175886-1.75%
  • bitcoin-cashBitcoin Cash(BCH)$248.79-0.76%
  • daiDai(DAI)$1.000.00%
  • CantonCanton(CC)$0.113958-7.79%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.01%
  • uniswapUniswap(UNI)$6.2211.82%
  • litecoinLitecoin(LTC)$49.410.67%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.32-5.42%
  • hedera-hashgraphHedera(HBAR)$0.074195-1.23%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • avalanche-2Avalanche(AVAX)$7.25-0.92%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.24%
  • suiSui(SUI)$0.73-1.03%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • crypto-com-chainCronos(CRO)$0.055496-1.92%
  • tether-goldTether Gold(XAUT)$4,323.88-2.52%
  • nearNEAR Protocol(NEAR)$1.88-5.31%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • MemeCoreMemeCore(M)$1.05-3.76%
  • okbOKB(OKB)$110.10-1.97%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.01%
  • BittensorBittensor(TAO)$221.20-4.10%
  • aaveAave(AAVE)$134.886.10%
  • AsterAster(ASTER)$0.700.44%
  • pax-goldPAX Gold(PAXG)$4,331.26-2.51%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0573480.02%
  • MorphoMorpho(MORPHO)$2.622.97%
  • mantleMantle(MNT)$0.54-0.52%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Meet StyleAvatar3D: A New AI Method for Generating Stylized 3D Avatars Using Image-Text Diffusion Models and a GAN-based 3D Generation Network

June 2, 2023
in AI & Technology
Reading Time: 5 mins read
A A
Meet StyleAvatar3D: A New AI Method for Generating Stylized 3D Avatars Using Image-Text Diffusion Models and a GAN-based 3D Generation Network
ShareShareShareShareShare

Since the advent of large-scale image-text pairings and sophisticated generative model topologies like diffusion models, generative models have made tremendous progress in producing high-fidelity 2D pictures. These models eliminate manual involvement by allowing users to create realistic visuals from text cues. Due to the lack of diversity and accessibility of 3D learning models compared to their 2D counterparts, 3D generative models continue to confront significant problems. The availability of high-quality 3D models is constrained by the arduous and highly specialized manual development of 3D assets in software engines. 

Researchers have lately investigated pre-trained image-text generative methods for creating high-fidelity 3D models to address this issue. These models include detailed priors of item geometry and appearance, which may make it easier to create realistic and varied 3D models. In this study researchers from Tencent, Nanyang Technological University, Fudan University and  Zhejiang University present a unique method for creating 3D-styled avatars that use text-to-image diffusion models that have already undergone training and allow users to choose avatars’ styles and facial features via text prompts. They use EG3D, a GAN-based 3D generation network, specifically because it has several benefits. 

First, EG3D uses calibrated photos rather than 3D data for training, making it possible to continuously increase the variety and realism of 3D models using improved image data. This feat is quite simple for 2D photographs. Second, they can produce each view independently, effectively controlling the randomness during picture formation because the images used for training do not require stringent multi-view uniformity in appearance. Their method uses ControlNet based upon StableDiffusion, which permits picture production directed by predetermined postures, to create calibrated 2D training images for training EG3D. 

🚀 JOIN the fastest ML Subreddit Community

Reusing camera characteristics from posture photographs for learning purposes enables these poses to be synthesized or retrieved from avatars in current engines. Even when utilizing accurate stance photographs as guidance, ControlNet frequently struggles to create views with enormous angles, such as the back of the head. The generation of complete 3D models needs to be improved by these failed outputs. They have taken two separate approaches to the problem to address it. First, they have created view-specific prompts for various views during picture production to reduce failure occurrences dramatically. The synthesized photos might partially match the stance photographs, even with view-specific cues. 

To address this mismatch, they have created a coarse-to-fine discriminator for 3D GAN training. Each picture data in their system has a coarse and fine posture annotation. They select a training annotation at random during GAN training. They give a high chance of adopting good posture annotation for confident views like the front face, but learning for the rest of the opinions relies more heavily on coarse ideas. This method can produce more accurate and varied 3D models even when the input photos include cluttered annotations. Additionally, they have created a latent diffusion model in the latent style space of StyleGAN to enable conditional 3D creation using an image input. 

The diffusion model can be trained quickly because of the style code’s low dimensions, great expressiveness, and compactness. They directly sample image and style code pairings from their trained 3D generators to learn the diffusion model. They ran comprehensive tests on many massive datasets to gauge the efficacy of their suggested strategy. Their findings show that their method exceeds current cutting-edge techniques regarding visual quality and variety. In conclusion, this research introduces a unique method that uses trained image-text diffusion models to produce high-fidelity 3D avatars. 

Their architecture considerably increases the versatility of avatar production by allowing styles and facial features to be determined by text prompts. To address the issue of picture-position misalignment, they have also suggested a coarse-to-fine pose-aware discriminator, which will allow for better use of image data with erroneous pose annotations. Last but not least, they have created an additional conditional generation module that enables conditional 3D creation using picture input in the latent style space. This module further increases the framework’s adaptability and allows users to create 3D models that are customized to their tastes. They also plan to open-source their code. 


Check Out The Paper and Github link. Don’t forget to join our 22k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]

🚀 Check Out 100’s AI Tools in AI Tools Club


YOU MAY ALSO LIKE

This Is The Best Setting And Placement For Your Dolby Atmos Soundbar

Aramco Digital and Avathon Partner on Autonomous Operations AI – Unite.AI

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


➡️ Ultimate Guide to Data Labeling in Machine Learning

Credit: Source link

ShareTweetSendSharePin

Related Posts

This Is The Best Setting And Placement For Your Dolby Atmos Soundbar
AI & Technology

This Is The Best Setting And Placement For Your Dolby Atmos Soundbar

September 2, 2026
Aramco Digital and Avathon Partner on Autonomous Operations AI – Unite.AI
AI & Technology

Aramco Digital and Avathon Partner on Autonomous Operations AI – Unite.AI

September 1, 2026
Anthropic Debuts Claude Fable 5.1 and Mythos 5.1 With Split Safeguards – Unite.AI
AI & Technology

Anthropic Debuts Claude Fable 5.1 and Mythos 5.1 With Split Safeguards – Unite.AI

September 1, 2026
Anthropic Releases Claude Fable 5.1 and Claude Mythos 5.1: 52.6% on Terminal-Bench-Science and 75% Cheaper Cache Reads
AI & Technology

Anthropic Releases Claude Fable 5.1 and Claude Mythos 5.1: 52.6% on Terminal-Bench-Science and 75% Cheaper Cache Reads

September 1, 2026
Next Post
AI blamed for 3,900 people losing their jobs in May: report

AI blamed for 3,900 people losing their jobs in May: report

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Jackson Hole 2026: Discipline, Not Decisions

Jackson Hole 2026: Discipline, Not Decisions

September 1, 2026
Colombia struck with 7.4 magnitude earthquake, killing over 100

Colombia struck with 7.4 magnitude earthquake, killing over 100

August 26, 2026
Death toll rises in Nepal-China border disaster as lake poses new flood threat – AP News

Death toll rises in Nepal-China border disaster as lake poses new flood threat – AP News

August 28, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!