• bitcoinBitcoin(BTC)$76,961.00-2.27%
  • ethereumEthereum(ETH)$2,438.00-2.34%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$705.98-4.71%
  • rippleXRP(XRP)$1.34-5.36%
  • usd-coinUSDC(USDC)$1.000.02%
  • solanaSolana(SOL)$99.22-3.95%
  • tronTRON(TRX)$0.338381-0.58%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.030.36%
  • zcashZcash(ZEC)$1,119.27-12.75%
  • HyperliquidHyperliquid(HYPE)$79.27-7.31%
  • dogecoinDogecoin(DOGE)$0.083074-6.81%
  • RainRain(RAIN)$0.015867-2.92%
  • USDSUSDS(USDS)$1.00-0.02%
  • moneroMonero(XMR)$503.21-0.73%
  • whitebitWhiteBIT Coin(WBT)$79.51-2.30%
  • chainlinkChainlink(LINK)$11.53-4.13%
  • leo-tokenLEO Token(LEO)$9.190.05%
  • cardanoCardano(ADA)$0.206421-5.21%
  • stellarStellar(XLM)$0.176045-4.71%
  • daiDai(DAI)$1.00-0.02%
  • bitcoin-cashBitcoin Cash(BCH)$224.89-12.98%
  • Ethena USDeEthena USDe(USDE)$1.00-0.04%
  • USD1USD1(USD1)$1.00-0.03%
  • litecoinLitecoin(LTC)$51.82-4.27%
  • CantonCanton(CC)$0.098902-5.02%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.34-3.56%
  • uniswapUniswap(UNI)$5.94-9.84%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • hedera-hashgraphHedera(HBAR)$0.074767-4.73%
  • avalanche-2Avalanche(AVAX)$7.55-5.05%
  • nearNEAR Protocol(NEAR)$2.43-6.44%
  • suiSui(SUI)$0.74-7.52%
  • shiba-inuShiba Inu(SHIB)$0.000005-6.83%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • crypto-com-chainCronos(CRO)$0.056267-5.63%
  • tether-goldTether Gold(XAUT)$4,356.53-1.08%
  • MemeCoreMemeCore(M)$1.15-1.31%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.00-0.02%
  • okbOKB(OKB)$110.38-2.60%
  • BittensorBittensor(TAO)$238.58-7.48%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.10%
  • mantleMantle(MNT)$0.57-8.10%
  • pax-goldPAX Gold(PAXG)$4,358.55-1.09%
  • AsterAster(ASTER)$0.69-6.44%
  • aaveAave(AAVE)$120.98-6.68%
  • polkadotPolkadot(DOT)$1.08-4.45%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.055248-0.49%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

3D Body Models Now Have Sound: Meta AI Introduces an Artificial Intelligence Model that can Generate Accurate 3D Spatial Audio for Full Human Bodies

November 15, 2023
in AI & Technology
Reading Time: 4 mins read
A A
3D Body Models Now Have Sound: Meta AI Introduces an Artificial Intelligence Model that can Generate Accurate 3D Spatial Audio for Full Human Bodies
ShareShareShareShareShare

The constant development of intelligent systems replicating and comprehending human behavior has led to significant advancements in the complementary fields of Computer Vision and Artificial Intelligence (AI). Machine learning models are gaining immense popularity while bridging the gap between reality and virtuality. Although 3D human body modeling has received a lot of attention in the field of computer vision, the task of modeling the acoustic side and producing 3D spatial audio from speech and body motion is still a topic of discussion. The focus has always been on the visual fidelity of artificial representations of the human body.

Human perception is multi-modal in nature as it incorporates both auditory and visual cues into the comprehension of the environment. It is essential to simulate 3D sound that corresponds with the visual picture accurately in order to create a sense of presence and immersion in a 3D world. To address these challenges, a team of researchers from Shanghai AI Laboratory and Meta Reality Labs Research has introduced a model that produces accurate 3D spatial audio representations for entire human bodies.

The team has shared that the proposed technique uses head-mounted microphones and data on human body pose to synthesize 3D spatial sound precisely. The case study focuses on a telepresence scenario combining augmented reality and virtual reality (AR/VR) in which users communicate using full-body avatars. Egocentric audio data from head-mounted microphones and body posture data that is utilized to animate the avatar have been used as examples of input. 

Current methods for sound spatialization presume that the sound source is known and that it is captured there undisturbed. The suggested approach gets around these problems by using body pose data to train a multi-modal network that distinguishes between the sources of various noises and produces precisely spatialized signals. The sound area surrounding the body is the output, and the audio from seven head-mounted microphones and the subject’s posture make up the input.

The team has conducted an empirical evaluation, demonstrating that the model can reliably produce sound fields resulting from body movements when trained with a suitable loss function. The model’s code and dataset are available for public use on the internet, promoting openness, repeatability, and additional developments in this field. The GitHub repository can be accessed at https://github.com/facebookresearch/SoundingBodies. 

The primary contributions of the work have been summarized by the team as follows. 

  1. A unique technique has been introduced that uses head-mounted microphones and body poses to render realistic 3D sound fields for human bodies. 
  1. A comprehensive empirical evaluation has been shared that highlights the importance of body pose and a well-thought-out loss function.
  1. The team has shared a new dataset they have produced that combines multi-view human body data with spatial audio recordings from a 345-microphone array. 

Check out the Paper and GitHub Page. All credit for this research goes to the researchers of this project. Also, don’t forget to join our 33k+ ML SubReddit, 41k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..

We are also on Telegram and WhatsApp.


YOU MAY ALSO LIKE

Abacus.AI Releases Three Open-Weight Smaug Models for Agentic Workloads – Unite.AI

Yoto Just Announced Two New Audio Devices For Kids

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


🔥 Join The AI Startup Newsletter To Learn About Latest AI Startups

Credit: Source link

ShareTweetSendSharePin

Related Posts

Abacus.AI Releases Three Open-Weight Smaug Models for Agentic Workloads – Unite.AI
AI & Technology

Abacus.AI Releases Three Open-Weight Smaug Models for Agentic Workloads – Unite.AI

September 10, 2026
Yoto Just Announced Two New Audio Devices For Kids
AI & Technology

Yoto Just Announced Two New Audio Devices For Kids

September 10, 2026
Salesforce Unveils Six-Capability Trusted AI Harness for Enterprises – Unite.AI
AI & Technology

Salesforce Unveils Six-Capability Trusted AI Harness for Enterprises – Unite.AI

September 10, 2026
You Can Now Plan IRL Events On Snapchat
AI & Technology

You Can Now Plan IRL Events On Snapchat

September 10, 2026
Next Post
Google AI Proposes Easy End-to-End Diffusion-based Text to Speech E3-TTS: A Simple and Efficient End-to-End Text-to-Speech Model Based on Diffusion

Google AI Proposes Easy End-to-End Diffusion-based Text to Speech E3-TTS: A Simple and Efficient End-to-End Text-to-Speech Model Based on Diffusion

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Deadly downtown Minneapolis shooting started as a custody exchange — here's what else we know – kstp.com

Deadly downtown Minneapolis shooting started as a custody exchange — here's what else we know – kstp.com

September 4, 2026
Re-Shop Your Health Plan Every Open Enrollment

Re-Shop Your Health Plan Every Open Enrollment

September 9, 2026
Kentucky governor demands ‘honesty’ on McConnell’s health

Kentucky governor demands ‘honesty’ on McConnell’s health

September 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!