• bitcoinBitcoin(BTC)$78,416.00-0.64%
  • ethereumEthereum(ETH)$2,477.99-0.70%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$721.23-4.05%
  • rippleXRP(XRP)$1.39-2.81%
  • usd-coinUSDC(USDC)$1.00-0.02%
  • solanaSolana(SOL)$101.81-2.23%
  • tronTRON(TRX)$0.3399660.41%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.03-0.84%
  • zcashZcash(ZEC)$1,232.270.69%
  • HyperliquidHyperliquid(HYPE)$83.76-2.72%
  • dogecoinDogecoin(DOGE)$0.085794-4.81%
  • RainRain(RAIN)$0.0163511.85%
  • USDSUSDS(USDS)$1.00-0.01%
  • moneroMonero(XMR)$514.722.27%
  • whitebitWhiteBIT Coin(WBT)$80.93-0.94%
  • chainlinkChainlink(LINK)$11.84-5.20%
  • leo-tokenLEO Token(LEO)$9.190.08%
  • cardanoCardano(ADA)$0.213909-2.37%
  • stellarStellar(XLM)$0.180853-4.38%
  • bitcoin-cashBitcoin Cash(BCH)$250.61-3.13%
  • daiDai(DAI)$1.000.00%
  • Ethena USDeEthena USDe(USDE)$1.00-0.03%
  • USD1USD1(USD1)$1.00-0.02%
  • CantonCanton(CC)$0.104196-3.98%
  • litecoinLitecoin(LTC)$52.80-2.37%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.38-1.10%
  • uniswapUniswap(UNI)$6.09-11.45%
  • avalanche-2Avalanche(AVAX)$7.84-1.73%
  • hedera-hashgraphHedera(HBAR)$0.077021-2.56%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • nearNEAR Protocol(NEAR)$2.496.60%
  • suiSui(SUI)$0.77-5.41%
  • shiba-inuShiba Inu(SHIB)$0.000005-3.53%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • crypto-com-chainCronos(CRO)$0.057962-2.99%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • MemeCoreMemeCore(M)$1.233.73%
  • tether-goldTether Gold(XAUT)$4,417.600.44%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • BittensorBittensor(TAO)$254.60-0.89%
  • Ripple USDRipple USD(RLUSD)$1.00-0.02%
  • okbOKB(OKB)$113.25-0.95%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.03%
  • mantleMantle(MNT)$0.60-5.10%
  • AsterAster(ASTER)$0.73-4.21%
  • aaveAave(AAVE)$125.07-3.10%
  • pax-goldPAX Gold(PAXG)$4,422.730.46%
  • polkadotPolkadot(DOT)$1.11-6.74%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0567761.66%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from Stanford University and FAIR Meta Unveil CHOIS: A Groundbreaking AI Method for Synthesizing Realistic 3D Human-Object Interactions Guided by Language

December 11, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Researchers from Stanford University and FAIR Meta Unveil CHOIS: A Groundbreaking AI Method for Synthesizing Realistic 3D Human-Object Interactions Guided by Language
ShareShareShareShareShare

The problem of generating synchronized motions of objects and humans within a 3D scene has been addressed by researchers from Stanford University and FAIR Meta by introducing CHOIS. The system operates based on sparse object waypoints, an initial state of things and humans, and a textual description. It controls interactions between humans and objects by producing realistic and controllable motions for both entities in the specified 3D environment.

Leveraging large-scale, high-quality motion capture datasets like AMASS, interest in generative human motion modeling has risen, including action-conditioned and text-conditioned synthesis. While prior works used VAE formulations for diverse human motion generation from text, CHOIS focuses on human-object interactions. Unlike existing approaches that often center on hand motion synthesis, CHOIS considers full-body motions preceding object grasping and predicts object motion based on human movements, offering a comprehensive solution for interactive 3D scene simulations.

CHOIS addresses a critical need for synthesizing realistic human behaviors in 3D environments, crucial for computer graphics, embodied AI, and robotics. CHOIS advances the field by generating synchronized human and object motion based on language descriptions, initial states, and sparse object waypoints. It tackles challenges like realistic motion generation, accommodating environment clutter, and synthesizing interactions from language descriptions, presenting a comprehensive system for controllable human-object interactions in diverse 3D scenes.

The model uses a conditional diffusion approach to generate synchronized object and human motion based on language descriptions, object geometry, and initial states. Constraints are incorporated during the sampling process to ensure realistic human-object contact. The training phase uses a loss function to guide the model in predicting object transformations without explicitly enforcing contact constraints.

The CHOIS system is rigorously evaluated against baselines and ablations, showcasing superior performance on metrics like condition matching, contact accuracy, reduced hand-object penetration, and foot floating. On the FullBodyManipulation dataset, object geometry loss enhances the model’s capabilities. CHOIS outperforms baselines and ablations on the 3D-FUTURE dataset, demonstrating its generalization to new objects. Human perceptual studies highlight CHOIS’s better alignment with text input and superior interaction quality compared to the baseline. Quantitative metrics, including position and orientation errors, measure the deviation of generated results from ground truth motion.

In conclusion, CHOIS is a system that generates realistic human-object interactions based on language descriptions and sparse object waypoints. The procedure considers object geometry loss during training and employs effective guidance terms during sampling to enhance the realism of the results. The interaction module learned by CHOIS can be integrated into a pipeline for synthesizing long-term interactions given language and 3D scenes. CHOIS has significantly improved in generating realistic human-object interactions aligned with provided language descriptions.

Future research could explore enhancing CHOIS by integrating additional supervision, like object geometry loss, to improve the matching of generated object motion with input waypoints. Investigating advanced guidance terms for enforcing contact constraints may lead to more realistic results. Extending evaluations to diverse datasets and scenarios will test CHOIS’s generalization capabilities. Further human perceptual studies can provide deeper insights into generated interactions. Applying the learned interaction module to generate long-term interactions based on object waypoints from 3D scenes would also expand CHOIS’s applicability.


Check out the Paper and Project. All credit for this research goes to the researchers of this project. Also, don’t forget to join our 33k+ ML SubReddit, 41k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI

LandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity

Hello, My name is Adnan Hassan. I am a consulting intern at Marktechpost and soon to be a management trainee at American Express. I am currently pursuing a dual degree at the Indian Institute of Technology, Kharagpur. I am passionate about technology and want to create new products that make a difference.


🐝 [FREE AI WEBINAR] ‘Beginners Guide to LangChain: Chat with Your Multi-Model Data’ Dec 11, 2023 10 am PST

Credit: Source link

ShareTweetSendSharePin

Related Posts

Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI
AI & Technology

Anthropic Discloses Fourth Cyber Incident in Alignment Assessment – Unite.AI

September 10, 2026
LandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity
AI & Technology

LandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity

September 10, 2026
Apple Wallet Is Not The Same As Apple Pay: Here’s How They Differ
AI & Technology

Apple Wallet Is Not The Same As Apple Pay: Here’s How They Differ

September 9, 2026
Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce and Patch Vulnerabilities
AI & Technology

Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce and Patch Vulnerabilities

September 9, 2026
Next Post
This Morning’s Top Headlines – Oct. 2

This Morning’s Top Headlines – Oct. 2

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Flight passenger details moment he was partially sucked out

Flight passenger details moment he was partially sucked out

September 6, 2026
Frustrations grow over police shooting in Wisconsin

Frustrations grow over police shooting in Wisconsin

September 5, 2026
LIVE: Funeral services held for Sen. Lindsey Graham | NBC News

LIVE: Funeral services held for Sen. Lindsey Graham | NBC News

September 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!