• bitcoinBitcoin(BTC)$79,651.00-0.39%
  • ethereumEthereum(ETH)$2,496.84-0.39%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$746.36-2.30%
  • rippleXRP(XRP)$1.41-1.02%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$105.01-0.85%
  • tronTRON(TRX)$0.3358100.74%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.060.00%
  • zcashZcash(ZEC)$1,189.1210.81%
  • HyperliquidHyperliquid(HYPE)$86.21-0.07%
  • dogecoinDogecoin(DOGE)$0.089505-1.61%
  • RainRain(RAIN)$0.016676-2.93%
  • moneroMonero(XMR)$534.05-3.52%
  • USDSUSDS(USDS)$1.000.00%
  • chainlinkChainlink(LINK)$13.117.75%
  • whitebitWhiteBIT Coin(WBT)$73.45-0.42%
  • leo-tokenLEO Token(LEO)$9.25-0.87%
  • cardanoCardano(ADA)$0.219465-0.81%
  • stellarStellar(XLM)$0.1895972.02%
  • bitcoin-cashBitcoin Cash(BCH)$255.86-1.75%
  • daiDai(DAI)$1.00-0.01%
  • uniswapUniswap(UNI)$7.02-2.28%
  • Ethena USDeEthena USDe(USDE)$1.00-0.02%
  • CantonCanton(CC)$0.1099850.15%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$54.21-0.53%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.42-0.15%
  • hedera-hashgraphHedera(HBAR)$0.080799-0.96%
  • avalanche-2Avalanche(AVAX)$7.801.70%
  • suiSui(SUI)$0.80-0.12%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.70%
  • nearNEAR Protocol(NEAR)$2.419.14%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0575280.73%
  • tether-goldTether Gold(XAUT)$4,400.32-0.59%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.152.93%
  • BittensorBittensor(TAO)$271.1313.98%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$113.27-1.93%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.35%
  • mantleMantle(MNT)$0.6510.53%
  • AsterAster(ASTER)$0.79-0.10%
  • aaveAave(AAVE)$133.82-0.77%
  • pax-goldPAX Gold(PAXG)$4,403.10-0.66%
  • OndoOndo(ONDO)$0.3837832.38%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056621-0.72%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Microsoft Researchers Propose BioViL-T: A Novel Self-Supervised Framework Ushering in Enhanced Predictive Performance and Data Efficiency in Biomedical Applications

June 18, 2023
in AI & Technology
Reading Time: 5 mins read
A A
Microsoft Researchers Propose BioViL-T: A Novel Self-Supervised Framework Ushering in Enhanced Predictive Performance and Data Efficiency in Biomedical Applications
ShareShareShareShareShare

Artificial Intelligence (AI) has emerged as a significant disruptive force across numerous industries, from how technological businesses operate to how innovation is unlocked in different subdomains in the healthcare sector. In particular, the biomedical field has witnessed significant advancements and transformation with the introduction of AI. One such noteworthy progress can be boiled down to using self-supervised vision-language models in radiology. Radiologists rely heavily on radiology reports to convey imaging observations and provide clinical diagnoses. It is noteworthy that prior imaging studies frequently play a key role in this decision-making process because they provide crucial context for assessing the course of illnesses and establishing suitable medication choices. However, current AI solutions in the mark cannot successfully align images with report data due to limited access to previous scans. Furthermore, these methods frequently do not consider the chronological development of illnesses or imaging findings typically present in biological datasets. This lack of contextual information poses risks in downstream applications like automated report generation, where models may generate inaccurate temporal content without access to past medical scans.

With the introduction of vision-language models, researchers aim to generate informative training signals by utilizing image-text pairs, thus, eliminating the need for manual labels. This approach enables the models to learn how to precisely identify and pinpoint discoveries in the images and establish connections with the information presented in radiology reports. Microsoft Research has continually worked to improve AI for reporting and radiography. Their prior research on multimodal self-supervised learning of radiology reports and images has produced encouraging results in identifying medical problems and localizing these findings within the images. As a contribution to this wave of research, Microsoft released BioViL-T, a self-supervised training framework that considers earlier images and reports when available during training and fine-tuning. BioViL-T achieves breakthrough results on various downstream benchmarks, such as progression classification and report creation, by utilizing the existing temporal structure present in datasets. The study will be presented at the prestigious Computer Vision and Pattern Recognition Conference (CVPR) in 2023.

The distinguishing characteristic of BioViL-T lies in its explicit consideration of previous images and reports throughout the training and fine-tuning processes rather than treating each image-report pair as a separate entity. The researchers’ rationale behind incorporating prior images and reports was primarily to maximize the utilization of available data, resulting in more comprehensive representations and enhanced performance across a broader range of tasks. BioViL-T introduces a unique CNN-Transformer multi-image encoder that is jointly trained with a text model. This novel multi-image encoder serves as the fundamental building block of the pre-training framework, addressing challenges such as the absence of previous images and pose variations in images over time.

🚀 JOIN the fastest ML Subreddit Community

A CNN and a transformer model were chosen to create the hybrid multi-image encoder to extract spatiotemporal features from image sequences. When previous images are available, the transformer is in charge of capturing patch embedding interactions across time. On the other hand, CNN is in order of giving visual token properties of individual images. This hybrid image encoder improves data efficiency, making it suitable for datasets of even smaller sizes. It efficiently captures static and temporal image characteristics, which is essential for applications like report decoding that call for dense-level visual reasoning over time. The pre-training procedure of the BioViL-T model can be divided into two main components: a multi-image encoder for extracting spatiotemporal features and a text encoder incorporating optional cross-attention with image features. These models are jointly trained using cross-modal global and local contrastive objectives. The model also utilizes multimodal fused representations obtained through cross-attention for image-guided masked language modeling., thereby effectively harnessing visual and textual information. This plays a central role in resolving ambiguities and enhancing language comprehension, which is of utmost importance for a wide range of downstream tasks.

The success of the Microsoft researchers’ strategy was aided by a variety of experimental evaluations that they conducted. The model achieves state-of-the-art performance for a variety of downstream tasks like progression categorization, phrase grounding, and report generation in single- and multi-image configurations. Additionally, it improves over previous models and yields appreciable results on tasks like disease classification and sentence similarity. Microsoft Research has made the model and source code available to the public to encourage the community to investigate their work further. A brand-new multimodal temporal benchmark dataset dubbed MS-CXR-T is also being made public by the researchers to stimulate additional research into quantifying how well vision-language representations can capture temporal semantics. 


Check Out The Paper and Microsoft Article. Don’t forget to join our 23k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]


Featured Tools From AI Tools Club

🚀 Check Out 100’s AI Tools in AI Tools Club


YOU MAY ALSO LIKE

Is It Safe To Buy A Refurbished iPhone From Walmart?

When Are Portable Apple CarPlay Screens Actually Worth It?

Khushboo Gupta is a consulting intern at MarktechPost. She is currently pursuing her B.Tech from the Indian Institute of Technology(IIT), Goa. She is passionate about the fields of Machine Learning, Natural Language Processing and Web Development. She enjoys learning more about the technical field by participating in several challenges.


➡️ Try: Ake: A Superb Residential Proxy Network (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

Is It Safe To Buy A Refurbished iPhone From Walmart?
AI & Technology

Is It Safe To Buy A Refurbished iPhone From Walmart?

September 7, 2026
When Are Portable Apple CarPlay Screens Actually Worth It?
AI & Technology

When Are Portable Apple CarPlay Screens Actually Worth It?

September 7, 2026
The Pros And Cons Of Using Wireless Android Auto
AI & Technology

The Pros And Cons Of Using Wireless Android Auto

September 6, 2026
H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
AI & Technology

H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder

September 6, 2026
Next Post
RXO: Stay Patient And Wait For The Right Time To Invest (NYSE:RXO)

RXO: Stay Patient And Wait For The Right Time To Invest (NYSE:RXO)

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Teen pleads guilty to deadly 2024 Georgia high school shooting

Teen pleads guilty to deadly 2024 Georgia high school shooting

September 5, 2026
Trump attends dignified transfer of U.S. troops

Trump attends dignified transfer of U.S. troops

September 6, 2026
Miley, Beyoncé, ADÉLA & More: New Music Friday Guide – Billboard

Miley, Beyoncé, ADÉLA & More: New Music Friday Guide – Billboard

September 4, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!