• bitcoinBitcoin(BTC)$77,264.000.17%
  • ethereumEthereum(ETH)$2,506.22-0.53%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$721.30-0.69%
  • rippleXRP(XRP)$1.36-0.61%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$100.96-0.31%
  • tronTRON(TRX)$0.3411300.43%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.00-0.31%
  • zcashZcash(ZEC)$1,089.12-2.76%
  • HyperliquidHyperliquid(HYPE)$78.12-2.15%
  • dogecoinDogecoin(DOGE)$0.084171-0.74%
  • RainRain(RAIN)$0.015301-3.20%
  • moneroMonero(XMR)$530.67-0.54%
  • USDSUSDS(USDS)$1.00-0.01%
  • whitebitWhiteBIT Coin(WBT)$80.15-0.03%
  • chainlinkChainlink(LINK)$11.41-0.61%
  • leo-tokenLEO Token(LEO)$9.06-0.54%
  • cardanoCardano(ADA)$0.2081640.26%
  • stellarStellar(XLM)$0.179265-0.40%
  • Ethena USDeEthena USDe(USDE)$1.00-0.02%
  • daiDai(DAI)$1.00-0.01%
  • bitcoin-cashBitcoin Cash(BCH)$224.04-0.73%
  • USD1USD1(USD1)$1.000.00%
  • litecoinLitecoin(LTC)$54.862.11%
  • uniswapUniswap(UNI)$6.24-1.54%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.36-1.52%
  • CantonCanton(CC)$0.095602-1.60%
  • hedera-hashgraphHedera(HBAR)$0.0760582.15%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • avalanche-2Avalanche(AVAX)$7.410.55%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.96%
  • nearNEAR Protocol(NEAR)$2.33-0.70%
  • suiSui(SUI)$0.72-0.46%
  • crypto-com-chainCronos(CRO)$0.058109-1.31%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,346.36-0.07%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.14-3.18%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$113.57-0.11%
  • BittensorBittensor(TAO)$234.461.11%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.13%
  • aaveAave(AAVE)$126.22-0.02%
  • BitwayBitway(BTW)$0.7026.78%
  • AsterAster(ASTER)$0.701.92%
  • pax-goldPAX Gold(PAXG)$4,348.84-0.13%
  • mantleMantle(MNT)$0.56-0.78%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056937-1.56%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

This AI Paper from Microsoft Introduces a New Approach to Training Language Models: Mimicking Human Reading Comprehension for Enhanced Performance in Biomedicine, Finance, and Law

September 27, 2023
in AI & Technology
Reading Time: 4 mins read
A A
This AI Paper from Microsoft Introduces a New Approach to Training Language Models: Mimicking Human Reading Comprehension for Enhanced Performance in Biomedicine, Finance, and Law
ShareShareShareShareShare

Domain-specific big language models have emerged due to the oversaturation of general large language models (LLMs). Three main categories may be used to group existing methodologies. The first builds models from scratch using a combination of generic and domain-specific corpora. Even though this naturally produces domain-specific LLMs, the large computational and data needs cause serious issues. The second method, which is more economical, refines the language model using supervised datasets. However, it needs to be determined how well-tuned LLMs can understand domain knowledge that can be utilized across all domain-specific activities. In the third, recovered domain information is used to motivate the general language model, which may be seen as an application of LLM rather than a direct improvement to the LLM itself. 

Researchers from Microsoft try domain-adaptive pretraining, or ongoing pretraining on domain-specific corpora, which they believe is useful in customizing different natural language processing models to certain domains. By combining domain-specific knowledge with broad ability, this method benefits downstream domain-specific activities while incurring less expense. This drives their research into whether ongoing pretraining is similarly advantageous for extensive generative models. They undertake preliminary experiments on three domains, biology, finance, and law, and find that further training on the raw corpora drastically reduces prompting performance while maintaining benefits for fine-tuning assessment and knowledge probing tests. This leads us to the conclusion that domain-adaptive pretraining using raw corpora teaches the LLM about the domain while impairing its capacity to prompt. 

Figure 1 shows a condensed example of a reading comprehension text. The raw text is followed by a series of tasks that are built from it, such as summarization (purple), word-to-text (blue), natural language inference (red), common sense reasoning (teal), paraphrase detection (yellow), and text completion (green). 

They offer a straightforward approach for converting massive raw corpora into reading comprehension texts to use domain-specific knowledge and improve prompting performance. Each raw text is enhanced with several tasks pertinent to its topic, as shown in Figure 1. These exercises are intended to support the model’s continued capacity to respond to queries in natural language, depending on the context of the original text. To further improve prompting ability, they provide a variety of generic directions to the reading comprehension texts. Their tests in biology, economics, and law demonstrate how well their method enhances model performance on numerous domain-specific tasks. They call the final model, which stands for Adapted Large Language Model, AdaptLLM. In the future, they see this process expanded to include creating a generic big language model, adding to the ever-expanding canvas of jobs across additional domains. 

In conclusion, their contributions consist of: 

• In their investigation of ongoing pretraining for big language models, they find that while continuing to train the model on domain-specific raw corpora can provide domain knowledge, it severely degrades its capacity to prompt. 

• To efficiently learn the domain knowledge while concurrently maintaining prompting performance, they present a straightforward recipe that mechanically turns massive raw corpora into reading comprehension texts. Their tests demonstrate that their approach regularly enhances model performance in three distinct fields: biology, finance, and law.


Check out the Paper and Github. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 30k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..


YOU MAY ALSO LIKE

Car Manufacturers Are Ditching CarPlay In 2026: Here’s Why

A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 96 Blocks per Token with Unbounded Temporal Depth

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


🚀 The end of project management by humans (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

Car Manufacturers Are Ditching CarPlay In 2026: Here’s Why
AI & Technology

Car Manufacturers Are Ditching CarPlay In 2026: Here’s Why

September 13, 2026
A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 96 Blocks per Token with Unbounded Temporal Depth
AI & Technology

A Princeton Researcher Proposes Recurrent Looped Transformer (RLT) that Carries Decoder State across Every Token, Fixing 96 Blocks per Token with Unbounded Temporal Depth

September 13, 2026
If Your Laptop Trackpad Is Popping Out, Stop Using It Immediately
AI & Technology

If Your Laptop Trackpad Is Popping Out, Stop Using It Immediately

September 13, 2026
How To Get Your Cut Of PlayStation’s .85 Million Settlement
AI & Technology

How To Get Your Cut Of PlayStation’s $7.85 Million Settlement

September 13, 2026
Next Post
Steve Pagliuca Says Owning Celtics Is a ‘Labor of Love’

Steve Pagliuca Says Owning Celtics Is a 'Labor of Love'

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
How These XL Phones Compete

How These XL Phones Compete

September 10, 2026
How To Watch The Flame Fatales 2026 Speedrunning Marathon

How To Watch The Flame Fatales 2026 Speedrunning Marathon

September 10, 2026
Breaking Down Apple’s First Foldable iPhone With Mark Gurman

Breaking Down Apple’s First Foldable iPhone With Mark Gurman

September 12, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!