• bitcoinBitcoin(BTC)$79,816.000.47%
  • ethereumEthereum(ETH)$2,459.720.33%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$769.287.36%
  • rippleXRP(XRP)$1.411.04%
  • usd-coinUSDC(USDC)$1.000.01%
  • solanaSolana(SOL)$102.981.52%
  • tronTRON(TRX)$0.3342141.24%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.27%
  • HyperliquidHyperliquid(HYPE)$85.360.89%
  • zcashZcash(ZEC)$1,015.852.86%
  • dogecoinDogecoin(DOGE)$0.0876103.87%
  • RainRain(RAIN)$0.0169402.14%
  • moneroMonero(XMR)$540.603.58%
  • USDSUSDS(USDS)$1.00-0.02%
  • chainlinkChainlink(LINK)$11.942.46%
  • whitebitWhiteBIT Coin(WBT)$73.270.35%
  • leo-tokenLEO Token(LEO)$9.27-0.20%
  • cardanoCardano(ADA)$0.2177092.28%
  • stellarStellar(XLM)$0.1835122.54%
  • bitcoin-cashBitcoin Cash(BCH)$251.800.19%
  • daiDai(DAI)$1.00-0.01%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.1099482.26%
  • USD1USD1(USD1)$1.00-0.01%
  • litecoinLitecoin(LTC)$53.736.95%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.424.76%
  • uniswapUniswap(UNI)$6.351.43%
  • hedera-hashgraphHedera(HBAR)$0.0804213.98%
  • suiSui(SUI)$0.806.23%
  • avalanche-2Avalanche(AVAX)$7.552.54%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.0000054.76%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • nearNEAR Protocol(NEAR)$2.1812.14%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0564590.67%
  • tether-goldTether Gold(XAUT)$4,424.96-0.15%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.121.17%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$112.894.68%
  • BittensorBittensor(TAO)$235.965.52%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.03%
  • AsterAster(ASTER)$0.8110.43%
  • aaveAave(AAVE)$131.720.38%
  • pax-goldPAX Gold(PAXG)$4,432.41-0.12%
  • mantleMantle(MNT)$0.580.89%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056972-0.02%
  • OndoOndo(ONDO)$0.3687054.30%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from USC and Microsoft Propose UniversalNER: A New AI Model Trained with Targeted Distillation Recognizing 13k+ Entity Types and Outperforming ChatGPT’s NER Accuracy by 9% F1 on 43 Datasets

August 13, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Researchers from USC and Microsoft Propose UniversalNER: A New AI Model Trained with Targeted Distillation Recognizing 13k+ Entity Types and Outperforming ChatGPT’s NER Accuracy by 9% F1 on 43 Datasets
ShareShareShareShareShare

ChatGPT and other large language models (LLMs) have shown impressive generalization abilities, but their training and inference costs are often prohibitive. Additionally, white-box access to model weights and inference probabilities is frequently crucial for explainability and confidence in mission-critical applications like healthcare. As a result, instruction tuning has gained popularity as a method for condensing LLMs into more affordable and transparent student models. These student models have shown convincing skills to mimic ChatGPT, as Alpaca and Vicuna showed. Close examination reveals that they still need to catch up to the ideal LLM, particularly in downstream applications that are specifically targeted. 

Because of the restricted computing available, a generic distillation can only create a superficial approximation of the original LLM across all conceivable applications. Instead, they investigate targeted distillation in this research, where they train student models through mission-focused instruction adjustment for a diverse application class like open information extraction. They demonstrate that while maintaining its generalizability across semantic types and domains, this may maximally reproduce LLM’s capabilities for the specified application class. Since named entity recognition (NER) is one of the most fundamental problems in natural language processing, they chose it for their case study. Recent research demonstrates that LLMs still need to catch up to the most advanced supervised system for an entity type when there are many annotated instances. 

There needs to be music little-annotable for most object kinds, though. Developing annotated examples is costly and time-consuming, especially in high-value sectors like biology, where annotation requires specialized knowledge. New entity types are continually emerging. Supervised NER models also show poor generalizability for new domains and entity types since they are trained on pre-specified entity types and domains. They outline a generic process for LLM targeted distillation and show how open-domain NER may use it. Researchers from the University of Southern California and Microsoft Research demonstrate how to utilize ChatGPT to create instruction-tuning data for NER from large amounts of unlabeled online text and use LLaMA to create the UniversalNER models (abbreviated UniNER). 

They put up the biggest and most varied NER benchmark to date (UniversalNER benchmark), which consists of 43 datasets from 9 different disciplines, including medical, programming, social media, law, and finance. LLaMA and Alpaca score badly on this benchmark (around 0 F1) on zero-shot NER. Vicuna performs significantly better in comparison, yet in average F1, it is still behind ChatGPT by more than 20 absolute points. In contrast, UniversalNER outperforms Vicuna by over 30 absolute points in average F1 and achieves state-of-the-art NER accuracy across tens of thousands of entity types in the UniversalNER benchmark. In addition to replicating ChatGPT’s capacity to recognize any entity with a small number of parameters (7–13 billion), UniversalNER also beats its NER accuracy by 7-9 absolute points in average F1. 

Surprisingly, UniversalNER significantly surpasses state-of-the-art multi-task instruction-tuned systems like InstructUIE, which uses supervised NER instances. They also undertake extensive ablation tests to evaluate the effects of different distillation components like the instruction prompts and negative sampling. They will provide their distillation recipe, data, and the UniversalNER model and present an interactive demo to aid further study on targeted distillation.

Build your personal brand with Taplio! 🚀 The 1st AI-powered tool to grow on LinkedIn (Sponsored)

Check out the Paper, Github, and Project Page. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 28k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.


YOU MAY ALSO LIKE

How To Check Your MacBook’s Hard Drive Health

Remote Work As A Worm, Colorful Platformers And Other New Indie Games Worth Checking Out

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


🔥 Use SQL to predict the future (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To Check Your MacBook’s Hard Drive Health
AI & Technology

How To Check Your MacBook’s Hard Drive Health

September 5, 2026
Remote Work As A Worm, Colorful Platformers And Other New Indie Games Worth Checking Out
AI & Technology

Remote Work As A Worm, Colorful Platformers And Other New Indie Games Worth Checking Out

September 5, 2026
OpenAI Plans Misalignment Incident Reporting Framework After Wiki Incident – Unite.AI
AI & Technology

OpenAI Plans Misalignment Incident Reporting Framework After Wiki Incident – Unite.AI

September 5, 2026
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
AI & Technology

Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus

September 5, 2026
Next Post
The Fabulously Wealthy Do This More Than Less Affluent Demographics

The Fabulously Wealthy Do This More Than Less Affluent Demographics

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Aerial photos show destruction by wildfires in Spokane

Aerial photos show destruction by wildfires in Spokane

August 30, 2026
Walmart mangoes recalled over potential salmonella contamination

Walmart mangoes recalled over potential salmonella contamination

September 1, 2026
Battling wildfires across Europe

Battling wildfires across Europe

September 2, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!