• bitcoinBitcoin(BTC)$77,293.00-1.15%
  • ethereumEthereum(ETH)$2,413.45-1.90%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$686.41-0.24%
  • rippleXRP(XRP)$1.34-2.03%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$99.76-2.80%
  • tronTRON(TRX)$0.323143-2.48%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.033.18%
  • HyperliquidHyperliquid(HYPE)$82.28-1.20%
  • zcashZcash(ZEC)$830.79-2.08%
  • dogecoinDogecoin(DOGE)$0.081532-1.49%
  • RainRain(RAIN)$0.0168070.37%
  • USDSUSDS(USDS)$1.000.01%
  • moneroMonero(XMR)$519.790.15%
  • leo-tokenLEO Token(LEO)$9.27-0.22%
  • whitebitWhiteBIT Coin(WBT)$71.11-1.45%
  • chainlinkChainlink(LINK)$11.20-1.43%
  • cardanoCardano(ADA)$0.196424-1.06%
  • stellarStellar(XLM)$0.174487-1.31%
  • bitcoin-cashBitcoin Cash(BCH)$246.77-0.13%
  • daiDai(DAI)$1.000.02%
  • CantonCanton(CC)$0.112215-7.96%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • USD1USD1(USD1)$1.00-0.02%
  • uniswapUniswap(UNI)$6.299.91%
  • litecoinLitecoin(LTC)$49.040.86%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.32-4.64%
  • hedera-hashgraphHedera(HBAR)$0.073901-0.48%
  • Global DollarGlobal Dollar(USDG)$1.000.01%
  • avalanche-2Avalanche(AVAX)$7.20-0.78%
  • shiba-inuShiba Inu(SHIB)$0.0000051.10%
  • suiSui(SUI)$0.720.17%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • crypto-com-chainCronos(CRO)$0.055033-2.08%
  • tether-goldTether Gold(XAUT)$4,325.19-1.42%
  • nearNEAR Protocol(NEAR)$1.86-3.80%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • MemeCoreMemeCore(M)$1.05-2.30%
  • okbOKB(OKB)$109.49-1.65%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.33%
  • BittensorBittensor(TAO)$218.75-3.61%
  • aaveAave(AAVE)$130.613.84%
  • AsterAster(ASTER)$0.71-0.16%
  • pax-goldPAX Gold(PAXG)$4,335.11-1.36%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0572780.60%
  • mantleMantle(MNT)$0.551.39%
  • MorphoMorpho(MORPHO)$2.611.64%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Exploring Instruction-Tuning Language Models: Meet Tülu-A Suite of Fine-Tuned Large Language Models (LLMs)

June 13, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Exploring Instruction-Tuning Language Models: Meet Tülu-A Suite of Fine-Tuned Large Language Models (LLMs)
ShareShareShareShareShare

The well-famous ChatGPT developed by OpenAI is one of the best examples of Large Language Models (LLMs) that have been recently released. LLMs like ChatGPT have taken the world by storm with their unmatchable potential and ability to imitate humans in performing various tasks. These models have mostly adopted instruction fine-tuning to help get the model into the habit of performing some common tasks. This approach involves training the models on supervised input and output pairs, which can be derived from other models. 

Various open instruction-following datasets are being used for the current advancements in instruction-tuning language models. Though open models can compete with cutting-edge proprietary models, these assertions are frequently only backed by a restricted evaluation, which makes it difficult to compare models in-depth and determine the value of various resources. To address this, a team of researchers from the Allen Institute for AI and the University of Washington has introduced a wide range of instruction-tuned models with parameter sizes ranging from 6.7 billion to 65 billion.

These models are trained on 12 instruction datasets ranging from synthetic and distilled datasets like Alpaca to hand-curated datasets like OpenAssistant. The models are carefully tested in a variety of areas, including reasoning, multilingualism, coding, factual knowledge, and open-ended instruction-following skills. In order to provide a thorough study, the evaluation is carried out utilizing a collection of automatic, model-based, and human-based metrics.

🚀 JOIN the fastest ML Subreddit Community

The team has also introduced TÜLU, which is a suite of large language models fine-tuned on a combination of data sources. These models are fine-tuned using a combination of high-quality open resources. The team has examined the performance of various instruction-tuning datasets and their effect on particular skills through various evaluations. They discovered that different datasets could reveal or improve particular skills and that neither a single dataset nor a set of datasets offers the highest performance across all evaluations.

The team has mentioned that an interesting finding from the research is that benchmark-based evaluations fail to capture differences in model capabilities that are shown by model comparisons. The best model in any given evaluation averaged 83% of ChatGPT’s performance and 68% of GPT-4’s performance. The team has stated that TÜLU, with 65 billion parameters, is the largest publicly-released, fully-instruction tuned variant, trained on seven popular available datasets. It has achieved the best average performance while staying within 15% of the best-performing model on each individual task.

Some of the key contributions mentioned in the research paper are – 

  1. Specific domain and capability-specific instruction datasets are very successful at enhancing model performance.
  1. Larger or pre-trained-for-longer base models consistently perform better after instruction tuning.
  1. The best average performance across benchmarks was attained by TÜLU, the fine-tuned LLaMa on a mixture of existing instruction datasets, although it is not the best when comparing various evaluation settings separately.
  1. Even a very big 65B parameter model that has been optimized on a huge variety of instruction datasets falls short of ChatGPT, although it outperforms comparable smaller models by a significant margin.
  1. Strong correlations between model-based preference evaluation on open-ended instruction following and the typical number of unique tokens produced by a model indicate that model-based preference evaluation contains biases that may mask variations in model capabilities.

Check Out The Paper and Github link. Don’t forget to join our 23k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more. If you have any questions regarding the above article or if we missed anything, feel free to email us at [email protected]

🚀 Check Out 100’s AI Tools in AI Tools Club


YOU MAY ALSO LIKE

GTA VI Extended Look Got 31 Million Views On Netflix Despite Six-Hour Exclusivity Window

Meta Superintelligence Labs Releases Muse Voice Transcribe: One Real-Time Model for Streaming ASR, Diarization, and Endpointing

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


➡️ Try: Criminal IP: AI-based Phishing Link Checker Chrome Extension

Credit: Source link

ShareTweetSendSharePin

Related Posts

GTA VI Extended Look Got 31 Million Views On Netflix Despite Six-Hour Exclusivity Window
AI & Technology

GTA VI Extended Look Got 31 Million Views On Netflix Despite Six-Hour Exclusivity Window

September 2, 2026
Meta Superintelligence Labs Releases Muse Voice Transcribe: One Real-Time Model for Streaming ASR, Diarization, and Endpointing
AI & Technology

Meta Superintelligence Labs Releases Muse Voice Transcribe: One Real-Time Model for Streaming ASR, Diarization, and Endpointing

September 2, 2026
This Is The Best Setting And Placement For Your Dolby Atmos Soundbar
AI & Technology

This Is The Best Setting And Placement For Your Dolby Atmos Soundbar

September 2, 2026
Aramco Digital and Avathon Partner on Autonomous Operations AI – Unite.AI
AI & Technology

Aramco Digital and Avathon Partner on Autonomous Operations AI – Unite.AI

September 1, 2026
Next Post
El Salvador Makes Bitcoin Legal Tender

El Salvador Makes Bitcoin Legal Tender

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Odd Lots: Why Many Communities Are Skeptical of the AI Boom

Odd Lots: Why Many Communities Are Skeptical of the AI Boom

August 29, 2026
FIFA boss Gianni Infantino cancels planned speech at swanky Hamptons bash amid legal threat

FIFA boss Gianni Infantino cancels planned speech at swanky Hamptons bash amid legal threat

August 28, 2026
Mamdani reacts to El-Sayed’s Michigan primary win

Mamdani reacts to El-Sayed’s Michigan primary win

August 28, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!