• bitcoinBitcoin(BTC)$79,599.00-1.37%
  • ethereumEthereum(ETH)$2,453.29-1.87%
  • tetherTether(USDT)$1.000.01%
  • binancecoinBNB(BNB)$720.89-0.01%
  • rippleXRP(XRP)$1.40-2.80%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$101.94-1.44%
  • tronTRON(TRX)$0.3320740.77%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.54%
  • HyperliquidHyperliquid(HYPE)$84.09-3.69%
  • zcashZcash(ZEC)$1,019.208.61%
  • dogecoinDogecoin(DOGE)$0.084839-2.13%
  • RainRain(RAIN)$0.016457-3.78%
  • moneroMonero(XMR)$527.603.29%
  • USDSUSDS(USDS)$1.00-0.01%
  • chainlinkChainlink(LINK)$11.69-1.10%
  • whitebitWhiteBIT Coin(WBT)$73.14-0.75%
  • leo-tokenLEO Token(LEO)$9.22-0.98%
  • cardanoCardano(ADA)$0.210981-5.35%
  • stellarStellar(XLM)$0.179166-1.80%
  • bitcoin-cashBitcoin Cash(BCH)$246.44-3.12%
  • daiDai(DAI)$1.000.00%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.107675-3.87%
  • USD1USD1(USD1)$1.000.00%
  • litecoinLitecoin(LTC)$51.701.40%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.402.78%
  • uniswapUniswap(UNI)$6.17-2.50%
  • hedera-hashgraphHedera(HBAR)$0.0788810.60%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • avalanche-2Avalanche(AVAX)$7.40-1.27%
  • suiSui(SUI)$0.77-0.39%
  • shiba-inuShiba Inu(SHIB)$0.000005-1.81%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • nearNEAR Protocol(NEAR)$2.1913.06%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,428.11-0.82%
  • crypto-com-chainCronos(CRO)$0.055848-2.57%
  • Circle USYCCircle USYC(USYC)$1.140.04%
  • MemeCoreMemeCore(M)$1.139.16%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • okbOKB(OKB)$109.410.65%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.00%
  • BittensorBittensor(TAO)$227.640.57%
  • aaveAave(AAVE)$130.42-1.76%
  • AsterAster(ASTER)$0.732.59%
  • pax-goldPAX Gold(PAXG)$4,435.90-0.84%
  • mantleMantle(MNT)$0.570.87%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056612-1.61%
  • OndoOndo(ONDO)$0.3671491.47%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from Imperial College London and DeepMind Designed an AI Framework that Uses Language as the Core Reasoning Tool of an RL Agent

July 28, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Researchers from Imperial College London and DeepMind Designed an AI Framework that Uses Language as the Core Reasoning Tool of an RL Agent
ShareShareShareShareShare

In recent years, there have been significant breakthroughs in the field of Deep Learning, particularly in the popular sub-fields of Artificial Intelligence, including Natural Language Processing (NLP), Natural Language Understanding (NLU) and Computer Vision (CV). Large Language Models (LLMs) have been created in the framework of NLP and demonstrate outstanding language processing and text production skills that are on par with human talents. On the other hand, without any explicit guidance, CV’s Vision Transformers (ViTs) have been able to learn meaningful representations from photos and videos. Vision-linguistic Models (VLMs) have also been developed, which can connect visual inputs with linguistic descriptions or the other way around.

Foundation Models behind a wide range of downstream applications involving various input modalities have been pre-trained on vast amounts of textual and visual data, leading to the emergence of significant attributes like common sense reasoning, proposing and sequencing sub-goals, and visual understanding. The prospect of utilizing Foundation Models’ capabilities to create more effective and all-encompassing reinforcement learning (RL) agents is the topic of research for researchers. RL agents often pick up knowledge through interacting with their surroundings and getting rewards as feedback, but this method of learning by trial and error can be time-consuming and unworkable.

To address the limitations, a team of researchers has proposed a framework that places language at the core of reinforcement learning robotic agents, particularly in scenarios where learning from scratch is required. The core contribution of their work is to demonstrate that by utilizing LLMs and VLMs, they can effectively address several fundamental problems in particularly four RL settings.

  1. Efficient Exploration in Sparse-Reward Settings: It is difficult for RL agents to learn the best behavior because they frequently find it difficult to explore settings with few rewards. The suggested approach makes exploration and learning in these contexts more effective by utilizing the knowledge kept in Foundation Models.
  1. Reusing gathered Data for Sequential Learning: The framework allows RL agents to build on previously gathered data rather than beginning from scratch each time a new task is met, aiding the sequential learning of new tasks.
  1. Scheduling learned abilities for NewTasks: The framework supports the scheduling of learned abilities, enabling agents to handle novel tasks with their current knowledge efficiently.
  1. Learning from Observations of Expert Agents: By using Foundation Models to learn from observations of expert agents, learning processes can become more efficient and quick.

The team has summarized the main contributions as follows –

  1. The framework has been made in a way that enables the RL agent to reason and make judgments more effectively based on textual information by using language models and vision language models as the fundamental reasoning tools. The agent’s capacity to comprehend challenging tasks and settings is improved by this method.
  1. The proposed framework shows its efficiency in resolving fundamental RL problems that in the past needed distinct, specially created algorithms.
  1. The new framework outperforms conventional baseline techniques in the sparse-reward robotic manipulation setting.
  2. The framework also shows that it can efficiently use previously taught skills to complete tasks. The RL agent’s generalization and adaptability are enhanced by the ability to transfer learned information to new situations.
  1. It demonstrates how the RL agent may accurately learn from observable demonstrations by imitating films of human experts.

In conclusion, the study shows that language models and vision language models have the ability to serve as the core components of reinforcement learning agents’ reasoning.


Check out the Paper. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 26k+ ML SubReddit, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.


YOU MAY ALSO LIKE

How To See What’s Taking Up Space On Your Windows PC

OpenAI Commits $1B to Frontline Cyber Defense, Launches MS-ISAC Pilot – Unite.AI

Tanya Malhotra is a final year undergrad from the University of Petroleum & Energy Studies, Dehradun, pursuing BTech in Computer Science Engineering with a specialization in Artificial Intelligence and Machine Learning.
She is a Data Science enthusiast with good analytical and critical thinking, along with an ardent interest in acquiring new skills, leading groups, and managing work in an organized manner.


🔥 Gain a competitive
edge with data: Actionable market intelligence for global brands, retailers, analysts, and investors. (Sponsored)

Credit: Source link

ShareTweetSendSharePin

Related Posts

How To See What’s Taking Up Space On Your Windows PC
AI & Technology

How To See What’s Taking Up Space On Your Windows PC

September 4, 2026
OpenAI Commits B to Frontline Cyber Defense, Launches MS-ISAC Pilot – Unite.AI
AI & Technology

OpenAI Commits $1B to Frontline Cyber Defense, Launches MS-ISAC Pilot – Unite.AI

September 4, 2026
Flock Cameras Are Officially Banned On State Roads In Florida
AI & Technology

Flock Cameras Are Officially Banned On State Roads In Florida

September 4, 2026
Researchers Document OpenAI Agent Swarm That Repurposed German Wiki – Unite.AI
AI & Technology

Researchers Document OpenAI Agent Swarm That Repurposed German Wiki – Unite.AI

September 4, 2026
Next Post
Facebook Is Working to Combat Online Sex Trafficking, Says Zuckerberg

Facebook Is Working to Combat Online Sex Trafficking, Says Zuckerberg

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Meta Settlement a ‘Pittance,’ Law Professor Says

Meta Settlement a ‘Pittance,’ Law Professor Says

August 29, 2026
Tron Carter | Invested in the Game Podcast Clip

Tron Carter | Invested in the Game Podcast Clip

September 1, 2026
US strikes Iranian launchers on Larak Island in first known attack in weeks – BBC

US strikes Iranian launchers on Larak Island in first known attack in weeks – BBC

August 31, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!