• bitcoinBitcoin(BTC)$78,721.00-0.86%
  • ethereumEthereum(ETH)$2,493.110.07%
  • tetherTether(USDT)$1.00-0.02%
  • binancecoinBNB(BNB)$755.111.46%
  • rippleXRP(XRP)$1.40-0.13%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$103.68-1.24%
  • tronTRON(TRX)$0.3382000.43%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.060.00%
  • zcashZcash(ZEC)$1,144.54-3.83%
  • HyperliquidHyperliquid(HYPE)$84.39-3.45%
  • dogecoinDogecoin(DOGE)$0.0902350.61%
  • RainRain(RAIN)$0.0170112.86%
  • USDSUSDS(USDS)$1.000.01%
  • moneroMonero(XMR)$520.25-3.15%
  • chainlinkChainlink(LINK)$12.76-4.81%
  • whitebitWhiteBIT Coin(WBT)$78.797.72%
  • leo-tokenLEO Token(LEO)$9.211.31%
  • cardanoCardano(ADA)$0.2205480.96%
  • stellarStellar(XLM)$0.1915290.39%
  • bitcoin-cashBitcoin Cash(BCH)$257.930.57%
  • daiDai(DAI)$1.000.02%
  • uniswapUniswap(UNI)$7.152.29%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • litecoinLitecoin(LTC)$55.72-0.18%
  • USD1USD1(USD1)$1.00-0.02%
  • CantonCanton(CC)$0.105356-2.63%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.40-1.65%
  • hedera-hashgraphHedera(HBAR)$0.080608-0.03%
  • avalanche-2Avalanche(AVAX)$8.102.16%
  • suiSui(SUI)$0.831.86%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.0000050.08%
  • nearNEAR Protocol(NEAR)$2.34-0.28%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0589722.27%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • tether-goldTether Gold(XAUT)$4,398.34-0.15%
  • MemeCoreMemeCore(M)$1.162.83%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • BittensorBittensor(TAO)$255.59-5.19%
  • okbOKB(OKB)$116.332.68%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.23%
  • AsterAster(ASTER)$0.77-2.87%
  • mantleMantle(MNT)$0.63-1.50%
  • aaveAave(AAVE)$131.93-1.57%
  • pax-goldPAX Gold(PAXG)$4,401.60-0.16%
  • OndoOndo(ONDO)$0.382048-1.85%
  • polkadotPolkadot(DOT)$1.0911.18%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Researchers from Google and the University of Toronto Introduce Groundbreaking Zero-Shot Agent for Autonomous Learning and Task Execution in Live Computer Environments

October 25, 2023
in AI & Technology
Reading Time: 4 mins read
A A
Researchers from Google and the University of Toronto Introduce Groundbreaking Zero-Shot Agent for Autonomous Learning and Task Execution in Live Computer Environments
ShareShareShareShareShare

Large language models (LLMs) for action production in various live contexts, such as ALFWORLD and ALPHACODE, have shown promise in earlier efforts. Examples include SAYCAN, REACT, TOOLFORMER, and SWIFTSAGE. LLMs are used similarly to follow expert trails, understand environmental changes, plan and carry out future activities, and compose API requests. Several studies, including REFLEXION and SELF-REFINE, have demonstrated that repeatedly performing a task with numerous rounds of self-reflection may significantly enhance task completion. LLMs are asked to modify a previous execution plan in light of environmental feedback. Such adjustments are incorporated into the action generator’s prompt for the subsequent round. 

MINIWOB++ has recently been utilized as a testbed to evaluate LLM’s performance on modularized computing workloads. Using comprehensive trace examples of the task for direct supervision (WebGUM), self-supervision, or few/many shot prompting (SYNAPSE) are standard methods for learning a task. They have completed dozens of computer jobs with a task completion rate greater than 90%, seemingly solving the computer control issue. Nonetheless, the need for expert traces constrains the agent’s capacity to learn new jobs. Can an agent independently know and enhance its control over a computer without utilizing well-chosen traces as guidance? Researchers from Google Research and the University of Toronto suggest a zero-shot agent to answer this query. 

Their agent is built on top of PaLM2, a recent LLM, and it uses a single set of instruction prompts for all activities rather than task-specific prompts. Additionally, contemporary efforts like RCI, ADAPLANNER, and SYNAPSE use screen representations that might include a lot more data than what is displayed to the user on the screen. For instance, Fig. 1 illustrates items that are contained in the HTML that are provided to the LLM but are not displayed on the screen. Arbitrarily, using this new knowledge makes the agent’s ability to complete the task easier. However, in typical usage scenarios, such information might not be easily accessible and, depending on it, could limit how widely the agent can be applied. 

Figure 1 shows disparate displays on screens. Fig. 1a–1c shows the social media task before and after pressing the “more” button (seed=2). HTML has already made the material visible before clicking. Fig. 1d-1e: The click-tab-2 (seed=0) has a similar problem.

13 rather difficult jobs on MINIWOB++ that are meant to span many screens were carefully evaluated, and they discovered that 5 of them included HTML that contained such information—multi-screen information in a single observation. These are the contributions they made: First, in comparison to earlier studies, they adopt a condensed screen depiction, which makes the test environment more all-encompassing and realistic. Second, they provide a straightforward but effective action planner that, in a single pass, precisely plans out executable operations on a state. They demonstrate that such a “naive” approach can complete nearly all the simple tasks on the MINIWOB++ benchmark using the most recent LLM capacity. 

To help the agent successfully learn from exploratory failures and advance in more difficult tasks, they suggest a systematic thought management technique that draws influence from Reflexion. Their agent achieves performance equivalent to previous few/many-shot state-of-the-art after a few rounds of tries. Their agent is the first zero-shot design for computer control tasks that they are aware of, according to research.


Check out the Paper. All Credit For This Research Goes To the Researchers on This Project. Also, don’t forget to join our 31k+ ML SubReddit, 40k+ Facebook Community, Discord Channel, and Email Newsletter, where we share the latest AI research news, cool AI projects, and more.

If you like our work, you will love our newsletter..

We are also on WhatsApp. Join our AI Channel on Whatsapp..


YOU MAY ALSO LIKE

Renault Is Building Its €17,900 Dacia Spring EV In Europe To Qualify For Local Subsidies

An Attractive ‘Mid-Size’ Foldable With Powerful Specs

Aneesh Tickoo is a consulting intern at MarktechPost. He is currently pursuing his undergraduate degree in Data Science and Artificial Intelligence from the Indian Institute of Technology(IIT), Bhilai. He spends most of his time working on projects aimed at harnessing the power of machine learning. His research interest is image processing and is passionate about building solutions around it. He loves to connect with people and collaborate on interesting projects.


▶️ Now Watch AI Research Updates On Our Youtube Channel [Watch Now]

Credit: Source link

ShareTweetSendSharePin

Related Posts

Renault Is Building Its €17,900 Dacia Spring EV In Europe To Qualify For Local Subsidies
AI & Technology

Renault Is Building Its €17,900 Dacia Spring EV In Europe To Qualify For Local Subsidies

September 8, 2026
An Attractive ‘Mid-Size’ Foldable With Powerful Specs
AI & Technology

An Attractive ‘Mid-Size’ Foldable With Powerful Specs

September 8, 2026
How Long Before a Real Crackdown on AI Model Decensoring? – Unite.AI
AI & Technology

How Long Before a Real Crackdown on AI Model Decensoring? – Unite.AI

September 8, 2026
Reducto Releases r-1: A Single Pass Document Parsing Model That Cuts Errors 20% at 1 Cent Per Page
AI & Technology

Reducto Releases r-1: A Single Pass Document Parsing Model That Cuts Errors 20% at 1 Cent Per Page

September 8, 2026
Next Post
NBC News’ Pete Williams Retires After Nearly 30 Years With Network

NBC News’ Pete Williams Retires After Nearly 30 Years With Network

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Hochul says AI companies are ‘flooding the zone’ with data centers

Hochul says AI companies are ‘flooding the zone’ with data centers

September 5, 2026
Ukraine targeted a warehouse belonging to Wildberries, Russia’s equivalent of Amazon

Ukraine targeted a warehouse belonging to Wildberries, Russia’s equivalent of Amazon

September 7, 2026
Family of Nolan Wells to review audio of sinking boat call made by his friends

Family of Nolan Wells to review audio of sinking boat call made by his friends

September 1, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!