• bitcoinBitcoin(BTC)$79,699.00-0.19%
  • ethereumEthereum(ETH)$2,494.940.46%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$748.99-2.82%
  • rippleXRP(XRP)$1.41-0.30%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$105.221.66%
  • tronTRON(TRX)$0.3353990.36%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.060.00%
  • zcashZcash(ZEC)$1,223.6920.21%
  • HyperliquidHyperliquid(HYPE)$86.811.52%
  • dogecoinDogecoin(DOGE)$0.089861-0.81%
  • RainRain(RAIN)$0.016722-1.85%
  • moneroMonero(XMR)$538.42-2.43%
  • USDSUSDS(USDS)$1.000.03%
  • chainlinkChainlink(LINK)$12.856.77%
  • whitebitWhiteBIT Coin(WBT)$73.50-0.01%
  • leo-tokenLEO Token(LEO)$9.370.91%
  • cardanoCardano(ADA)$0.219669-0.06%
  • stellarStellar(XLM)$0.1851390.16%
  • bitcoin-cashBitcoin Cash(BCH)$257.18-0.83%
  • daiDai(DAI)$1.00-0.01%
  • uniswapUniswap(UNI)$7.170.89%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • CantonCanton(CC)$0.1091530.14%
  • USD1USD1(USD1)$1.000.01%
  • litecoinLitecoin(LTC)$54.690.10%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.42-0.72%
  • hedera-hashgraphHedera(HBAR)$0.0808070.46%
  • avalanche-2Avalanche(AVAX)$7.751.72%
  • suiSui(SUI)$0.800.22%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.29%
  • nearNEAR Protocol(NEAR)$2.4311.14%
  • paypal-usdPayPal USD(PYUSD)$1.000.01%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • crypto-com-chainCronos(CRO)$0.0572741.01%
  • tether-goldTether Gold(XAUT)$4,420.00-0.16%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • MemeCoreMemeCore(M)$1.13-0.14%
  • BittensorBittensor(TAO)$262.3011.14%
  • Ripple USDRipple USD(RLUSD)$1.000.01%
  • okbOKB(OKB)$112.82-0.36%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.25%
  • AsterAster(ASTER)$0.78-0.10%
  • aaveAave(AAVE)$134.17-0.12%
  • mantleMantle(MNT)$0.602.74%
  • pax-goldPAX Gold(PAXG)$4,424.17-0.19%
  • OndoOndo(ONDO)$0.3784131.76%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.056535-1.39%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

NVIDIA AI Introduces ASPIRE: A Self-Improving Robotics Framework Reaching 31% Zero-Shot on LIBERO-Pro Long Tasks

July 4, 2026
in AI & Technology
Reading Time: 14 mins read
A A
NVIDIA AI Introduces ASPIRE: A Self-Improving Robotics Framework Reaching 31% Zero-Shot on LIBERO-Pro Long Tasks
ShareShareShareShareShare

Traditional robot programming is hard to scale. It requires orchestrating multimodal perception, physical contact dynamics, diverse configurations, and execution failures by hand. Code-as-policy systems let language models compose these into executable robot programs. That makes robot behavior inspectable, editable, and debuggable.

But existing robotic coding agents run in naive execution environments. They receive only coarse, task-level feedback. A failed rollout signals that the task failed, not why. The root cause can be perception, motion planning, grasping, contact dynamics, or long-horizon coordination. These systems also discard fixes once a task ends. So the agent solving its hundredth task is no more experienced than at its first.

A team of researchers from NVIDIA, University of Michigan, UIUC, UC Berkeley, and CMU introduces ASPIRE (Agentic Skill Programming through Iterative Robot Exploration). It is a continual learning system that writes and refines robot control programs. It also distills validated fixes into a reusable, transferable skill library.

How ASPIRE works

ASPIRE runs an open-ended learning loop with three components. It uses a coordinator–actor architecture. A central coordinator manages the shared skill library and dispatches actor coding agents to tasks. Actors do not exchange full chat histories or raw trajectories. Only distilled skills move between them.

Closed-loop robot execution engine: This replaces coarse rollout feedback with per-primitive multimodal traces. For each perception, planning, and control call, it stores inputs, outputs, and return status. It also stores RGB keyframes, overlays, grasp candidates, object poses, and motion-planning results. The agent inspects only the calls implicated by a failure. It then localizes the fault and validates a repair through re-execution.

Skill library: Reusable knowledge is rarely an entire task program. So the library stores heterogeneous fixes. These include localization heuristics, perception prompts, grasping constraints, motion primitives, and debugging workflows. Each skill is compact in-context guidance. It holds a failure signature, a when-to-apply condition, a repair strategy, and often a code sketch. The coordinator admits only patterns that pass debug validation and API-policy checks.

Evolutionary search: Trace-guided debugging alone can collapse into local repair loops. The agent keeps patching the same failed strategy. To broaden exploration, ASPIRE proposes K candidate programs each round. Candidates condition on top-performing prior programs and their remaining failure traces. The next round explores distinct strategies rather than refining one solution.

In simulation, the coding agent is Claude Code with Claude Opus 4.6 and a 1M-token context window. Programs are written in CaP-X, an open-source code-as-policy framework built on MuJoCo Playground. The agent cannot read simulator ground truth. Reading physics-engine state or asset files like .bddl, .xml, or .urdf is forbidden. The rule is simple. If a real robot with a camera could do it, it is allowed.

Interactive Explainer


A worked example: the Multi-Angle Approach skill

Consider a BEHAVIOR-1K task where a robot must pick up a radio near a table. Perception returns the radio pose, but repeated navigate_to_pose calls fail. The generated goal lies within about 20 centimeters of the table edge. That falls inside the table’s collision-avoidance buffer, and cuRobo returns PLANNING_ERROR.

The agent reads the trace and localizes the cause. The failure is target infeasibility, not perception or grasping. It then writes a repair that samples standoff poses around the radio.

YOU MAY ALSO LIKE

H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder

How To Send High-Quality Images And Videos From Android To iPhone

# radio_pos, safe_navigate() and dist_to() are provided by ASPIRE's robot API
for angle_deg in [180, -90, 90, -45, 45]:
    angle = np.radians(angle_deg)
    tx = radio_pos[0] + 0.7 * np.cos(angle)      # standoff 0.7 m from the radio
    ty = radio_pos[1] + 0.7 * np.sin(angle)
    face_yaw = np.arctan2(radio_pos[1] - ty, radio_pos[0] - tx)
    moved = safe_navigate([tx, ty, face_yaw], f"ang_{angle_deg}")
    if moved and dist_to(radio_pos[:2]) < 0.8:   # reached a pose within 0.8 m
        break

Each angle puts the goal on a different side of the object. When one side is blocked, another is often open. Here the 180-degree pose clears the buffer. The validated fix is admitted as a reusable navigation-recovery skill.

Benchmarks and results

ASPIRE is evaluated on three benchmark families. LIBERO-Pro tests short-horizon robustness under object, goal, and spatial perturbations. Robosuite covers contact-rich single- and dual-arm manipulation. BEHAVIOR-1K covers long-horizon household mobile manipulation. The primary coding-agent baseline is CaP-Agent0. It uses visual differencing, a predefined skill library, and per-episode test-time retries. The comparison also includes end-to-end vision-language-action policies: OpenVLA, π0, and π0.5.

On LIBERO-Pro, ASPIRE gains up to 77 points on the Object suite. That figure averages both perturbation axes over the strongest baseline. It also gains 41.5 points on Goal and 42.5 points on Spatial. On Robosuite, bimanual handover rises from 20% to 92%. On BEHAVIOR-1K, the radio pickup task rises from 56% to 88%.

The zero-shot result is notable. Reusing skills accumulated on LIBERO-90, ASPIRE reaches about 31% on held-out LIBERO-Pro Long tasks. Prior methods saturate near 4%.

Dimension End-to-end VLAs (OpenVLA, π0, π0.5) CaP-Agent0 ASPIRE
Paradigm Learned-weight policy Code-as-policy agent Code-as-policy agent
Cross-task experience None (frozen weights) Discarded after each task Distilled into a skill library
Failure feedback None at test time Coarse scene-level summaries Per-primitive multimodal traces
Test-time strategy Direct inference Per-seed reasoning + retries One program per task
LIBERO-Pro overall 0–13% 18% 72%
LIBERO-Pro Long zero-shot 0–5% ~4% ~31%

Real-robot skill transfer

The research team tests three simulation-discovered skills on a real bimanual YAM station. The real-robot coding agent is OpenAI Codex GPT-5.5. The embodiment and API differ from simulation. Transferred skills reduce debugging cost. Soda-can lifting improved from 13/20 to 19/20 while using about 10x fewer tokens. Drawer opening moved from 0/20 to 11/20, where the no-skill baseline never succeeded.

Key Takeaways

  • ASPIRE writes and debugs robot programs, then saves validated fixes as reusable in-context skills.
  • Per-primitive multimodal traces let the agent localize failures instead of guessing from rollout outcomes.
  • It gains up to 77 points on LIBERO-Pro and lifts Robosuite handover from 20% to 92%.
  • Zero-shot transfer reaches about 31% on LIBERO-Pro Long, against about 4% for prior methods.
  • Simulation-discovered skills reduced real-robot debugging cost across a different embodiment and API.

Check out the Paper and Project Page. Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

Need to partner with us for promoting your GitHu


Credit: Source link

ShareTweetSendSharePin

Related Posts

H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
AI & Technology

H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder

September 6, 2026
How To Send High-Quality Images And Videos From Android To iPhone
AI & Technology

How To Send High-Quality Images And Videos From Android To iPhone

September 6, 2026
What Is Vibe Coding And Why Does It Get So Much Hate?
AI & Technology

What Is Vibe Coding And Why Does It Get So Much Hate?

September 6, 2026
My Content Tracker Idea Became a Real App – Unite.AI
AI & Technology

My Content Tracker Idea Became a Real App – Unite.AI

September 6, 2026
Next Post
Iran Begins Khamenei’s Funeral in Show of Defiance Against the U.S. – WSJ

Iran Begins Khamenei’s Funeral in Show of Defiance Against the U.S. - WSJ

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Father of boy rescued from pounding California surf says he is beyond grateful

Father of boy rescued from pounding California surf says he is beyond grateful

September 2, 2026
Perplexity Releases Hybrid Compute on Mac: Cloud Agents Orchestrate Down to a Local Model, Gated On Device

Perplexity Releases Hybrid Compute on Mac: Cloud Agents Orchestrate Down to a Local Model, Gated On Device

September 2, 2026
Full Episode: TODAY Show – July 29

Full Episode: TODAY Show – July 29

September 3, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!