• bitcoinBitcoin(BTC)$76,778.001.42%
  • ethereumEthereum(ETH)$2,468.453.25%
  • tetherTether(USDT)$1.00-0.03%
  • binancecoinBNB(BNB)$728.732.26%
  • rippleXRP(XRP)$1.303.11%
  • usd-coinUSDC(USDC)$1.00-0.02%
  • solanaSolana(SOL)$101.304.35%
  • tronTRON(TRX)$0.333810-0.57%
  • zcashZcash(ZEC)$1,478.5517.93%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.022.18%
  • HyperliquidHyperliquid(HYPE)$82.174.38%
  • dogecoinDogecoin(DOGE)$0.0820133.69%
  • moneroMonero(XMR)$512.514.24%
  • USDSUSDS(USDS)$1.000.03%
  • whitebitWhiteBIT Coin(WBT)$79.201.84%
  • RainRain(RAIN)$0.012872-2.02%
  • chainlinkChainlink(LINK)$11.386.16%
  • leo-tokenLEO Token(LEO)$8.920.76%
  • cardanoCardano(ADA)$0.2025395.51%
  • stellarStellar(XLM)$0.1870577.58%
  • uniswapUniswap(UNI)$7.6523.52%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • bitcoin-cashBitcoin Cash(BCH)$233.898.41%
  • daiDai(DAI)$1.00-0.04%
  • USD1USD1(USD1)$1.00-0.02%
  • litecoinLitecoin(LTC)$53.836.40%
  • CantonCanton(CC)$0.10058510.48%
  • nearNEAR Protocol(NEAR)$3.0122.22%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.342.51%
  • avalanche-2Avalanche(AVAX)$7.625.24%
  • hedera-hashgraphHedera(HBAR)$0.0761784.55%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.0000058.93%
  • suiSui(SUI)$0.747.30%
  • crypto-com-chainCronos(CRO)$0.0579004.61%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • tether-goldTether Gold(XAUT)$4,354.790.32%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • MemeCoreMemeCore(M)$1.162.69%
  • BittensorBittensor(TAO)$229.076.19%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • okbOKB(OKB)$112.262.62%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.150.29%
  • AsterAster(ASTER)$0.748.34%
  • aaveAave(AAVE)$127.9511.03%
  • BitwayBitway(BTW)$0.70-6.38%
  • pax-goldPAX Gold(PAXG)$4,355.850.23%
  • mantleMantle(MNT)$0.574.68%
  • Pump.funPump.fun(PUMP)$0.0039557.64%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Microsoft Research Introduces CORPGEN To Manage Multi Horizon Tasks For Autonomous AI Agents Using Hierarchical Planning and Memory

February 27, 2026
in AI & Technology
Reading Time: 5 mins read
A A
Microsoft Research Introduces CORPGEN To Manage Multi Horizon Tasks For Autonomous AI Agents Using Hierarchical Planning and Memory
ShareShareShareShareShare

Microsoft researchers have introduced CORPGEN, an architecture-agnostic framework designed to manage the complexities of realistic organizational work through autonomous digital employees. While existing benchmarks evaluate AI agents on isolated, single tasks, real-world corporate environments require managing dozens of concurrent, interleaved tasks with complex dependencies. The research team identifies this distinct problem class as Multi-Horizon Task Environments (MHTEs).

The Performance Gap in MHTEs

Empirical testing reveals that baseline computer using agents (CUAs) experience significant performance degradation when moved from single-task scenarios to MHTEs. Using three independent CUA implementations, completion rates dropped from 16.7% at 25% load to 8.7% at 100% load.

YOU MAY ALSO LIKE

Razer Refreshes The One-Handed Tartarus Pro Keyboard With Improved Switches

OceanStor M900 Brings PB-Scale Context Memory to Huawei SuperPoDs – Unite.AI

The research team identified four fundamental failure modes causing this decline:

  • Context Saturation: Context requirements grow O(N) with task count rather than O(1), rapidly exceeding the token window capacity.
  • Memory Interference: Information from one task often contaminates reasoning about another when multiple tasks share a single context window.
  • Dependency Graph Complexity: Corporate tasks form Directed Acyclic Graphs (DAGs) rather than linear chains, requiring complex topological reasoning.
  • Reprioritization Overhead: Decision complexity increases to O(N) per cycle because agents must constantly re-evaluate priorities across all active tasks.
https://arxiv.org/pdf/2602.14229

The CORPGEN Architecture

To address these failures, CORPGEN implements Multi-Objective Multi-Horizon Agent (MOMA) capabilities through four primary architectural mechanisms.

(a) Hierarchical Planning

Strategic coherence is maintained through goal decomposition across three temporal scales:

  • Strategic Objectives (Monthly): High-level goals and milestones based on agent identity and role.
  • Tactical Plans (Daily): Actionable tasks for specific applications with priority rankings.
  • Operational Actions (Per-Cycle): Individual tool calls selected based on current state and retrieved memory.

(b) Sub-Agent Isolation

Complex operations, such as GUI automation or research, are isolated into modular sub-agents. These autonomous agents operate in their own context scopes and return only structured results to the host agent, preventing cross-task memory contamination.

(c) Tiered Memory Architecture

The system utilizes a three-layer memory structure to manage state:

  • Working Memory: Intended for immediate reasoning, this layer resets each cycle.
  • Structured Long-Term Memory (LTM): Stores typed artifacts such as plans, summaries, and reflections.
  • Semantic Memory: Uses Mem0 to support similarity-based retrieval over unstructured past context using embeddings.

(d) Adaptive Summarization

To bound context growth, CORPGEN employs rule-based compression. When context length exceeds 4,000 tokens, ‘critical content’ (such as tool calls and state changes) is preserved verbatim, while ‘routine content’ (intermediate reasoning) is compressed into structured summaries.

Experimental Results and Learning

Across three CUA backends (UFO2, OpenAI CUA, and hierarchical), CORPGEN achieved up to a 3.5x improvement over baselines, reaching a 15.2% completion rate compared to 4.3% for standalone UFO2 at 100% load.

Ablation studies indicate that experiential learning provides the largest performance gains. This mechanism distills successful task executions into canonical trajectories which are then indexed in a FAISS database. At execution time, similar trajectories are retrieved as few-shot examples to bias action selection toward validated patterns.

The research TEAM observed a significant discrepancy in evaluation methods. Artifact-based judgment (inspecting generated files and outputs) achieved a 90% agreement rate with human labels. In contrast, trace-based LLM judgment (relying on screenshots and execution logs) only achieved 40% agreement. This suggests that current benchmarks may systematically underestimate agent performance by relying on limited visual traces rather than the actual artifacts produced.

Key Takeaways

  • Identification of Multi-Horizon Task Environments (MHTEs): The research team defines a new class of problems called MHTEs, where agents must manage dozens of interleaved, long-horizon tasks (45+ tasks, 500-1500+ steps) within a single persistent context. This differs from traditional benchmarks that evaluate single tasks in isolation.
  • Discovery of Catastrophic Performance Degradation: Standard computer-using agents (CUAs) experience a ‘catastrophic’ drop in performance when task load increases, with completion rates falling from 16.7% at 25% load to 8.7% at 100% load.
  • Four Fundamental Failure Modes: The researchers identified why current agents fail under load: context saturation (O(N) growth), memory interference (task conflation), dependency complexity (managing Directed Acyclic Graphs), and reprioritization overhead (O(N) decision complexity).
  • Architectural Mitigation via CORPGEN: The CORPGEN framework addresses these failures through four core mechanisms: hierarchical planning for goal alignment, sub-agent isolation to prevent memory contamination, tiered memory (working, structured, and semantic), and adaptive summarization to manage token limits.
  • Significant Performance Gains through Experiential Learning: Evaluation across multiple backends showed that CORPGEN can improve performance by up to 3.5x over baselines. Ablation studies revealed that experiential learning—reusing verified successful trajectories—provides the largest performance boost among all architectural components.

Check out the Paper and Technical details. Also, feel free to follow us on Twitter and don’t forget to join our 120k+ ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

The post Microsoft Research Introduces CORPGEN To Manage Multi Horizon Tasks For Autonomous AI Agents Using Hierarchical Planning and Memory appeared first on MarkTechPost.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Razer Refreshes The One-Handed Tartarus Pro Keyboard With Improved Switches
AI & Technology

Razer Refreshes The One-Handed Tartarus Pro Keyboard With Improved Switches

September 17, 2026
OceanStor M900 Brings PB-Scale Context Memory to Huawei SuperPoDs – Unite.AI
AI & Technology

OceanStor M900 Brings PB-Scale Context Memory to Huawei SuperPoDs – Unite.AI

September 17, 2026
Europe’s EU Kids Act Would Ban Social Media Access For Children Under 13
AI & Technology

Europe’s EU Kids Act Would Ban Social Media Access For Children Under 13

September 17, 2026
NVIDIA And Google’s New Coalition Wants To Speed Up AI Data Center Power Grid Connections
AI & Technology

NVIDIA And Google’s New Coalition Wants To Speed Up AI Data Center Power Grid Connections

September 17, 2026
Next Post
Breezy Johnson reflects on taking home first U.S. gold medal of 2026 Olympics

Breezy Johnson reflects on taking home first U.S. gold medal of 2026 Olympics

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Helena Foulkes wins Rhode Island Democratic governor primary, NBC News projects

Helena Foulkes wins Rhode Island Democratic governor primary, NBC News projects

September 14, 2026
Reddington says Trump has the ability to put pressure on DA Tim Cruz

Reddington says Trump has the ability to put pressure on DA Tim Cruz

September 15, 2026
Hubbell Stock: The Grid’s Small Components Can Deliver Large Returns (NYSE:HUBB)

Hubbell Stock: The Grid’s Small Components Can Deliver Large Returns (NYSE:HUBB)

September 11, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!