• bitcoinBitcoin(BTC)$80,359.00-0.83%
  • ethereumEthereum(ETH)$2,573.29-2.04%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$751.54-1.35%
  • rippleXRP(XRP)$1.38-2.78%
  • usd-coinUSDC(USDC)$1.000.00%
  • solanaSolana(SOL)$108.35-2.96%
  • tronTRON(TRX)$0.3401190.76%
  • zcashZcash(ZEC)$1,448.76-7.32%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.02-1.32%
  • HyperliquidHyperliquid(HYPE)$91.11-1.93%
  • dogecoinDogecoin(DOGE)$0.085041-2.44%
  • moneroMonero(XMR)$522.22-8.59%
  • whitebitWhiteBIT Coin(WBT)$81.77-1.68%
  • USDSUSDS(USDS)$1.000.00%
  • RainRain(RAIN)$0.013394-2.23%
  • chainlinkChainlink(LINK)$11.98-2.99%
  • cardanoCardano(ADA)$0.219965-1.28%
  • leo-tokenLEO Token(LEO)$8.900.16%
  • stellarStellar(XLM)$0.190333-1.59%
  • uniswapUniswap(UNI)$8.67-5.30%
  • bitcoin-cashBitcoin Cash(BCH)$246.25-0.45%
  • Ethena USDeEthena USDe(USDE)$1.00-0.01%
  • daiDai(DAI)$1.000.00%
  • nearNEAR Protocol(NEAR)$3.44-7.70%
  • litecoinLitecoin(LTC)$56.82-0.69%
  • USD1USD1(USD1)$1.000.00%
  • avalanche-2Avalanche(AVAX)$9.7214.47%
  • CantonCanton(CC)$0.104236-5.30%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.381.02%
  • MemeCoreMemeCore(M)$1.6729.85%
  • hedera-hashgraphHedera(HBAR)$0.0808072.44%
  • suiSui(SUI)$0.820.48%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • shiba-inuShiba Inu(SHIB)$0.000005-0.39%
  • crypto-com-chainCronos(CRO)$0.058594-0.87%
  • BittensorBittensor(TAO)$252.78-0.80%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.01%
  • tether-goldTether Gold(XAUT)$4,368.43-0.07%
  • Circle USYCCircle USYC(USYC)$1.140.00%
  • okbOKB(OKB)$115.49-0.77%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.57%
  • aaveAave(AAVE)$137.00-4.30%
  • OndoOndo(ONDO)$0.4083122.86%
  • AsterAster(ASTER)$0.74-2.59%
  • EthenaEthena(ENA)$0.1949538.79%
  • mantleMantle(MNT)$0.59-2.47%
  • pax-goldPAX Gold(PAXG)$4,360.83-0.05%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

ByteDance Unveils ToolTrain: A New Tool-Integrated Reinforcement Learning RL Framework that Redefines Repo Deep Search

August 14, 2025
in AI & Technology
Reading Time: 3 mins read
A A
ByteDance Unveils ToolTrain: A New Tool-Integrated Reinforcement Learning RL Framework that Redefines Repo Deep Search
ShareShareShareShareShare

Issue localization involves identifying exact code locations that require modification to fix software problems, a process that often demands significant manual effort from developers, especially in large repositories. Due to its complexity and time-intensive nature, automating this task has become a key research focus. LLM-based agents enable language models to use various tools for dynamic repository exploration. However, these models face challenges in performing Repo Deep Search, a sequential navigation task that requires multi-step reasoning and effective tool usage. Current LLMs struggle with these high demands, often resulting in incorrect tool calls or a breakdown in maintaining coherent reasoning chains during the exploration process.

Existing work includes fault localization and agentic training. In fault localization, methods like DeepFL and DeepRL4FL utilize deep neural networks and CNNs to identify faulty code by analyzing test coverage, data dependencies, and static code representations. More recent advancements include LLMs, such as Agentless, to narrow down code locations. LLMs often lack the complexity needed for complex reasoning and tool usage in repository exploration. To address this, agentic training methods, such as SWE-Gym and SEAlign, fine-tune LLMs using high-quality trajectories. Another approach, LocAgent, constructs the ground truth for issue localization based on functions modified by golden patches from GitHub.

YOU MAY ALSO LIKE

How Long Can You Expect Your Old Cassette Tapes To Last?

How To Record Audio On Your iPhone

Researchers from Peking University, ByteDance, and Beijing Institute of Technology have proposed ToolTrain, a tool-integrated training framework to enhance the multi-hop reasoning capabilities of LLMs during issue localization. ToolTrain introduces RepoSearcher, a lightweight agent equipped with simple retrieval tools that enable LLMs to locate function or class definitions by name. To help the LLMs use these tools for multi-hop reasoning, the researchers construct labeled data from open-source repositories and follow a two-stage process: rejection-sampled SFT and tool-integrated RL. This approach ensures the model learns to use tools strategically, avoiding redundant explorations while focusing on promising code paths.

Researchers construct their evaluation dataset using SWE-Bench-Verified, a benchmark derived from real GitHub issues and manually verified by professional developers. This dataset provides ground-truth answers for issue localization by identifying functions and files modified in golden patches. To evaluate RepoSearcher’s performance, metrics such as Recall@k, MAP, MRR, nDCG@k, and %Resolved are used. Moreover, ToolTrain is applied to two models, Qwen-7B and Qwen-32B, which are then compared against four state-of-the-art frameworks: Agentless, CrcaLoca, CoSIL, and LocAgent. These baselines represent diverse design philosophies, ensuring a detailed evaluation of ToolTrain’s effectiveness for precise and strategic code exploration.

The RepoSearcher with ToolTrain achieves state-of-the-art performance among models of similar size and even outperforms larger commercial models on specific metrics. For instance, RepoSearcher with ToolTrain-32B achieves a function-level Recall@5 score of 68.55, surpassing Claude-3.7-Sonnet (66.38). The 7B-parameter model outperforms other frameworks using 32B models, enhancing the tool-calling capabilities of ToolTrain in smaller models. In issue resolution, RepoSearcher with ToolTrain-7B achieves a Recall@5 of 62.38 and a resolution rate of 14.00, the best among 7B models. However, disparities arise when using different patch generation models, as seen in the resolution rates of 14.00 (ToolTrain-7B) versus 31.60 (ToolTrain-32B), despite similar localization results.

In conclusion, researchers introduced ToolTrain to enhance the issue localization of LLMs. By combining SFT with RL, ToolTrain equips models like RepoSearcher to navigate code repositories effectively and perform precise multi-hop reasoning. Evaluated on real-world benchmarks, ToolTrain-trained models achieve state-of-the-art performance among similarly sized models and even outperform larger commercial models like Claude-3.7 on specific tasks. This shows its ability to optimize tool usage and reasoning in smaller models, reducing redundancy and improving efficiency. The study emphasizes the potential of ToolTrain to transform issue localization and provide an efficient solution for complex software challenges.


Check out the Paper and GitHub Page here. Feel free to check out our GitHub Page for Tutorials, Codes and Notebooks. Also, feel free to follow us on Twitter and don’t forget to join our 100k+ ML SubReddit and Subscribe to our Newsletter.

🇬 Star us on GitHub
🇸 Sponsor us

The post ByteDance Unveils ToolTrain: A New Tool-Integrated Reinforcement Learning RL Framework that Redefines Repo Deep Search appeared first on MarkTechPost.

Credit: Source link

ShareTweetSendSharePin

Related Posts

How Long Can You Expect Your Old Cassette Tapes To Last?
AI & Technology

How Long Can You Expect Your Old Cassette Tapes To Last?

September 20, 2026
How To Record Audio On Your iPhone
AI & Technology

How To Record Audio On Your iPhone

September 20, 2026
What Is The Difference Between Apple CarPlay And CarPlay Ultra?
AI & Technology

What Is The Difference Between Apple CarPlay And CarPlay Ultra?

September 19, 2026
The Pros And Cons Of Using Wired Vs. Wireless Xbox Controllers
AI & Technology

The Pros And Cons Of Using Wired Vs. Wireless Xbox Controllers

September 19, 2026
Next Post
MongoDB Is Closer To Its GARP Moment

MongoDB Is Closer To Its GARP Moment

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Legal showdown over pro players in college football

Legal showdown over pro players in college football

September 18, 2026
Condoleezza Rice recounts her memories from the 9/11 attacks

Condoleezza Rice recounts her memories from the 9/11 attacks

September 13, 2026
Trump Proposes Renaming Artificial Intelligence, Announces AI Force – Unite.AI

Trump Proposes Renaming Artificial Intelligence, Announces AI Force – Unite.AI

September 19, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!