Researchers have developed RefactorPlatform, an open-source tool designed to evaluate the performance of AI agents in large-scale code refactoring tasks. This platform standardizes the evaluation environment, allowing for controlled comparisons of different design choices, including model backbones like GitHub Copilot CLI and OpenRouter, various execution strategies, and prompt variations. Initial tests on 100 multi-file tasks demonstrated that AST-aware chunking improves accuracy by 25-30% over naive token-window chunking, and a single retrieval-augmented agent outperformed a multi-agent delegation approach. AI
IMPACT Enables reproducible evaluation of AI code refactoring agents, potentially accelerating development and improving their reliability.
RANK_REASON The cluster describes a new open-source research platform and its initial evaluation results, published on arXiv. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- CatalyzeX Code Finder for Papers
- CORE Recommender
- DagsHub
- GitHub Copilot CLI
- Gotit.pub
- Hugging Face
- Influence Flower
- OpenRouter
- RefactorBench
- RefactorPlatform
- ScienceCast
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →