Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments
Scaling coding agents requires a continuing supply of executable data for training, benchmarking, and continuous evaluation. Each task must couple a realistic software state with a specification, development tools, and reliable verification. To expand this supply, we present Change2Task, a system grounded in repository history that converts merged pull requests into verified tasks on healthy modern revisions of the same repository. It aligns historical evidence with evolved code, reconstructs ta
Record details
Published: 30 July 2026
Source: arXiv cs.LG
Category: Research
Topics: Healthcare · Agents & autonomy · Environment
Retrieved: 31 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control
arXiv cs.CY · 27 July 2026
DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents
arXiv cs.AI · 24 July 2026
Clinical Pathways as Safety Specifications for Physical AI in Hospital Wards
arXiv cs.CY · 23 July 2026
ResidencyRL: Reinforcement Learning in Simulated Clinical Environments
arXiv · 7 August 2026
SR-Agent: An Experience-Driven Agentic Framework for Post-Ranking Strategies Refinement in E-Commerce Recommendation
arXiv · 20 July 2026
CTBench: Evaluating Troubleshooting Capabilities of AI Agents in Realistic Telecom Network Operations
arXiv cs.AI · 12 August 2026
How to cite this record
ethics.ai (30 July 2026), “Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments,” evidence record 15251, https://ethics.ai/record/15251 (originally published by arXiv cs.LG).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.