STAR-PólyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision
Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, long-horizon reasoning, existing systems continue to suffer from fundamental reliability issues: hallucination accumulation, memory fragmentation, and imbalanced reasoning-tool trade-offs. In this paper, we introduce STAR-PólyaMath, a multi-agent framework that systematically addresses these challenges through meta-level supervision and structured R
Record details
Published: 19 May 2026
Source: arXiv
Category: Research
Topics: Agents & autonomy
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
SkillEvolver: Skill Learning as a Meta-Skill
arXiv · 11 May 2026
The Meta-Agent Challenge: Are Current Agents Capable of Autonomous Agent Development?
arXiv · 3 June 2026
Knowledge Reutilization in Meta-Reinforcement Learning
arXiv · 16 June 2026
Owner-Harm: A Missing Threat Model for AI Agent Safety
arXiv · 20 April 2026
Counsel: A Meta-Evaluation Dataset for Agentic Tasks
arXiv · 19 June 2026
Eligibility-Aware Evidence Synthesis: An Agentic Framework for Clinical Trial Meta-Analysis
arXiv · 3 April 2026
How to cite this record
ethics.ai (19 May 2026), “STAR-PólyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision,” evidence record 4054, https://ethics.ai/record/4054 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.