Math Education Digital Shadows for Investigating Learning with GenAI: Mathematics Performance, Anxiety, and Confidence in LLMs
arXiv:2604.27618v2 Announce Type: replace-cross Abstract: Understanding the impact of large language models (LLMs) on mathematics education requires data on LLMs' mathematical performance and biases. To this end, we introduce Math Education Digital Shadows (MEDS), a dataset mapping how LLMs reason about mathematics across human- and AI-like personifications. MEDS comprises 28,000 runs from 14 LLMs (i.e., Mistral, Qwen, DeepSeek, IBM Granite, Microsoft Phi, and xAI Grok) generated under human-sha
Record details
Published: 27 July 2026
Source: arXiv cs.CY
Category: Research
Topics: Children & education · Finance, VC & PE
Retrieved: 27 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Do language families matter? Evaluating LLMs for sentiment analysis through a hierarchical cross-lingual lens
Frontiers in Artificial Intelligence · 24 July 2026
The Shibboleth Effect: Auditing the Cross-Lingual Distributional Skew of Large Language Models
arXiv · 9 June 2026
Understanding Content Moderation in Large Language Models through Restricted Books: From Refusal to Warning
arXiv · 12 August 2026
Can AI Make Conflicts Worse? An Alignment Failure in LLM Deployment Across Conflict Contexts
arXiv · 21 May 2026
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
arXiv · 5 May 2026
Generative AI in Academic Writing: A Comparison of DeepSeek, Qwen, ChatGPT, Gemini, Llama, Mistral, and Gemma
OpenAlex · 30 April 2026
How to cite this record
ethics.ai (27 July 2026), “Math Education Digital Shadows for Investigating Learning with GenAI: Mathematics Performance, Anxiety, and Confidence in LLMs,” evidence record 13594, https://ethics.ai/record/13594 (originally published by arXiv cs.CY).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.