NL2Scratch: An Executable Benchmark and Evaluation for Block-Based Programming
Block-based programming environments such as Scratch are widely used in early programming education, yet natural-language-to-code (NL2Code) research has focused primarily on text-based languages. Scratch programs are event-driven, visually compositional, and distributed across concurrent scripts, making conventional NL2Code assumptions and evaluation insufficient. We introduce NL2Scratch, an executable benchmark for natural-language-to-Scratch generation comprising 311,648 parser-valid NL--progr
Record details
Published: 20 June 2026
Source: arXiv
Category: Research
Topics: Children & education · Environment
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Self-Efficacy and Favorability Shape Learning from Tutoring Systems and Paper Practice
arXiv · 16 June 2026
Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillation
arXiv · 4 June 2026
Commenting with Copilot: A Taxonomy and Multi-Year Analysis of Student Code-Generation Specifications
arXiv cs.CY · 14 July 2026
Conceptualizations of GenAI and students’ professionalization: Within the multi-layered environment of learning for higher education
Computers and Education: Artificial Intelligence · 14 July 2026
Multi-Turn On-Policy Distillation with Prefix Replay
HuggingFace Daily Papers · 15 July 2026
Large Language Models in Architecture Studio: A Framework for Learning Outcomes
arXiv cs.CY · 21 July 2026
How to cite this record
ethics.ai (20 June 2026), “NL2Scratch: An Executable Benchmark and Evaluation for Block-Based Programming,” evidence record 746, https://ethics.ai/record/746 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.