EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs
Audio Large Language Models (ALLMs) are highly vulnerable to real-world noise, which often induces severe semantic drift and hallucinations. Existing robustness methods primarily rely on waveform-level acoustic enhancement, answer-level supervision, or the internal suppression of noise representations. To address these issues, we propose echodistill, an alignment-based noisy-to-clean self-distillation framework. Echodistill leverages a frozen clean-audio teacher to provide semantic references fo
Record details
Published: 11 May 2026
Source: arXiv
Category: Research
Topics: Safety & alignment
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance
arXiv · 11 May 2026
Guided Streaming Stochastic Interpolant Policy
arXiv · 11 May 2026
Metis: Learning to Jailbreak LLMs via Self-Evolving Metacognitive Policy Optimization
arXiv · 11 May 2026
ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design
arXiv · 11 May 2026
TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment
arXiv · 11 May 2026
E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability
arXiv · 11 May 2026
How to cite this record
ethics.ai (11 May 2026), “EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs,” evidence record 4588, https://ethics.ai/record/4588 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.