Skip to content
arXiv cs.AI · Papers

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning

arXiv:2607.20166v1 Announce Type: cross Abstract: Large Audio Language models (LALMs) have made rapid progress on acoustic understanding, yet they still struggle with fine-grained audio reasoning (e.g., recognizing event order, repetitions and duration). Existing post-training methods heavily rely on expensive external