arXiv cs.AI
· Papers
Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning
arXiv:2607.20166v1 Announce Type: cross Abstract: Large Audio Language models (LALMs) have made rapid progress on acoustic understanding, yet they still struggle with fine-grained audio reasoning (e.g., recognizing event order, repetitions and duration). Existing post-training methods heavily rely on expensive external