Discrete vs. Continuous: A Comprehensive Study of Unified Audio Understanding in LALMs

arXiv:2609.22851v2 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) utilize either continuous features or discrete tokens, yet the optimal representation paradigm for general audio understanding remains debated. Existing benchmarks often focus on narrow domains or evaluate…

science

Sources