Discrete vs. Continuous: A Comprehensive Study of Unified Audio Understanding in LALMs
arXiv:2609.22851v2 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) utilize either continuous features or discrete tokens, yet the optimal representation paradigm for general audio understanding remains debated. Existing benchmarks often focus on narrow domains or evaluate…