arXiv Computation and Language By Mohammad Javad Ranjbar Kalahroodi, Mohammad Amini, Parmis Bathayan, Heshaam Faili, Azadeh Shakery

PARSA-Bench: A Comprehensive Persian Audio-Language Model Benchmark

Read the original on arXiv Computation and Language →

PARSA‑Bench is the first dedicated benchmark for evaluating large audio‑language models on Persian, addressing unique challenges such as classical poetry, traditional music, and code‑switching. It comprises 16 tasks—10 of which are new—covering speech understanding, paralinguistic analysis, and culturally grounded audio reasoning. Across most tasks, text‑only baselines outperform audio‑based models, indicating that audio understanding remains the main limitation, except for Persian poetry where prosody provides additional information that audio beats text.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computation and Language.