arxiv
PublishedSeptember 22, 2026 at 4:00 AM
—neutral
MENASpeechBank: A Reference Voice Bank with Persona-Conditioned Multi-Turn Conversations for AudioLLMs
Publisher summary· verbatim
arXiv:2602.07036v2 Announce Type: replace-cross Abstract: Audio large language models (AudioLLMs) enable instruction following over speech and general audio, but progress is limited by the scarcity of diverse, conversational, and instruction-aligned speech--text data. This gap is particularly pronou
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivLoRA Enhanced Contrastive Learning with SAS Vision Transformers39marxivHallucination-R1: Robustness-Oriented Paraphrase Generation for Factual Consistency39marxivRuntime Authorization for Resources Acquired by AI Agents39marxivLearning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications39mThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗