Alibaba releases Qwen Audio 3.1, a new suite of speech models that cuts AI audio costs by up to 95 percent. The lineup includes five models covering automatic speech recognition, text‑to‑speech, and real‑time interaction, with ASR-Next adding multi‑speaker detection, timestamps, emotion, and ambient noise recognition. Multilingual synthesis and filler‑word cleanup improve usability across languages and dialects, making high‑quality audio AI more affordable for developers.
Opening Kapyn…