יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה MarkTechPost ·

Alibaba Qwen משחרר Qwen-Audio-3.1-Realtime: מודל קולי מלא-דו-כיווני

Alibaba Qwen Releases Qwen-Audio-3.1-Realtime: A Full-Duplex Voice Model Trained to Think, Act, and Decide When to Speak
Alibaba Qwen השיק Qwen-Audio-3.1-Realtime, מודל קולי מלא-דו-כיווני לגורמי קול. המודל כולל 2 מודלים: חלוקת תפקידים ומודל רנדר קולי. המודל ניתן להפעלה כ-API ניהולית. קיימים 3 רמות: חשוב, פעול, ודבר. המודל ניתן להפעלה כ-API ניהולית.
תקציר מקורי באנגליתAlibaba’s Qwen team has released Qwen-Audio-3.1 , a 5-model audio stack spanning ASR, TTS and realtime interaction. The main model is Qwen-Audio-3.1-Realtime , a full-duplex speech model built for voice agents that call tools. Qwen also cut prices : about 85% on Realtime, about 70% on TTS and up to 95% on ASR. Is it deployable? Yes, as a managed API. qwen-audio-3.1-realtime-plus is live on QwenCloud over WebSocket. No open weights were announced. What Ships on QwenCloud The model page lists text and audio as both input and output. Context is 262K tokens, with 245K max input and 16K max output. Default limits are 60 requests and 100K tokens per minute. Pricing is $6.4 per 1M audio input tokens and $0.8 per 1M text input tokens. Text and audio output costs $24 per 1M tokens, with outpu
קרא במקור המקורי