QW
qwen

Qwen: Qwen3 ASR Flash

qwen/qwen3-asr-flash-2026-02-10
تاریخ انتشار
۱۴۰۵/۲/۲۴
تعداد ارائه‌دهنده
۱
ورودی‌ها
صوت
خروجی‌ها
transcription

جزئیات قیمت

Audio Seconds۱۰تومان / ثانیه

درباره مدل

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages — including Chinese (with Cantonese, Sichuanese, Minnan, and Wu dialects), English, Arabic, French, German, Spanish, Italian, Portuguese, Russian, Japanese, and Korean — with automatic language detection so no manual configuration is needed for mixed-language audio. The model is designed for difficult acoustic conditions: it transcribes lyrics over background music, handles noisy and far-field recordings, filters silence and non-speech audio, and accepts arbitrary context text (names, jargon, domain terminology) to bias recognition toward specific vocabulary.

قابلیت‌های مدل

ورودی‌های پشتیبانی‌شده
صوت
خروجی‌های پشتیبانی‌شده
transcription
معماری ورودی/خروجیaudio->transcription

ارائه‌دهندگان (۱)

ارائه‌دهندهورودی (تومان / ۱M)خروجی (تومان / ۱M)پنجره متنحداکثر خروجیکوانتیزهوضعیت
Alibaba—————فعال

استفاده از طریق API

شناسه این مدل را در درخواست‌های API هوشگر استفاده کنید:

POST https://api.hooshgar.ir/v1/chat/completions
{
  "model": "qwen/qwen3-asr-flash-2026-02-10",
  "messages": [
    {
      "role": "user",
      "content": "سلام"
    }
  ]
}
مستندات کامل API

مدل‌های مرتبط

اطلاعات مدل بر اساس داده‌های OpenRouter