OP
openai

OpenAI: Whisper Large V3

openai/whisper-large-v3
تاریخ انتشار
۱۴۰۵/۲/۱۱
تعداد ارائه‌دهنده
۳
ورودی‌ها
صوت
خروجی‌ها
transcription

جزئیات قیمت

Audio Seconds۲٫۲۵تومان / ثانیه

درباره مدل

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm, flac, and ogg. With 1,550M parameters, it achieves a 10.3% word error rate and is well-suited for noise-robust, multilingual transcription in demanding conditions. Supports timestamp granularities at word and segment levels.

قابلیت‌های مدل

ورودی‌های پشتیبانی‌شده
صوت
خروجی‌های پشتیبانی‌شده
transcription
معماری ورودی/خروجیaudio->transcription

ارائه‌دهندگان (۳)

ارائه‌دهندهورودی (تومان / ۱M)خروجی (تومان / ۱M)پنجره متنحداکثر خروجیکوانتیزهوضعیت
DeepInfra—————فعال
Together—————فعال
Groq—————فعال

استفاده از طریق API

شناسه این مدل را در درخواست‌های API هوشگر استفاده کنید:

POST https://api.hooshgar.ir/v1/chat/completions
{
  "model": "openai/whisper-large-v3",
  "messages": [
    {
      "role": "user",
      "content": "سلام"
    }
  ]
}
مستندات کامل API

مدل‌های مرتبط

اطلاعات مدل بر اساس داده‌های OpenRouter