GO
google

Google: Gemini 3.1 Flash TTS Preview

google/gemini-3.1-flash-tts-preview
پنجره متن
۳۲٬۷۶۸ توکن
حداکثر خروجی
۱۶٬۳۸۴ توکن
قیمت ورودی
۳۰۰٬۰۰۰ تومان / ۱ میلیون توکن
قیمت خروجی
۶٬۰۰۰٬۰۰۰ تومان / ۱ میلیون توکن
تاریخ انتشار
۱۴۰۵/۲/۴
تعداد ارائه‌دهنده
۱
ورودی‌ها
متن
خروجی‌ها
speech

جزئیات قیمت

Input Price۳۰۰٬۰۰۰تومان / M text tokens
Output Price۶٬۰۰۰٬۰۰۰تومان / M audio tokens

درباره مدل

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output across 70+ languages — nearly 3× the language coverage of its predecessor. The headline addition is a system of 200+ inline audio tags (e.g. `[whispers]`, `[laughs]`, `[excited]`) that let developers steer delivery, emotion, and pacing mid-sentence, alongside a "director's chair" workflow in Google AI Studio for defining per-character Audio Profiles and scene-level context. It supports up to two speakers with independent voice and style configuration per speaker, outputs PCM audio at 24 kHz / 16-bit mono, and automatically watermarks all output with SynthID. Context window is 32k tokens.

قابلیت‌های مدل

ورودی‌های پشتیبانی‌شده
متن
خروجی‌های پشتیبانی‌شده
speech
معماری ورودی/خروجیtext->speech

ارائه‌دهندگان (۱)

ارائه‌دهندهورودی (تومان / ۱M)خروجی (تومان / ۱M)پنجره متنحداکثر خروجیکوانتیزهوضعیت
Google۳۰۰٬۰۰۰۶٬۰۰۰٬۰۰۰۳۲٬۷۶۸۱۶٬۳۸۴—فعال

استفاده از طریق API

شناسه این مدل را در درخواست‌های API هوشگر استفاده کنید:

POST https://api.hooshgar.ir/v1/chat/completions
{
  "model": "google/gemini-3.1-flash-tts-preview",
  "messages": [
    {
      "role": "user",
      "content": "سلام"
    }
  ]
}
مستندات کامل API

مدل‌های مرتبط

اطلاعات مدل بر اساس داده‌های OpenRouter