HE
heygen

HeyGen: Avatar IV

heygen/avatar-iv
تاریخ انتشار
۱۴۰۵/۶/۲
تعداد ارائه‌دهنده
۱
ورودی‌ها
متن، تصویر، صوت
خروجی‌ها
ویدیو

جزئیات قیمت

Video Output۱۵٬۰۰۰تومان / ثانیه

درباره مدل

HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal tone, rhythm, and emotion of the audio to drive head motion, facial expression, and gestures, producing output at up to 1080p. The spoken audio comes from one of two inputs: a text script, which the model voices with HeyGen text-to-speech, or a supplied audio track, which the image is lip-synced to directly. Passthrough parameters let you choose a voice, tune voice settings, set expressiveness, prompt specific motion, replace or remove the background, add captions, and title the video.

قابلیت‌های مدل

ورودی‌های پشتیبانی‌شده
متن
تصویر
صوت
خروجی‌های پشتیبانی‌شده
ویدیو
معماری ورودی/خروجیtext,image,audio->video

ارائه‌دهندگان (۱)

ارائه‌دهندهورودی (تومان / ۱M)خروجی (تومان / ۱M)پنجره متنحداکثر خروجیکوانتیزهوضعیت
HeyGen—————فعال

استفاده از طریق API

شناسه این مدل را در درخواست‌های API هوشگر استفاده کنید:

POST https://api.hooshgar.ir/v1/chat/completions
{
  "model": "heygen/avatar-iv",
  "messages": [
    {
      "role": "user",
      "content": "سلام"
    }
  ]
}
مستندات کامل API

مدل‌های مرتبط

اطلاعات مدل بر اساس داده‌های OpenRouter