QW
qwen

Qwen: Qwen3 VL 235B A22B Instruct

qwen/qwen3-vl-235b-a22b-instruct
فراخوانی ابزار
خروجی ساختاریافته
خروجی JSON
خروجی تکرارپذیر (seed)
پنجره متن
۲۶۲٬۱۴۴ توکن
حداکثر خروجی
۱۶٬۳۸۴ توکن
قیمت ورودی
۶۰٬۰۰۰ تومان / ۱ میلیون توکن
قیمت خروجی
۲۶۴٬۰۰۰ تومان / ۱ میلیون توکن
تاریخ انتشار
۱۴۰۴/۷/۱
تعداد ارائه‌دهنده
۵
ورودی‌ها
متن، تصویر
خروجی‌ها
متن

جزئیات قیمت

Input Price۶۰٬۰۰۰تومان / ۱ میلیون توکن
Output Price۲۶۴٬۰۰۰تومان / ۱ میلیون توکن
Cache Read۳۳٬۰۰۰تومان / ۱ میلیون توکن

درباره مدل

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table extraction, multilingual OCR). The series emphasizes robust perception (recognition of diverse real-world and synthetic categories), spatial understanding (2D/3D grounding), and long-form visual comprehension, with competitive results on public multimodal benchmarks for both perception and reasoning. Beyond analysis, Qwen3-VL supports agentic interaction and tool use: it can follow complex instructions over multi-image, multi-turn dialogues; align text to video timelines for precise temporal queries; and operate GUI elements for automation tasks. The models also enable visual coding workflows—turning sketches or mockups into code and assisting with UI debugging—while maintaining strong text-only performance comparable to the flagship Qwen3 language models. This makes Qwen3-VL suitable for production scenarios spanning document AI, multilingual OCR, software/UI assistance, spatial/embodied tasks, and research on vision-language agents.

قابلیت‌های مدل

ورودی‌های پشتیبانی‌شده
متن
تصویر
خروجی‌های پشتیبانی‌شده
متن
قابلیت‌ها
فراخوانی ابزار
خروجی ساختاریافته
خروجی JSON
خروجی تکرارپذیر (seed)
معماری ورودی/خروجیtext,image->text
پارامترهای پشتیبانی‌شده
max_tokenstemperaturetop_pstopfrequency_penaltypresence_penaltyrepetition_penaltytop_kseedmin_plogit_biasstructured_outputstoolstool_choiceresponse_format

ارائه‌دهندگان (۵)

ارائه‌دهندهورودی (تومان / ۱M)خروجی (تومان / ۱M)پنجره متنحداکثر خروجیکوانتیزهوضعیت
DeepInfra۶۰٬۰۰۰۲۶۴٬۰۰۰۲۶۲٬۱۴۴۱۶٬۳۸۴fp8فعال
Venice۶۳٬۰۰۰۵۷۰٬۰۰۰۱۲۸٬۰۰۰۱۶٬۳۸۴fp8فعال
Parasail۶۳٬۰۰۰۵۷۰٬۰۰۰۱۳۱٬۰۷۲۳۲٬۷۶۸fp8فعال
Alibaba۷۸٬۰۰۰۳۱۲٬۰۰۰۱۳۱٬۰۷۲۳۲٬۷۶۸—فعال
Novita۹۰٬۰۰۰۴۵۰٬۰۰۰۱۳۱٬۰۷۲۳۲٬۷۶۸bf16اختلال

استفاده از طریق API

این مدل را با همان کلید API هوشگر و از طریق نقطه پایانی سازگار با OpenAI فراخوانی کنید:

POST https://api.hooshgar.ir/v1/chat/completions
{
  "model": "qwen/qwen3-vl-235b-a22b-instruct",
  "messages": [
    {
      "role": "user",
      "content": "سلام"
    }
  ]
}
مستندات کامل API

مدل‌های مرتبط

اطلاعات مدل بر اساس داده‌های OpenRouter