QW
qwen
Qwen: Qwen3 VL 235B A22B Instruct
qwen/qwen3-vl-235b-a22b-instructفراخوانی ابزار
خروجی ساختاریافته
خروجی JSON
خروجی تکرارپذیر (seed)
پنجره متن
۲۶۲٬۱۴۴ توکن
حداکثر خروجی
۱۶٬۳۸۴ توکن
قیمت ورودی
۶۰٬۰۰۰ تومان / ۱ میلیون توکن
قیمت خروجی
۲۶۴٬۰۰۰ تومان / ۱ میلیون توکن
تاریخ انتشار
۱۴۰۴/۷/۱
تعداد ارائهدهنده
۵
ورودیها
متن، تصویر
خروجیها
متن
جزئیات قیمت
Input Price۶۰٬۰۰۰تومان / ۱ میلیون توکن
Output Price۲۶۴٬۰۰۰تومان / ۱ میلیون توکن
Cache Read۳۳٬۰۰۰تومان / ۱ میلیون توکن
درباره مدل
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table extraction, multilingual OCR). The series emphasizes robust perception (recognition of diverse real-world and synthetic categories), spatial understanding (2D/3D grounding), and long-form visual comprehension, with competitive results on public multimodal benchmarks for both perception and reasoning.
Beyond analysis, Qwen3-VL supports agentic interaction and tool use: it can follow complex instructions over multi-image, multi-turn dialogues; align text to video timelines for precise temporal queries; and operate GUI elements for automation tasks. The models also enable visual coding workflows—turning sketches or mockups into code and assisting with UI debugging—while maintaining strong text-only performance comparable to the flagship Qwen3 language models. This makes Qwen3-VL suitable for production scenarios spanning document AI, multilingual OCR, software/UI assistance, spatial/embodied tasks, and research on vision-language agents.
قابلیتهای مدل
ورودیهای پشتیبانیشده
متن
تصویر
خروجیهای پشتیبانیشده
متن
قابلیتها
فراخوانی ابزار
خروجی ساختاریافته
خروجی JSON
خروجی تکرارپذیر (seed)
معماری ورودی/خروجیtext,image->text
پارامترهای پشتیبانیشده
max_tokenstemperaturetop_pstopfrequency_penaltypresence_penaltyrepetition_penaltytop_kseedmin_plogit_biasstructured_outputstoolstool_choiceresponse_formatارائهدهندگان (۵)
| ارائهدهنده | ورودی (تومان / ۱M) | خروجی (تومان / ۱M) | پنجره متن | حداکثر خروجی | کوانتیزه | وضعیت |
|---|---|---|---|---|---|---|
| DeepInfra | ۶۰٬۰۰۰ | ۲۶۴٬۰۰۰ | ۲۶۲٬۱۴۴ | ۱۶٬۳۸۴ | fp8 | فعال |
| Venice | ۶۳٬۰۰۰ | ۵۷۰٬۰۰۰ | ۱۲۸٬۰۰۰ | ۱۶٬۳۸۴ | fp8 | فعال |
| Parasail | ۶۳٬۰۰۰ | ۵۷۰٬۰۰۰ | ۱۳۱٬۰۷۲ | ۳۲٬۷۶۸ | fp8 | فعال |
| Alibaba | ۷۸٬۰۰۰ | ۳۱۲٬۰۰۰ | ۱۳۱٬۰۷۲ | ۳۲٬۷۶۸ | — | فعال |
| Novita | ۹۰٬۰۰۰ | ۴۵۰٬۰۰۰ | ۱۳۱٬۰۷۲ | ۳۲٬۷۶۸ | bf16 | اختلال |
استفاده از طریق API
این مدل را با همان کلید API هوشگر و از طریق نقطه پایانی سازگار با OpenAI فراخوانی کنید:
POST https://api.hooshgar.ir/v1/chat/completions{
"model": "qwen/qwen3-vl-235b-a22b-instruct",
"messages": [
{
"role": "user",
"content": "سلام"
}
]
}مستندات کامل APIمدلهای مرتبط
Qwen: Qwen3.8 Max Prime
qwen/qwen3.8-max-prime
۱٬۰۰۰٬۰۰۰ توکن
Qwen: Qwen3.8 Omni Flash
qwen/qwen3.8-omni-flash
۱٬۰۰۰٬۰۰۰ توکن
Qwen: Qwen3.8 Max (0902)
qwen/qwen3.8-max-0902
۱٬۰۰۰٬۰۰۰ توکن
Qwen: Qwen3.8 Flash
qwen/qwen3.8-flash
۱٬۰۰۰٬۰۰۰ توکن
Qwen: Qwen3.8 27B (free)
qwen/qwen3.8-27b:free
۲۶۲٬۱۴۴ توکن
Qwen: Qwen3.8 27B
qwen/qwen3.8-27b
۲۶۲٬۱۴۴ توکن
اطلاعات مدل بر اساس دادههای OpenRouter
