Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
| Тип | Модель |
| Категория | Модели и платформы / text+image->text |
| Цена | только платно (от $0.12/мес) |
| Платформа | Своё развёртывание |
| Системы | api, python, self-hosted |
| Хостинг | cloud |
| Язык сайта | en |
| Вендор | Qwen |
| Запуск | 2025-10-14 |