Qwen3-VL
Found in 4 packages: transformers, vllm, comfyui, unsloth
transformers(3 releases)
v5.14.0BreakingThis release introduces the Inkling multimodal model and TIPSv2, alongside significant improvements in kernel performance, generation capabilities, and bug fixes across various components. It also includes breaking changes for GPTNeoX and GPTBigCode related to weight remapping and attention backend compatibility.
v5.11.0This release introduces two major new models, DiffusionGemma and DeepSeek-V3.2, alongside significant enhancements to the custom kernel API for better fusion capabilities. Numerous bug fixes address issues across various models, CI stability, and dependency compatibility.
v4.57.0This release introduces support for several next-generation model architectures, including the high-efficiency Qwen3-Next and Qwen3-VL series, the privacy-focused VaultGemma, and the high-speed Longcat Flash MoE.
vllm(3 releases)
v0.24.0Breakingv0.24.0 introduces extensive support and performance optimizations for new models like MiniMax-M3 and DeepSeek-V4, matures the Model Runner V2 with default quantization support, and overhauls device selection by removing internal use of CUDA_VISIBLE_DEVICES.
v0.20.2vLLM v0.20.2 is a small patch release focused on bug fixes for DeepSeek V4, gpt-oss, and Qwen3-VL models.
v0.18.0Breakingv0.18.0 introduces major features like gRPC serving, GPU-less render serving, and significant improvements to KV cache offloading and Elastic Expert Parallelism. Ray is now an optional dependency, and numerous model-specific fixes and kernel optimizations have been integrated.
comfyui(1 releases)
unsloth(1 releases)
Track Symbol Changes
Use the Change8 MCP server or GitHub Action to get notified when Qwen3-VL changes.
Learn More