Alibaba
Qwen3-VL
Qwen3-VL is Alibaba's September 2025 vision-language model family, supporting images, video and OCR across 32 languages. The repository documents 256K native context and extension to 1M. Model Studio has scheduled multiple Qwen3-VL hosted endpoints for retirement on October 10, 2026 at 00:00 UTC+8, subject to rollout timing. That notice does not withdraw the published model weights.
- Provider
- Alibaba
- Context
- 256KCatalog · As of Aug 24, 2026Within 30d
- yno subscription tier
- Pro
- Released
- Sep 23, 2025
- Speed
- Medium
- Reasoning
- Advanced
- Modality
- Multimodal
Benchmarks
Quality
- AA Intelligence Index
- LiveCodeBench
- 64.6%Artificial Analysis · Date unknownAge unknown
- GPQA Diamond
- 77.2%Artificial Analysis · Date unknownAge unknown
Speed and latency
Not available
API pricing
- API input cost
- API output cost
Benchmark variant: Qwen3 VL 235B A22B (Reasoning)
AA data retrieved Sep 11, 2026 · Artificial Analysis
Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology
Capabilities
- Vision
- Video
- OCR
- Document Analysis
- Multilingual
Use cases
- Document processing
- Video Q&A
- Multilingual OCR
- Visual agents
Strengths
- Image and video understanding
- OCR across 32 languages
- 256K native context; extensible to 1M
- Published vision-language weights
Best for
Teams deploying Qwen3-VL weights for documents or video and reviewing hosted migration needs
No credit card required
Related models
Frequently asked
- What is Qwen3-VL?
- Qwen3-VL is Alibaba's September 2025 vision-language model family, supporting images, video and OCR across 32 languages. The repository documents 256K native context and extension to 1M. Model Studio has scheduled multiple Qwen3-VL hosted endpoints for retirement on October 10, 2026 at 00:00 UTC+8, subject to rollout timing. That notice does not withdraw the published model weights.
- How much does Qwen3-VL cost in yno.ai?
- Qwen3-VL is available on the Pro tier of yno.ai.
- What can Qwen3-VL do?
- Qwen3-VL is best for Teams deploying Qwen3-VL weights for documents or video and reviewing hosted migration needs. Its main capabilities include Vision, Video, OCR, Document Analysis, Multilingual.
- How does Qwen3-VL compare to other models?
- Qwen3-VL excels at Image and video understanding and is recommended for Document processing, Video Q&A, Multilingual OCR, Visual agents. See related models below.
- How do I use Qwen3-VL in yno.ai?
- Sign up for yno.ai, select Qwen3-VL from the model picker, and start chatting.