Skip to main content

    Alibaba

    Qwen3-VL

    Qwen3-VL is Alibaba's September 2025 vision-language model family, supporting images, video and OCR across 32 languages. The repository documents 256K native context and extension to 1M. Model Studio has scheduled multiple Qwen3-VL hosted endpoints for retirement on October 10, 2026 at 00:00 UTC+8, subject to rollout timing. That notice does not withdraw the published model weights.

    256K tokens
    Multimodal
    Pro
    Provider
    Alibaba
    Context
    256K
    Catalog · As of Aug 24, 2026Within 30d
    yno subscription tier
    Pro
    Released
    Sep 23, 2025
    Speed
    Medium
    Reasoning
    Advanced
    Modality
    Multimodal

    Benchmarks

    Quality

    AA Intelligence Index
    13.4points
    Artificial Analysis · Date unknownAge unknown
    LiveCodeBench
    64.6%
    Artificial Analysis · Date unknownAge unknown
    GPQA Diamond
    77.2%
    Artificial Analysis · Date unknownAge unknown

    Speed and latency

    Not available

    API pricing

    API input cost
    $0.40/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    API output cost
    $4.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown

    Benchmark variant: Qwen3 VL 235B A22B (Reasoning)

    AA data retrieved Sep 11, 2026 · Artificial Analysis

    Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology

    Capabilities

    • Vision
    • Video
    • OCR
    • Document Analysis
    • Multilingual

    Use cases

    • Document processing
    • Video Q&A
    • Multilingual OCR
    • Visual agents

    Strengths

    • Image and video understanding
    • OCR across 32 languages
    • 256K native context; extensible to 1M
    • Published vision-language weights

    Best for

    Teams deploying Qwen3-VL weights for documents or video and reviewing hosted migration needs

    Use Qwen3-VL in yno.ai

    No credit card required

    Related models

    Frequently asked

    What is Qwen3-VL?
    Qwen3-VL is Alibaba's September 2025 vision-language model family, supporting images, video and OCR across 32 languages. The repository documents 256K native context and extension to 1M. Model Studio has scheduled multiple Qwen3-VL hosted endpoints for retirement on October 10, 2026 at 00:00 UTC+8, subject to rollout timing. That notice does not withdraw the published model weights.
    How much does Qwen3-VL cost in yno.ai?
    Qwen3-VL is available on the Pro tier of yno.ai.
    What can Qwen3-VL do?
    Qwen3-VL is best for Teams deploying Qwen3-VL weights for documents or video and reviewing hosted migration needs. Its main capabilities include Vision, Video, OCR, Document Analysis, Multilingual.
    How does Qwen3-VL compare to other models?
    Qwen3-VL excels at Image and video understanding and is recommended for Document processing, Video Q&A, Multilingual OCR, Visual agents. See related models below.
    How do I use Qwen3-VL in yno.ai?
    Sign up for yno.ai, select Qwen3-VL from the model picker, and start chatting.