Skip to main content

    Alibaba

    Qwen3.8-27B

    Alibaba's August 14, 2026 dense 27B vision-language model, with Apache 2.0 weights and optional thinking. It accepts text, images and video. Published weights have a native 262,144-token context, extensible to 1M with YaRN; QwenCloud provides a hosted 1M-context endpoint at $0.50/$3 per MTok (input/output). Hosted limits and prices are separate from local deployment requirements.

    262K tokens (up to 1M)
    Multimodal
    Pro
    New
    Provider
    Alibaba
    Context
    262K
    Catalog · As of Aug 18, 2026Within 30d
    yno subscription tier
    Pro
    Released
    Aug 14, 2026
    Speed
    Fast
    Reasoning
    Advanced
    Modality
    Multimodal

    Benchmarks

    Quality

    AA Intelligence Index
    33.9points
    Artificial Analysis · Date unknownAge unknown
    AA Coding Index
    68.1points
    Artificial Analysis · Date unknownAge unknown
    GPQA Diamond
    90.5%
    Artificial Analysis · Date unknownAge unknown

    Speed and latency

    Output speed
    44.45tokens/s
    Configuration: Prompt length: 1,000 · Parallel queries: 1
    Artificial Analysis · Date unknownAge unknown
    Time to first token
    1.27s
    Configuration: Prompt length: 1,000 · Parallel queries: 1
    Artificial Analysis · Date unknownAge unknown
    Time to first answer token
    46.26s
    Configuration: Prompt length: 1,000 · Parallel queries: 1
    Artificial Analysis · Date unknownAge unknown

    API pricing

    API input cost
    $0.50/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    API output cost
    $3.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown

    Benchmark variant: Qwen3.8 27B (xhigh)

    AA data retrieved Sep 11, 2026 · Artificial Analysis

    Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology

    Capabilities

    • Open Source
    • Code
    • Vision
    • Agentic
    • Local Deployment
    • Multilingual

    Use cases

    • Self-hosted coding agents
    • Local multimodal AI
    • Apache-2 licensed deployments
    • Edge agentic workloads

    Strengths

    • Apache 2.0 weights
    • Dense 27B model
    • Vision and optional thinking
    • 262K native context; extensible to 1M

    Best for

    Teams deploying a 27B vision-language model locally or through QwenCloud

    Use Qwen3.8-27B in yno.ai

    No credit card required

    Related models

    Frequently asked

    What is Qwen3.8-27B?
    Alibaba's August 14, 2026 dense 27B vision-language model, with Apache 2.0 weights and optional thinking. It accepts text, images and video. Published weights have a native 262,144-token context, extensible to 1M with YaRN; QwenCloud provides a hosted 1M-context endpoint at $0.50/$3 per MTok (input/output). Hosted limits and prices are separate from local deployment requirements.
    How much does Qwen3.8-27B cost in yno.ai?
    Qwen3.8-27B is available on the Pro tier of yno.ai.
    What can Qwen3.8-27B do?
    Qwen3.8-27B is best for Teams deploying a 27B vision-language model locally or through QwenCloud. Its main capabilities include Open Source, Code, Vision, Agentic, Local Deployment, Multilingual.
    How does Qwen3.8-27B compare to other models?
    Qwen3.8-27B excels at Apache 2.0 weights and is recommended for Self-hosted coding agents, Local multimodal AI, Apache-2 licensed deployments, Edge agentic workloads. See related models below.
    How do I use Qwen3.8-27B in yno.ai?
    Sign up for yno.ai, select Qwen3.8-27B from the model picker, and start chatting.