Skip to main content

    Alibaba · Alibaba

    Qwen3.8-2.4T-A95B vs Qwen3.8-Max

    Compare Qwen3.8-2.4T-A95B (Alibaba) and Qwen3.8-Max (Alibaba) on benchmarks, capabilities, and pricing.

    At a glance

    Qwen3.8-2.4T-A95BQwen3.8-Max
    ProviderAlibabaAlibaba
    Context262K tokens (up to 1M)1M tokens
    ModalityText onlyMultimodal
    PricingMaxMax
    SpeedSlowMedium
    ReasoningExpertExpert

    Benchmarks

    Qwen3.8-2.4T-A95B

    AA Intelligence Index
    40points
    Artificial Analysis · Date unknownAge unknown
    Output speed
    40.21tokens/s
    Configuration: Prompt length: 1,000 · Parallel queries: 1
    Artificial Analysis · Date unknownAge unknown
    API input cost
    $2.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    API output cost
    $6.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    Context window
    262K
    Catalog · As of Aug 14, 2026Within 30d

    Benchmark variant: Qwen3.8 2.4T A95B

    AA data retrieved Sep 11, 2026 · Artificial Analysis

    Qwen3.8-Max

    AA Intelligence Index
    40.3points
    Artificial Analysis · Date unknownAge unknown
    Output speed
    40.84tokens/s
    Configuration: Prompt length: 1,000 · Parallel queries: 1
    Artificial Analysis · Date unknownAge unknown
    API input cost
    $2.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    API output cost
    $6.00/ 1M tokens
    Artificial Analysis · Date unknownAge unknown
    Context window
    1M
    Catalog · As of Aug 13, 2026Within 30d

    Benchmark variant: Qwen3.8 Max

    AA data retrieved Sep 11, 2026 · Artificial Analysis

    Try both on yno.ai

    Switch between Qwen3.8-2.4T-A95B and Qwen3.8-Max per task with one account.

    Frequently asked

    Which is faster, Qwen3.8-2.4T-A95B or Qwen3.8-Max?
    Qwen3.8-2.4T-A95B is rated slow and Qwen3.8-Max is rated medium. yno.ai surfaces output-speed scores from Artificial Analysis on each model's detail page so you can compare exact tokens-per-second figures for your workload.
    Which is cheaper, Qwen3.8-2.4T-A95B or Qwen3.8-Max?
    Qwen3.8-2.4T-A95B is on the max tier in yno.ai; Qwen3.8-Max is on the max tier. See the pricing page for the latest per-tier limits.
    Which is better for Self-hosted frontier agents?
    Both models support Self-hosted frontier agents. Qwen3.8-2.4T-A95B brings Published post-trained weights; Qwen3.8-Max brings 2.4T MoE / 95B active. Run a side-by-side eval on your prompts in yno.ai to see which fits your workload.
    Can I use both Qwen3.8-2.4T-A95B and Qwen3.8-Max in yno.ai?
    Yes. Both are available on yno.ai under your single account; you can route different stages of an agent to different models or A/B test them on the same prompt without per-provider boilerplate.

    Related comparisons