Alibaba
Qwen3.8-27B
Alibaba's August 14, 2026 dense 27B vision-language model, with Apache 2.0 weights and optional thinking. It accepts text, images and video. Published weights have a native 262,144-token context, extensible to 1M with YaRN; QwenCloud provides a hosted 1M-context endpoint at $0.50/$3 per MTok (input/output). Hosted limits and prices are separate from local deployment requirements.
262K tokens (up to 1M)
Multimodal
Pro
New
- Provider
- Alibaba
- Context
- 262KCatalog · As of Aug 18, 2026Within 30d
- yno subscription tier
- Pro
- Released
- Aug 14, 2026
- Speed
- Fast
- Reasoning
- Advanced
- Modality
- Multimodal
Benchmarks
Quality
- AA Intelligence Index
- AA Coding Index
- GPQA Diamond
- 90.5%Artificial Analysis · Date unknownAge unknown
Speed and latency
- Output speed
- 44.45tokens/sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
- Time to first token
- 1.27sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
- Time to first answer token
- 46.26sConfiguration: Prompt length: 1,000 · Parallel queries: 1Artificial Analysis · Date unknownAge unknown
API pricing
- API input cost
- API output cost
Benchmark variant: Qwen3.8 27B (xhigh)
AA data retrieved Sep 11, 2026 · Artificial Analysis
Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology
Capabilities
- Open Source
- Code
- Vision
- Agentic
- Local Deployment
- Multilingual
Use cases
- Self-hosted coding agents
- Local multimodal AI
- Apache-2 licensed deployments
- Edge agentic workloads
Strengths
- Apache 2.0 weights
- Dense 27B model
- Vision and optional thinking
- 262K native context; extensible to 1M
Best for
Teams deploying a 27B vision-language model locally or through QwenCloud
Use Qwen3.8-27B in yno.ai
No credit card required
Related models
Frequently asked
- What is Qwen3.8-27B?
- Alibaba's August 14, 2026 dense 27B vision-language model, with Apache 2.0 weights and optional thinking. It accepts text, images and video. Published weights have a native 262,144-token context, extensible to 1M with YaRN; QwenCloud provides a hosted 1M-context endpoint at $0.50/$3 per MTok (input/output). Hosted limits and prices are separate from local deployment requirements.
- How much does Qwen3.8-27B cost in yno.ai?
- Qwen3.8-27B is available on the Pro tier of yno.ai.
- What can Qwen3.8-27B do?
- Qwen3.8-27B is best for Teams deploying a 27B vision-language model locally or through QwenCloud. Its main capabilities include Open Source, Code, Vision, Agentic, Local Deployment, Multilingual.
- How does Qwen3.8-27B compare to other models?
- Qwen3.8-27B excels at Apache 2.0 weights and is recommended for Self-hosted coding agents, Local multimodal AI, Apache-2 licensed deployments, Edge agentic workloads. See related models below.
- How do I use Qwen3.8-27B in yno.ai?
- Sign up for yno.ai, select Qwen3.8-27B from the model picker, and start chatting.