Your Go-Anywhere Daily AI

Create AI agents, automate business workflows, and build custom plugins.

Works with 73+ AI models including GPT-5.6 Sol, Claude Opus 5, and Gemini 3.7 Flash to help you work smarter, not harder.

New here? Create a free account — no credit card required.

Trusted by businesses worldwide

Build AI Agents Without Code

Create powerful AI agents and automate your workflows with our intuitive platform. No technical expertise required.

Dynamic Context API

Inject real-time business data and external information directly into your AI conversations.

Prompt Library

Save and organize your best prompts, share with your team, and reuse proven workflows.

Visual Workflow Builder

Create complex AI workflows with our drag-and-drop interface - no coding required.

Multi-Model Support

Switch between 73+ AI models from OpenAI, Anthropic, Google, and other leading providers seamlessly in one platform.

Real-Time Collaboration

Work together on AI projects, share conversations, and build team knowledge bases.

Secret Vault Infisical Integration

Enterprise-grade secret management with automatic API key injection for your AI models.

Enterprise Security

Your data stays private with SOC2 compliance, end-to-end encryption, and regular audits.

Harness the future of intelligence today

Step into tomorrow with revolutionary AI models that redefine what's possible. From OpenAI's GPT-5.3 Codex and reasoning powerhouse GPT-5.6 Sol, to Anthropic's Claude Sonnet 5 with 1M context, and Google's breakthrough Gemini 3.7 Flash. Our platform gives you instant access to cutting-edge capabilities from the world's leading AI labs.

GLM-5.3

Zhipu AI

NEW

Best for

Coding teams building long-running agents with hosted or self-hosted deployment

Read full description GLM-5.3

Zhipu's August 14, 2026 coding and agentic model, built on the GLM-5.2 base with further post-training. Z.ai reports Terminal-Bench 3.0 of 28.3 and DeepSWE v1.1 of 66.9. Published weights use the custom GLM-5.3 license; hosted access is available through Z.ai and the GLM Coding Plan. Text-only input, a documented 1M-token context window and up to 128K output, with reasoning always enabled. Useful for coding teams that need long-running agents or self-hosted deployment under the model license.

Billed monthly
AA Intelligence Index
44.9points
Artificial Analysis · Age unknown
Output speed
53.45tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$1.40/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$4.40/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Within 30d

Benchmark variant: GLM-5.3 (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
44.9points
Artificial Analysis · Date unknownAge unknown
Output speed
53.45tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$1.40/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$4.40/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 24, 2026Within 30d

Benchmark variant: GLM-5.3 (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Gemini 3.7 Flash

Google

NEW

Best for

Teams that want a Google coding-and-agents workhorse at intro Flash pricing

Read full description Gemini 3.7 Flash

Google's Flash model for coding and agent workflows, released August 13, 2026. At launch, Google reported DeepSWE v1.1 65.3%, FrontierCode 1.1 Main 43.6%, WebDev Arena 1588 Elo and AutomationBench 30.4%. It accepts text, images, audio, video and PDFs, with a 1,048,576-token input limit and a 65,536-token text-output limit. Thinking levels are low, medium and high. Standard Gemini API rates are $0.75/$3.75 per MTok (input/output) through December 31, 2026, then $1.50/$7.50 from January 1, 2027.

Billed monthly
AA Intelligence Index
39.4points
Artificial Analysis · Age unknown
Output speed
333.34tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.75/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$3.75/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Within 30d

Benchmark variant: Gemini 3.7 Flash (high)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
39.4points
Artificial Analysis · Date unknownAge unknown
Output speed
333.34tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.75/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$3.75/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 14, 2026Within 30d

Benchmark variant: Gemini 3.7 Flash (high)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Grok 4.6

xAI

NEW

Best for

Teams that want frontier long-horizon agent capability at xAI's aggressive pricing

Read full description Grok 4.6

xAI's August 12, 2026 model for coding, agentic tasks and knowledge work. It builds on Grok 4.5 with supplemental training, regenerated supervised fine-tuning trajectories and agentic reinforcement learning. At launch, xAI reported DeepSWE v1.1 65.9% and APEX-Agents 57.5% at high reasoning effort. It accepts text and images, produces text and has a 500K-token context window, with low/medium/high/xhigh reasoning effort. Standard xAI API input/output rates are $2/$6 per MTok (input/output) below 200K prompt tokens and $4/$12 at or above 200K; cached input is $0.50/$1 respectively. Microsoft Foundry availability was announced August 26, 2026.

Max plan only
AA Intelligence Index
44.4points
Artificial Analysis · Age unknown
Output speed
67.86tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$2.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$6.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
500K
Catalog · Within 30d

Benchmark variant: Grok 4.6 (high)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
44.4points
Artificial Analysis · Date unknownAge unknown
Output speed
67.86tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$2.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$6.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
500K
Catalog · As of Aug 13, 2026Within 30d

Benchmark variant: Grok 4.6 (high)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Muse Spark 1.2

Meta

NEW

Best for

Teams using Muse Code or Meta Model API for coding and multimodal workflows

Read full description Muse Spark 1.2

Meta's August 5, 2026 coding-focused update to Muse Spark 1.1, released in Muse Code and Meta Model API. Co-trained with Muse Code for code generation, debugging, codebase understanding and long-horizon developer workflows. Supports image and video reasoning and audiovisual workflows. Meta launched the later Muse Spark 1.3 generation on September 2, 2026.

Billed monthly
AA Intelligence Index
39.8points
Artificial Analysis · Age unknown
Output speed
250.08tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$1.25/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$4.25/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Estimated · Within 30d

Benchmark variant: Muse Spark 1.2 (xhigh)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
39.8points
Artificial Analysis · Date unknownAge unknown
Output speed
250.08tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$1.25/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$4.25/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Estimated · As of Aug 13, 2026Within 30d

Benchmark variant: Muse Spark 1.2 (xhigh)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Qwen3.8-Max

Alibaba

NEW

Best for

Teams building hosted coding agents and document or video analysis workflows

Read full description Qwen3.8-Max

Alibaba's hosted Qwen3.8-Max accepts text, images and video within a 1M-token context window. QwenCloud lists $2/$6 per MTok (input/output). Alibaba Cloud added the dated qwen3.8-max-0902 snapshot on September 2, 2026, also named qwen3.8-max-2026-09-02. The related Qwen3.8-2.4T-A95B weights are a text-only post-trained model; their license and local context limits should not be conflated with the hosted service.

Max plan only
AA Intelligence Index
40.3points
Artificial Analysis · Age unknown
Output speed
40.84tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$2.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$6.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Within 30d

Benchmark variant: Qwen3.8 Max

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
40.3points
Artificial Analysis · Date unknownAge unknown
Output speed
40.84tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$2.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$6.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 13, 2026Within 30d

Benchmark variant: Qwen3.8 Max

AA data retrieved Sep 11, 2026 · Artificial Analysis

Claude Opus 5

Anthropic

NEW

Best for

Teams building coding agents and complex enterprise workflows with adjustable reasoning effort

Read full description Claude Opus 5

Anthropic's July 24, 2026 model for complex coding and enterprise work, with adjustable effort and thinking enabled by default. Provides a 1M-token context window and up to 128K output at $5 input and $25 output per million tokens, unchanged from Opus 4.8. Optional Fast Mode is a research preview on the Claude API with access restrictions: up to 2.5x output-token throughput at $10/$50 per million tokens.

Max plan only
AA Intelligence Index
50.7points
Artificial Analysis · Age unknown
Output speed
58.1tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$5.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$25.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Claude Opus 5 (Adaptive Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
50.7points
Artificial Analysis · Date unknownAge unknown
Output speed
58.1tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$5.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$25.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Claude Opus 5 (Adaptive Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Gemini 3.5 Flash-Lite

Google

NEW

Best for

Teams running fleets of subagents or high-volume pipelines on Gemini

Read full description Gemini 3.5 Flash-Lite

Google's July 21, 2026 low-latency model for subagent tasks and high-volume document processing. Accepts text, images, video, audio and PDFs, with 1M input tokens and up to 65K text output. Standard API rates are $0.30 input and $2.50 output per million tokens; batch rates are $0.15/$1.25. Standard cached input costs $0.03 per million tokens, with cache storage billed separately.

Billed monthly
AA Intelligence Index
22.7points
Artificial Analysis · Age unknown
Output speed
349.95tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.30/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$2.50/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Gemini 3.5 Flash-Lite

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
22.7points
Artificial Analysis · Date unknownAge unknown
Output speed
349.95tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.30/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$2.50/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Gemini 3.5 Flash-Lite

AA data retrieved Sep 11, 2026 · Artificial Analysis

Kimi K3

Moonshot AI

NEW

Best for

Teams building coding and knowledge-work agents with native vision and long context

Read full description Kimi K3

Moonshot's Kimi K3 is a native vision model with 2.8T total and 104B active parameters and a 1,048,576-token context. Published weights use the Kimi K3 License. The first-party API supports text, image and video inputs; thinking is always enabled with low, high and max effort settings. It is available for coding and knowledge-work agents through Moonshot's API and self-hosted weights.

Max plan only
AA Intelligence Index
43.8points
Artificial Analysis · Age unknown
Output speed
37.03tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$3.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$15.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Kimi K3 (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
43.8points
Artificial Analysis · Date unknownAge unknown
Output speed
37.03tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$3.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$15.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Kimi K3 (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

GPT-5.6 Sol

OpenAI

Best for

Teams using the GPT-5.6 family for complex professional and coding workflows

Read full description GPT-5.6 Sol

OpenAI's July 9, 2026 GPT-5.6 model for complex professional work. Supports text and image input, text output and adjustable reasoning effort, with a 1.05M-token context window and up to 128K output. Promotional API rates are $4 input and $20 output per million tokens through at least November 21, 2026; cached input is $0.40. Above 272K input tokens, the full request costs 2x input and 1.5x output rates. Cache writes cost 1.25x the uncached input rate.

Max plan only
AA Intelligence Index
47.1points
Artificial Analysis · Age unknown
Output speed
72.34tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$4.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$20.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: GPT-5.6 Sol (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
47.1points
Artificial Analysis · Date unknownAge unknown
Output speed
72.34tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$4.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$20.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: GPT-5.6 Sol (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

GPT-5.6 Terra

OpenAI

Best for

Teams balancing model capability and API cost in GPT-5.6 production workloads

Read full description GPT-5.6 Terra

OpenAI's July 9, 2026 GPT-5.6 model balancing intelligence and cost. Supports text and image input, text output and adjustable reasoning effort, with a 1.05M-token context window and up to 128K output. API prices fell 20% on July 30 to $2 input and $12 output per million tokens; cached input is $0.20. Above 272K input tokens, the full request costs 2x input and 1.5x output rates. Cache writes cost 1.25x the uncached input rate.

Billed monthly
AA Intelligence Index
42.3points
Artificial Analysis · Age unknown
Output speed
103.11tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$2.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$12.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: GPT-5.6 Terra (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
42.3points
Artificial Analysis · Date unknownAge unknown
Output speed
103.11tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$2.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$12.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: GPT-5.6 Terra (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

GPT-5.6 Luna

OpenAI

Best for

Teams running cost-sensitive, high-volume workloads with the GPT-5.6 family

Read full description GPT-5.6 Luna

OpenAI's July 9, 2026 GPT-5.6 model for cost-sensitive, high-volume workloads. Supports text and image input, text output and adjustable reasoning effort, with a 1.05M-token context window and up to 128K output. API prices fell 80% on July 30 to $0.20 input and $1.20 output per million tokens; cached input is $0.02. Above 272K input tokens, the full request costs 2x input and 1.5x output rates. Cache writes cost 1.25x the uncached input rate.

Billed monthly
AA Intelligence Index
37.5points
Artificial Analysis · Age unknown
Output speed
125.89tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.20/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$1.20/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: GPT-5.6 Luna (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
37.5points
Artificial Analysis · Date unknownAge unknown
Output speed
125.89tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.20/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$1.20/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: GPT-5.6 Luna (max)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Qwen3.7-Plus

Alibaba

Best for

Teams building multimodal agents and long-context text or visual workflows

Read full description Qwen3.7-Plus

Alibaba's June 2026 Qwen3.7-Plus is a hosted model accepting text, images and video with a 1M-token context. Its preserve_thinking option retains reasoning across tool calls. Current provider rates depend on input length and promotional discounts, so a single price ratio against Max does not describe all requests.

Billed monthly
AA Intelligence Index
48points
Estimated · Stale · 31d old
Output speed
80tokens/s
Estimated · Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
1M
Estimated · Stale · 31d old

AA measurements unavailable

Measurement details & sources
AA Intelligence Index
48points
Estimated · As of Aug 11, 2026Stale · 31d old
Output speed
80tokens/s
Estimated · As of Aug 11, 2026Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
1M
Estimated · As of Aug 11, 2026Stale · 31d old

AA measurements unavailable

Claude Mythos 5

Anthropic

Best for

Project Glasswing participants who need Fable 5-class capability under restricted access

Read full description Claude Mythos 5

Anthropic's June 9, 2026 restricted-access counterpart to Fable 5, available by invitation through Project Glasswing. It shares Fable 5's capabilities but does not include the same safety classifiers, so behavior is not identical. It supports text and image input, a 1M-token context window, up to 128K output and always-on thinking. Mythos 5.1 was released September 1; Mythos 5 remains listed as active.

Max plan only
AA Intelligence Index
62points
Estimated · Stale · 31d old
Output speed
63tokens/s
Estimated · Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
1M
Estimated · Stale · 31d old

AA measurements unavailable

Measurement details & sources
AA Intelligence Index
62points
Estimated · As of Aug 11, 2026Stale · 31d old
Output speed
63tokens/s
Estimated · As of Aug 11, 2026Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
1M
Estimated · As of Aug 11, 2026Stale · 31d old

AA measurements unavailable

Claude Fable 5

Anthropic

Best for

Teams maintaining Fable 5 workflows or evaluating migration to Fable 5.1

Read full description Claude Fable 5

Anthropic's June 9, 2026 Fable model remains active as a prior-generation option after Fable 5.1 launched September 1. It accepts text and images, provides a 1M-token context window and up to 128K output, and uses always-on thinking. Standard Claude API prices are $10/$50 per MTok (input/output), with $1 cached input. Fable 5 includes safety classifiers that can refuse requests; its restricted Mythos 5 counterpart has the same capabilities without those classifiers.

Max plan only
AA Intelligence Index
49.7points
Artificial Analysis · Age unknown
Output speed
69.3tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$10.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$50.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
49.7points
Artificial Analysis · Date unknownAge unknown
Output speed
69.3tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$10.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$50.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)

AA data retrieved Sep 11, 2026 · Artificial Analysis

MAI-Thinking-1

Microsoft

Best for

Microsoft-stack teams that want a first-party frontier reasoning model

Read full description MAI-Thinking-1

Microsoft's first fully in-house reasoning model, unveiled at Build 2026 (June 2) — a sparse MoE with ~35B active of ~1T total params, trained without OpenAI distillation using data Microsoft describes as appropriately licensed. AIME 2025 97.0%, AIME 2026 94.5%, SWE-Bench Pro competitive with Claude Opus 4.6, and preferred over Sonnet 4.6 in blind human evals across 1,276 tasks. 256K context; public preview on Microsoft Foundry since August 12, 2026 with third-party availability via Fireworks AI, Baseten, and OpenRouter.

Max plan only
AA Intelligence Index
46points
Estimated · Stale · 31d old
Output speed
95tokens/s
Estimated · Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
256K
Estimated · Stale · 31d old

AA measurements unavailable

Measurement details & sources
AA Intelligence Index
46points
Estimated · As of Aug 11, 2026Stale · 31d old
Output speed
95tokens/s
Estimated · As of Aug 11, 2026Stale · 31d old
API input cost
Not available
API output cost
Not available
Context window
256K
Estimated · As of Aug 11, 2026Stale · 31d old

AA measurements unavailable

Claude Sonnet 5

Anthropic

Best for

Teams building coding and agent workflows with Sonnet 5 and adjustable reasoning effort

Read full description Claude Sonnet 5

Anthropic's June 30, 2026 Sonnet model for coding and agent workflows, with text and image input, a 1M-token context window and up to 128K output. Thinking is enabled by default with adjustable effort. Standard Claude API pricing is $2/$10 per MTok (input/output); the previously announced September 1 price increase was canceled. Anthropic's launch comparisons with Opus 4.8 depend on the task and effort setting.

Billed monthly
AA Intelligence Index
38.4points
Artificial Analysis · Age unknown
Output speed
83.7tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$2.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$10.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Claude Sonnet 5 (Adaptive Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
38.4points
Artificial Analysis · Date unknownAge unknown
Output speed
83.7tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$2.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$10.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Claude Sonnet 5 (Adaptive Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Mistral Medium 3.5

Mistral AI

Best for

Teams building coding and image-analysis agents with hosted or self-hosted deployment

Read full description Mistral Medium 3.5

Mistral's April 28, 2026 model (v26.04), combining instruction-following, reasoning, and coding in 128B dense parameters. It supports vision, function calling, configurable reasoning, and a 256K-token context. Published weights use a Modified MIT license. Mistral lists standard provider API pricing of $1.50/$7.50 per MTok for input/output. Its May 22 product announcement made the model the default in Vibe CLI and Le Chat.

Billed monthly
AA Intelligence Index
14.9points
Artificial Analysis · Age unknown
Output speed
134.09tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$1.50/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$7.50/ 1M tokens
Artificial Analysis · Age unknown
Context window
256K
Catalog · Stale · 31d old

Benchmark variant: Mistral Medium 3.5

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
14.9points
Artificial Analysis · Date unknownAge unknown
Output speed
134.09tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$1.50/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$7.50/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
256K
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Mistral Medium 3.5

AA data retrieved Sep 11, 2026 · Artificial Analysis

DeepSeek V4 Pro

DeepSeek

NEW

Best for

Teams using V4 Pro weights or preparing for the announced first-party API routing change

Read full description DeepSeek V4 Pro

DeepSeek-V4-Pro-0813 is the August 13, 2026 text-only V4 Pro release, with 1.6T total and 49B active parameters, a 1M-token context and MIT weights. As of September 10, the deepseek-v4-pro API still identifies this build. DeepSeek has announced that from September 14, 2026 at 12:00 Beijing time (04:00 UTC), this slug will temporarily route to DeepSeek-V4.1-Flash at Flash pricing until a future V4.1 Pro release.

Max plan only
AA Intelligence Index
36.3points
Artificial Analysis · Age unknown
Output speed
69.68tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$1.32/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$3.96/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Within 30d

Benchmark variant: DeepSeek V4 Pro 0813 (Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
36.3points
Artificial Analysis · Date unknownAge unknown
Output speed
69.68tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$1.32/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$3.96/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 13, 2026Within 30d

Benchmark variant: DeepSeek V4 Pro 0813 (Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

DeepSeek V4 Flash

DeepSeek

NEW

Best for

Historical comparison and self-hosted use of V4-Flash-0731 weights

Read full description DeepSeek V4 Flash

DeepSeek-V4-Flash-0731 is the July 31, 2026 text-only V4 Flash release, with 284B total and 13B active parameters, a 1M-token context and MIT weights. DeepSeek retired its first-party V4 Flash service on September 10, 2026. The legacy deepseek-v4-flash slug now temporarily routes to the newer, vision-capable DeepSeek-V4.1-Flash. These historical weights retain their own architecture and modalities.

Billed monthly
AA Intelligence Index
34.5points
Artificial Analysis · Age unknown
Output speed
239.38tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.44/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$1.32/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: DeepSeek V4 Flash 0731 (Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
34.5points
Artificial Analysis · Date unknownAge unknown
Output speed
239.38tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.44/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$1.32/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: DeepSeek V4 Flash 0731 (Reasoning, Max Effort)

AA data retrieved Sep 11, 2026 · Artificial Analysis

GPT-5.5 Pro

OpenAI

Best for

Research and analysis teams using additional GPT-5.5 reasoning compute for complex questions

Read full description GPT-5.5 Pro

A GPT-5.5 reasoning variant that uses additional compute for complex questions. Supports medium, high and xhigh reasoning effort, text and image input, and text output. Available through the Responses API, including Batch requests, with a 1M-token context window and up to 128K output tokens. Some requests can take several minutes; background mode supports longer tasks.

Max plan only
AA Intelligence Index
57points
Estimated · Stale · 31d old
Output speed
78tokens/s
Estimated · Stale · 31d old
API input cost
$0.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$0.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Estimated · Stale · 31d old

Benchmark variant: GPT-5.5 Pro (xhigh)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
57points
Estimated · As of Aug 11, 2026Stale · 31d old
Output speed
78tokens/s
Estimated · As of Aug 11, 2026Stale · 31d old
API input cost
$0.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$0.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Estimated · As of Aug 11, 2026Stale · 31d old

Benchmark variant: GPT-5.5 Pro (xhigh)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Gemma 4

Google (Gemma)

Best for

Organizations building self-hosted assistants or fine-tuning open-weight models for their own workloads

Read full description Gemma 4

Google DeepMind's open-weight Gemma 4 family, announced April 2, 2026 under Apache 2.0. The family now includes E2B, E4B, 12B, 26B A4B and 31B; 12B Unified was added June 3. E2B/E4B support 128K context, while 12B/26B A4B/31B support 256K. All accept text and images and produce text; audio input is supported by E2B, E4B and 12B. The family supports configurable thinking and function calling for local and server deployments.

Billed monthly
AA Intelligence Index
15.4points
Artificial Analysis · Age unknown
Output speed
34.57tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$0.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
256K
Catalog · Stale · 31d old

Benchmark variant: Gemma 4 31B (Reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
15.4points
Artificial Analysis · Date unknownAge unknown
Output speed
34.57tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$0.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
256K
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Gemma 4 31B (Reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

MiniMax M3

MiniMax

Best for

Teams building multimodal coding agents and computer-use workflows

Read full description MiniMax M3

MiniMax introduced M3 on June 1, 2026 with MiniMax Sparse Attention, native text, image and video input, and a 1M-token context. Published weights use the MiniMax community license. The standard provider API costs $0.30/$1.20 per MTok (input/output) for requests with at most 512K input tokens; both rates double above that threshold. Optional priority service costs 1.5 times the corresponding standard rate. MiniMax labels these standard rates a permanent 50% discount.

Billed monthly
AA Intelligence Index
29.6points
Artificial Analysis · Age unknown
Output speed
103.72tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.30/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$1.20/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: MiniMax-M3

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
29.6points
Artificial Analysis · Date unknownAge unknown
Output speed
103.72tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.30/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$1.20/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: MiniMax-M3

AA data retrieved Sep 11, 2026 · Artificial Analysis

Mistral Small 4

Mistral AI

Best for

Teams that want one model for reasoning, vision, AND agentic coding without operating three

Read full description Mistral Small 4

Mistral's March 16, 2026 unified model: merges Magistral (reasoning), Pixtral (vision), and Devstral (agentic coding) into a single 119B-total MoE with 6.5B active per token. 256K context, Apache 2.0, configurable reasoning effort.

Billed monthly
AA Intelligence Index
11.5points
Artificial Analysis · Age unknown
Output speed
169.4tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$0.15/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$0.60/ 1M tokens
Artificial Analysis · Age unknown
Context window
256K
Catalog · Stale · 31d old

Benchmark variant: Mistral Small 4 (Reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
11.5points
Artificial Analysis · Date unknownAge unknown
Output speed
169.4tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$0.15/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$0.60/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
256K
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Mistral Small 4 (Reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Gemini 3.1 Pro

Google

Best for

Research and enterprise workflows using reasoning, code analysis and multimodal input

Read full description Gemini 3.1 Pro

Google's reasoning model, released February 19, 2026 and available through the gemini-3.1-pro-preview API endpoint. It accepts text, images, video, audio and PDFs, with a 1,048,576-token input limit and a 65,536-token text-output limit. Google reported 77.1% on ARC-AGI-2 at launch, more than double Gemini 3 Pro's score on that benchmark. Standard Gemini API pricing is $2/$12 per MTok (input/output) for prompts up to 200K tokens, and $4/$18 for longer prompts.

Max plan only
AA Intelligence Index
30.4points
Artificial Analysis · Age unknown
Output speed
125.03tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$2.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$12.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
1M
Catalog · Stale · 31d old

Benchmark variant: Gemini 3.1 Pro Preview

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
30.4points
Artificial Analysis · Date unknownAge unknown
Output speed
125.03tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$2.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$12.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
1M
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Gemini 3.1 Pro Preview

AA data retrieved Sep 11, 2026 · Artificial Analysis

Claude Haiku 4.5

Anthropic

Best for

Teams running fast, cost-sensitive Claude workloads such as support, classification and summarization

Read full description Claude Haiku 4.5

Anthropic's October 15, 2025 Haiku model for fast, cost-sensitive work. It accepts text and images and produces text, with a 200K-token context window, up to 64K output and optional extended thinking. Standard Claude API rates are $1/$5 per MTok (input/output). Anthropic's launch evaluation described coding performance similar to Sonnet 4 at more than twice its speed; that comparison is specific to the vendor's launch testing.

Billed monthly
AA Intelligence Index
15.4points
Artificial Analysis · Age unknown
Output speed
92.11tokens/s
Artificial Analysis · Age unknownConfigured
API input cost
$1.00/ 1M tokens
Artificial Analysis · Age unknown
API output cost
$5.00/ 1M tokens
Artificial Analysis · Age unknown
Context window
200K
Catalog · Stale · 31d old

Benchmark variant: Claude 4.5 Haiku (Non-reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Measurement details & sources
AA Intelligence Index
15.4points
Artificial Analysis · Date unknownAge unknown
Output speed
92.11tokens/s
Configuration: Prompt length: 1,000 · Parallel queries: 1
Artificial Analysis · Date unknownAge unknown
API input cost
$1.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
API output cost
$5.00/ 1M tokens
Artificial Analysis · Date unknownAge unknown
Context window
200K
Catalog · As of Aug 11, 2026Stale · 31d old

Benchmark variant: Claude 4.5 Haiku (Non-reasoning)

AA data retrieved Sep 11, 2026 · Artificial Analysis

Extend with powerful plugins

Connect your AI agents to external tools and services. Build custom plugins or use pre-built ones.

Web Search

Search for information from the internet in real-time using Google.

Simple Calculator

Calculate a math expression. For example, "2 + 2" or "2 * 2".

Web Page Reader

Read the content of a web page via its URL.

Google Calendar

Return the next 10 events in the current user Google's calendar starting from a specific date.

Firecrawl Web Page Reader

Retrieves the content of a web page by scraping it using the Firecrawl API.

Send email with Zapier

Send an email to a specific email address with title and text content.

Render Chart

PREMIUM

Generate a Chart.js chart.

DALL-E 3

PREMIUM

Generate images using DALL-E 3 based on image descriptions. Adhere to content policy.

Interactive Canvas

PREMIUM

Render an interactive canvas with HTML source to the user interface. The HTML source should be complete.

Market News

PREMIUM

Fetches market news articles from Alpha Vantage. This plugin requires an API key.

SQLite Database

PREMIUM

Create, query, and manage SQLite databases with advanced data operations and analytics.

REST API Client

Make HTTP requests to any REST API with custom headers, authentication, and data processing.

Slack Integration

PREMIUM

Send messages, create channels, and manage Slack workspaces through MCP integration.

Excel Processor

PREMIUM

Read, write, and manipulate Excel files with advanced data analysis capabilities.

GitHub Manager

PREMIUM

Interact with GitHub repositories, create issues, manage pull requests, and analyze code.

AWS S3 Storage

PREMIUM

Upload, download, and manage files in Amazon S3 buckets with secure access controls.

Password Manager

PREMIUM

Generate secure passwords, store credentials safely, and manage authentication tokens.

Discord Bot

PREMIUM

Create and manage Discord bots, send messages, and interact with Discord servers.

Zapier Automation

Trigger Zapier workflows and automate tasks across thousands of applications.

Git Operations

PREMIUM

Perform Git operations, manage repositories, and track version control changes.

YouTube API

PREMIUM

Search YouTube videos, extract metadata, and manage playlists through YouTube API.

System Monitor

PREMIUM

Monitor system performance, track resource usage, and get system information.

Screenshot Tool

PREMIUM

Capture screenshots of web pages, applications, and desktop areas automatically.

PDF Processor

PREMIUM

Extract text from PDFs, merge documents, and convert between formats.

Google Maps

PREMIUM

Get location data, calculate distances, and access mapping services through Google Maps API.

E-commerce Analytics

PREMIUM

Track sales data, analyze customer behavior, and generate e-commerce reports.

CRM Integration

PREMIUM

Manage customer relationships, track leads, and sync data with popular CRM platforms.

Image Processing

PREMIUM

Resize, crop, filter, and optimize images with advanced computer vision capabilities.

Audio Transcription

PREMIUM

Convert audio files to text, analyze speech patterns, and generate transcripts.

How teams use yno.ai

Three usage patterns we see across engineering, research, and operations teams.

VP of Engineering

B2B SaaS scale-up

Engineering teams use yno.ai to run code review across multiple frontier models in one pipeline — Claude for refactor suggestions, GPT for test generation, Gemini for documentation. Setup that used to take a sprint runs in an afternoon.

Data Science Lead

Research-heavy analytics org

Research teams compose multi-step agent runs over long documents — switching between models per stage based on cost, latency, or capability. yno.ai's unified interface removes per-provider boilerplate from every notebook.

Product Operations

Mid-market ops team

Operations teams connect yno.ai's MCP plugin layer to internal tools — Notion, Linear, custom databases — so AI agents can act on real systems instead of generating disconnected text. New integrations land in hours, not weeks.

Simple, transparent pricing

Start free and upgrade as you grow. No hidden fees, cancel anytime.

Free

Free

Perfect for trying out AI agents and exploring the platform.

  • 50 messages per day
  • Basic chat interface
  • Community support
  • Export conversations
  • 1 custom agent
MOST POPULAR

Pro

$80
$16/month

Unlock all AI models and advanced features for your business.

  • Access to 39+ Pro AI models including GPT-5.6 Terra, Claude Sonnet 5, Gemini 3.7 Flash, GLM-5.2, Gemma 4
  • Unlimited messages
  • Priority response speed
  • Unlimited custom agents
  • Team collaboration (5 users)
  • API access
  • File uploads & analysis
  • Custom plugins
  • Chat history & organization
  • Priority support

* Introductory pricing valid for a limited time

Max

$85585% OFF
$128/month

For teams that need maximum power and enterprise features.

  • Plus 34+ Max-tier flagships: GPT-5.6 Sol, Claude Opus 5, DeepSeek V4 Pro, Qwen3.8-Max, Kimi K3
  • Everything in Pro
  • 500M+ tokens per month
  • 20x higher rate limits
  • Dedicated infrastructure
  • Advanced analytics dashboard
  • Unlimited team members
  • SSO & SAML
  • Custom model fine-tuning
  • White-label options
  • 24/7 phone support
  • SLA guarantee
  • On-premise deployment option

* Introductory pricing valid for a limited time

All plans include access to our web and mobile apps. Need a custom enterprise solution?Contact us