Skip to main content

    Google

    Gemini 3.8 Flash

    Google's September 2, 2026 Flash model for coding and agent workflows, described by Google as its best reasoning and coding model yet at the same speed and cost as 3.7 Flash. It accepts text, images, audio, video and PDFs, with a 1,048,576-token input limit and a 65,536-token text-output limit, and thinking levels of low, medium and high. Standard Gemini API rates are $0.75/$3.75 per MTok (input/output) through December 31, 2026, then $1.50/$7.50 from January 1, 2027, with cached input at $0.075. Google reports DeepSWE v1.1 73.7% and roughly 40% more work completed per task than 3.7 Flash.

    1M tokens
    Multimodal
    Pro
    New
    Provider
    Google
    Context
    1M
    Catalog · As of Sep 25, 2026Under 1d old
    yno subscription tier
    Pro
    Released
    Sep 2, 2026
    Speed
    Fast
    Reasoning
    Expert
    Modality
    Multimodal

    Benchmarks

    Quality

    AA Intelligence Index
    57points
    Artificial Analysis · As of Sep 25, 2026Under 1d old
    GPQA Diamond
    95%
    Artificial Analysis · As of Sep 25, 2026Under 1d old

    Speed and latency

    Output speed
    515tokens/s
    Artificial Analysis · As of Sep 25, 2026Under 1d old

    API pricing

    Not available

    Indexes use points; evaluations use accuracy percentages. API measurements do not measure yno application performance. Methodology

    Capabilities

    • Code
    • Agentic
    • Vision
    • Video
    • Speed
    • Multimodal

    Use cases

    • Agentic coding
    • Real-time apps
    • Document processing
    • Multimodal pipelines

    Strengths

    • Google-reported DeepSWE v1.1 73.7%
    • Same speed and cost as 3.7 Flash
    • Introductory API pricing through December 2026
    • Three thinking levels

    Best for

    Teams that want a Google coding-and-agents workhorse at intro Flash pricing

    Use Gemini 3.8 Flash in yno.ai

    No credit card required

    Related models

    Frequently asked

    What is Gemini 3.8 Flash?
    Google's September 2, 2026 Flash model for coding and agent workflows, described by Google as its best reasoning and coding model yet at the same speed and cost as 3.7 Flash. It accepts text, images, audio, video and PDFs, with a 1,048,576-token input limit and a 65,536-token text-output limit, and thinking levels of low, medium and high. Standard Gemini API rates are $0.75/$3.75 per MTok (input/output) through December 31, 2026, then $1.50/$7.50 from January 1, 2027, with cached input at $0.075. Google reports DeepSWE v1.1 73.7% and roughly 40% more work completed per task than 3.7 Flash.
    How much does Gemini 3.8 Flash cost in yno.ai?
    Gemini 3.8 Flash is available on the Pro tier of yno.ai.
    What can Gemini 3.8 Flash do?
    Gemini 3.8 Flash is best for Teams that want a Google coding-and-agents workhorse at intro Flash pricing. Its main capabilities include Code, Agentic, Vision, Video, Speed, Multimodal.
    How does Gemini 3.8 Flash compare to other models?
    Gemini 3.8 Flash excels at Google-reported DeepSWE v1.1 73.7% and is recommended for Agentic coding, Real-time apps, Document processing, Multimodal pipelines. See related models below.
    How do I use Gemini 3.8 Flash in yno.ai?
    Sign up for yno.ai, select Gemini 3.8 Flash from the model picker, and start chatting.