Model icon
gemini-3.8-flash
Chat
Model:
Gemini 3.8 Flash is Google's most intelligent Flash model for autonomous agents, long-horizon software engineering, and 1M-context knowledge workflows.
Chat

This Playground is for demo purposes. Data is only valid in the current window and will be cleared on refresh.

Type a message...

Configuration
1.0
1.0
Pricing details

Transparent pricing with no hidden fees. Pay as you go.

Googlegemini-3.8-flash
Input
PoYo price
$0.600/1M tokens
120 credits
Official price
$0.750/1M tokens
Official
You save
20%
Googlegemini-3.8-flash
Cached input
PoYo price
$0.060/1M tokens
12 credits
Official price
$0.075/1M tokens
Official
You save
20%
Googlegemini-3.8-flash
Output (including thinking tokens)
PoYo price
$3.00/1M tokens
600 credits
Official price
$3.75/1M tokens
Official
You save
20%

* Actual fees are based on the final output.

Available on PoYo

Complete guide to using Gemini 3.8 Flash API for Autonomous Agents & Long-Horizon Coding

Gemini 3.8 Flash API for Autonomous Agents & Long-Horizon Coding

Gemini 3.8 Flash is Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It supports a 1,048,576-token input context and up to 65,536 output tokens.

Use gemini-3.8-flash through /v1/chat/completions or the Gemini-native /v1beta/models path. PoYo charges 120 input, 12 cached-input, and 600 output credits per million tokens—20% below Google's introductory Standard rates.

Key Features of Gemini 3.8 Flash

Official Google demonstrations showcase Gemini 3.8 Flash building interactive 3D worlds, functional retro software, scientific visualizations, and hardware teardowns.

01

Build 3D Interactive Games from Prompts

Using looping instructions in Google Antigravity, Gemini 3.8 Flash generates full 3D interactive levels with puzzles, procedural textures, and complete environmental gameplay.

  • Multi-step 3D logic and debugging
  • Procedural texture integration
  • Autonomous agent iteration loops
Official Gemini 3.8 Flash 3D game generation demonstration

02

Recreate Functional Classic Software in a Single Prompt

Gemini 3.8 Flash builds a fully playable DOS version of Google Maps in a single prompt, complete with location search, turn-by-turn routing, and terminal-rendered Street View.

  • Single-prompt architectural design
  • Complex algorithm implementation
  • Retro terminal UI rendering
Official Gemini 3.8 Flash DOS Google Maps demonstration

03

Scientific Geographic Exploration & Realtime Modeling

Synthesizes real datasets from the U.S. Geological Survey into real-time 3D topographic slices, 2D projections, and domain-grounded scientific explanations.

  • Real-world scientific dataset integration
  • Dynamic 2D/3D spatial modeling
  • Domain-specific analytical narration
Official Gemini 3.8 Flash topographic map analysis demonstration

04

Interactive 3D Hardware Anatomy & Teardowns

Generates realistic Three.js visualizers of physically-proportioned hardware teardowns, automatically decomposing devices into inspectable exploded layers with interactive sliders.

  • Precise 3D spatial decomposition
  • Automated Three.js rendering code
  • Interactive inspection controls
Official Gemini 3.8 Flash hardware anatomy visualizer demonstration

What Can You Build with Gemini 3.8 Flash?

01

Long-Horizon Software Engineering

Resolve multi-file refactoring, autonomous debugging, and complex engineering tasks end-to-end with high diligence.

02

Autonomous Enterprise Agents

Deploy reliable agents capable of recursive reasoning, multi-step planning, tool orchestration, and self-verification.

03

Quantitative & Professional Analysis

Analyze complex financial models, legal documents, scientific reports, and enterprise data with high domain precision.

04

3D & Interactive Application Generation

Generate full 3D interactive applications, procedural games, and retro software with code and visual grounding.

05

Multimodal Document & Spatial Understanding

Process text, high-resolution images, video, audio, and technical PDFs with 1M-token context recall.

06

Tunable Thinking Levels

Select low, medium, or high thinking effort levels to flexibly balance execution speed, token cost, and reasoning depth.

Gemini 3.8 Flash vs Gemini 3.7 Flash vs Claude Sonnet 5 vs GPT-5.6 Terra

September 2026 benchmarks compare Gemini 3.8 Flash against Gemini 3.7 Flash, Claude Sonnet 5, and GPT-5.6 Terra across software engineering, agentic autonomy, document intelligence, and expert reasoning.

BenchmarkGemini 3.8 FlashGemini 3.7 FlashClaude Sonnet 5GPT-5.6 TerraNotes
DeepSWE v1.173.7%65.3%53.8%69.6%Long-horizon software engineering
Terminal-Bench 2.189.4%85.8%80.4%87.4%Agentic terminal coding
HLE-Verified54.9%53.6%31.0%51.1%Multidisciplinary expert reasoning
Vals Finance Agent v261.4%59.0%56.2%60.1%Financial workflow autonomy
Harvey Legal Agent Benchmark10.0%8.8%7.9%9.2%Full legal task resolution
WebDev Arena (Elo)1612158815411523Interactive web development
Artificial Analysis Intelligence Index59565557Composite intelligence
AutomationBench34.2%30.4%10.7%23.6%Enterprise workflow automation
GDP.pdf36.8%34.0%28.0%24.7%Complex document processing
GDM-MRCR v2 (128k)98.2%97.0%81.5%93.5%Long-context retrieval & recall

Official evaluation results as of September 2, 2026. Benchmarks follow standard published evaluation harnesses.

How to Use Gemini 3.8 Flash on PoYo

1. Create an API key. Sign in and create a PoYo key.

2. Choose a format. Use POST /v1/chat/completions for OpenAI-compatible clients, or POST /v1beta/models/gemini-3.8-flash:generateContent for Gemini-native payloads. Streaming is available through :streamGenerateContent.

3. Set the model. Send gemini-3.8-flash as the model ID. Read the Gemini API documentation.

Gemini 3.8 Flash API Frequently Asked Questions

Is Gemini 3.8 Flash available on PoYo now?

Yes. Use the gemini-3.8-flash model ID with either supported endpoint format.

Which endpoints are supported?

Use /v1/chat/completions for OpenAI-compatible requests or /v1beta/models/gemini-3.8-flash:generateContent for Gemini-native requests. Both non-streaming and streaming flows are supported.

How much does Gemini 3.8 Flash cost?

PoYo charges 120 credits per million input tokens, 12 credits per million cached-input tokens, and 600 credits per million output tokens (including thinking tokens)—matching our 3.7 Flash rates and 20% below Google.

What media can the model understand?

Google lists text, image, video, audio, and PDF input with text output. Native image generation, audio generation, and Live API output are not supported by this model.

How large is the context window?

The model supports up to 1,048,576 input tokens and 65,536 output tokens.

Which thinking levels are available?

Gemini 3.8 Flash supports low, medium, and high thinking levels. The minimal level is not supported.

Does cached input include storage charges?

No. The 12-credit rate applies to cached input tokens read by a request. Any separate context-cache storage duration charge is outside that rate.

How is Gemini 3.8 Flash different from 3.7 Flash?

Gemini 3.8 Flash is engineered for long-horizon coding and agentic workflows, achieving 73.7% on DeepSWE v1.1 (+8.4% over 3.7 Flash) and 89.4% on Terminal-Bench 2.1, with higher diligence in tool calling and multi-step verification.

What is the knowledge cutoff?

Google lists March 2026 for most domains, with some areas potentially limited to January 2025. Use Search Grounding or retrieval for current information.

Does Gemini 3.8 Flash support Computer Use or fine-tuning?

Computer Use is available as a Preview feature. Fine-tuning is not supported.

When does Google's introductory pricing end?

Google's introductory Standard rates end December 31, 2026. Google lists higher Standard rates beginning January 1, 2027; PoYo pricing will remain highly competitive.

Why Use Gemini 3.8 Flash on PoYo

01

Two API Formats

Use OpenAI-compatible Chat Completions or Gemini-native GenerateContent with one model ID.

02

One PoYo Key

Evaluate and ship Gemini alongside other models without managing another integration credential.

03

Transparent 20% Savings

See input, cached-input, and output rates in both credits and US-dollar equivalents.

04

Playground and Monitoring

Test both request formats, review service status, and track usage before moving to production.