Model icon
gemini-3.5-flash
Chat
模型:
Gemini 3.5 Flash API 支持对话和代码生成,适合生产级助手、代码 Agent、客服工具和结构化推理流程。
Chat

该 Playground 仅用于演示,数据只在当前窗口内有效,刷新页面后会被清空。

输入消息...

配置
1.0
1.0
定价详情

透明定价,无隐藏费用。按量付费,用多少付多少。

Googlegemini-3.5-flash
Input
PoYo价格
$0.900/1M tokens
180 积分
官方价格
$1.50/1M tokens
Official
为你节省
40%
Googlegemini-3.5-flash
Output
PoYo价格
$5.40/1M tokens
1080 积分
官方价格
$9.00/1M tokens
Official
为你节省
40%

* 实际费用以最终输出为准。

介绍

使用 PoYo 上的高性价比 Gemini 3.5 Flash API 的完整指南

PoYo 上的高性价比 Gemini 3.5 Flash API

无需 waitlist,即可在 PoYo 使用 Gemini 3.5 Flash API。Google DeepMind 将 Gemini 3.5 Flash 定位为面向高速 Agent 工作流、编码、多模态推理、1M context 和 64K text output 的 preview Flash 模型。支持 /v1/chat/completions 和现有 Gemini Native Format endpoint。

PoYo 上可用的 Gemini 3.5 Flash API 模型

01

Gemini 3.5 Flash

gemini-3.5-flash 是面向编码 Agent、长上下文多模态分析和工具型应用的快速 Gemini 模型。PoYo 按 1M tokens 计费:输入 180 credits,输出 1080 credits。
查看文档

Gemini 3.5 Flash API 核心特性

Gemini 3.5 Flash 重点面向速度、多模态上下文和 Agent 工作流,而不只是单轮聊天。

01

面向 Agent 工作流的快速 Flash 模型

Google DeepMind 将 Gemini 3.5 Flash 描述为 Gemini 3 Flash reasoning foundation 的下一次迭代,适合编码任务、Agent 工作流和需要更低延迟的企业流程。

  • Preview Flash 模型
  • 为 Agent 与编码优化
  • 速度与质量平衡良好
Gemini API

02

1M context 与多模态输入

官方 model card 列出 text、image、audio、video inputs 和最高 1M token context window。适合大型文档包、视频或音频分析和多模态产品工作流。

  • Text、image、audio、video 与 PDF inputs
  • 最高 1M token context window
  • 64K token text output
Gemini text generation

03

用 thinking levels 平衡质量、成本和延迟

Gemini 3.5 Flash 使用 thinking levels 控制质量、成本和延迟组合。系统可以在大多数回合快速响应,只在更难任务上开启更深推理。

  • 控制推理深度
  • 调节延迟与成本
  • 难任务使用更深 thinking
Gemini function calling

04

Gemini Native 与 OpenAI 兼容访问

PoYo 同时支持 Gemini 3.5 Flash 的 Gemini Native Format endpoint,可使用完整 Gemini 风格请求;也支持 /v1/chat/completions,方便偏好 OpenAI 兼容 Chat 客户端的团队接入。

  • /v1/chat/completions 用于兼容
  • /v1beta/models 用于 Gemini Native Format
  • 两个路径共用一个 API Key
Gemini API docs

你可以用 Gemini 3.5 Flash API 构建什么?

01

高速 AI Agent

构建能规划、调用工具并快速迭代的 Agent,同时比更重的 frontier reasoning 模型更省成本。

02

编码助手

用于 repo triage、debugging plans、代码解释,以及受益于快速模型回合的 Agentic coding 任务。

03

多模态分析

在需要比纯文本 Chat 更宽上下文的应用中处理 text、image、audio、video 和 PDF。

04

企业知识工作流

用 1M context 总结长文档包、基于内部文档推理,并自动化重复知识工作。

Gemini 3.5 Flash 与 GPT-5.5、Claude Opus 4.7 对比

将 Gemini 3.5 Flash 与两个高级推理模型对比:GPT-5.5 适合 OpenAI 风格编码 Agent 与 Responses API 工作流,Claude Opus 4.7 适合长上下文 hybrid reasoning、编码和 computer-use 任务。

特性Gemini 3.5 FlashGPT-5.5Claude Opus 4.7
提供商Google DeepMindOpenAIAnthropic
适合场景Gemini 3.5 Flash 适合快速 Agent 工作流、编码任务、多模态分析,以及需要速度与质量平衡的长上下文应用。高难度编码、规划、重工具 Agent 与高价值专业推理。深度推理、编码和 computer use
上下文窗口1M1M1M
最大输出64K128K128K
推理方式快速动态 thinkingxhigh GPT reasoningHybrid reasoning 与 X-High effort
Endpoint 支持Chat + Gemini NativeChat + ResponsesChat + Messages
PoYo 输入价格$0.90 / 1M$3.00 / 1M$4.00 / 1M
PoYo 输出价格$5.40 / 1M$18.00 / 1M$20.00 / 1M

Gemini 模型信息根据 2026 年 5 月 21 日检查的 Google DeepMind 与 Google AI 公开文档整理。GPT-5.5 与 Claude Opus 4.7 使用对应 PoYo 模型页配置。

如何在 PoYo 使用 Gemini 3.5 Flash API

第一步:创建 PoYo API Key
登录控制台,生成 API Key,并为 Chat API 使用充值 credits。
获取 API Key

第二步:选择 endpoint
需要 OpenAI 风格客户端时使用 /v1/chat/completions,需要 Gemini 原生请求结构时使用 /v1beta/models/gemini-3.5-flash:generateContent。

第三步:发送请求
curl --request POST \ --url https://api.poyo.ai/v1/chat/completions \ --header 'Authorization: Bearer YOUR_API_KEY' \ --header 'Content-Type: application/json' \ --data '{ "model": "gemini-3.5-flash", "messages": [{"role": "user", "content": "Plan a fast coding agent."}] }'
查看 API 文档

PoYo 上的 Gemini 3.5 Flash API 价格

计费项PoYo credits美元等价
Gemini 3.5 Flash input180 credits / 1M tokens$0.90 / 1M tokens
Gemini 3.5 Flash output1080 credits / 1M tokens$5.40 / 1M tokens

PoYo 定价约比官方便宜 40%。无需订阅费,可使用同一个 API Key、用量看板和按量付费 credits。

关于 Gemini 3.5 Flash API 的常见问题

Gemini 3.5 Flash 最适合什么?

Gemini 3.5 Flash 适合快速 Agent 工作流、编码任务、多模态分析,以及需要速度与质量平衡的长上下文应用。

PoYo 支持 Gemini 3.5 Flash 的哪些 endpoint?

PoYo 支持 /v1/chat/completions 和现有 Gemini Native Format endpoint:/v1beta/models/gemini-3.5-flash:generateContent。

Gemini 3.5 Flash 在 PoYo 上多少钱?

Gemini 3.5 Flash 在 PoYo 上输入 180 credits / 1M tokens,输出 1080 credits / 1M tokens,折合输入 $0.90、输出 $5.40 / 1M tokens。

Gemini 3.5 Flash 支持多模态输入吗?

支持。Google DeepMind 为 Gemini 3.5 Flash 列出 text、image、audio 和 video inputs,token context window 最高 1M,text output 最高 64K tokens。

为什么用 PoYo 访问 Gemini 3.5 Flash API

01

即时 Gemini 访问

无需单独 provider 配置,即可用同一个 PoYo dashboard 和 API key 开始使用 gemini-3.5-flash。

02

透明 token 定价

PoYo 展示清晰的输入与输出 token 价格,Gemini 3.5 Flash 约比官方价格便宜 40%。

03

两种集成方式

快速迁移可用 OpenAI 兼容 Chat;需要 Gemini 专用请求结构时使用 Gemini Native Format。

04

从 Playground 到生产

在模型页测试 prompts、监控用量,然后把同一个 model ID 放进生产环境。