
图像生成图像模型
图像生成图像 AI 模型
上传参考图后进行编辑、重绘、扩展或风格变体,让生成结果保留原图的主体、构图或视觉方向。
匹配模型
22 个模型

$0.04 / generation
nano-banana-pro
Google's advanced Gemini 3 Pro image model featuring 4K resolution, enhanced reasoning, real-time data integration, and multi-image composition capabilities.

$0.01 / generation
gpt-image-2
OpenAI's premium image generation family for text-to-image creation and single-image editing with simple prompt-driven controls.

$0.024 / generation
qwen-image-3
Alibaba's third-generation image model for long prompts, information-dense layouts, fine text rendering, multilingual visuals, and knowledge-rich image generation.

$0.075 / generation
seedream-5.0-pro
ByteDance's flagship image model for precise prompt following, dense layouts, multilingual text rendering, and region-precise editing.

$0.007 / generation
gpt-image-2.5-flare
OpenAI image generation and editing with high-fidelity reference images, precise inpainting, multi-turn consistency, and transparent backgrounds. Explore Flare and Sunburst and compare GPT Image 2.5 with GPT Image 2 on PoYo.

$0.04 / generation
grok-imagine-image-2.0
xAI image generation and editing with up to three input images, five supported aspect ratios, 1K or 2K resolution, selectable quality, and up to four outputs.

$0.03 / generation
qwen-image-2.1
Alibaba's 7B open-weight unified text-to-image and editing model with native RGBA transparency, up to 10 reference images, and 2K resolution.

$0.03 / generation
flux-2-pro
Black Forest Labs' production-grade model combining 4MP image generation and editing with multi-reference support, precise typography, and hex color control.

$0.01 / generation
gpt-image-1.5
OpenAI's latest image model with 4x speed, precision editing, and superior text rendering.

$0.025 / generation
nano-banana
Google's leaderboard-topping image model (Gemini 2.5 Flash) excelling in natural language editing, character consistency, and multi-image blending.

$0.025 / generation
seedream-4
ByteDance's multimodal image model for native 1K to 4K generation, reference-guided editing, and batch outputs up to 15 images.

$0.04 / generation
flux-kontext-pro
Black Forest Labs' in-context image generation and editing family for scene-preserving edits, text replacement, and prompt-driven visual changes.

$0.02 / generation
gpt-4o-image
OpenAI's native multimodal image generator with exceptional text rendering, precise prompt following, and conversational editing capabilities.

$0.025 / generation
seedream-4.5
ByteDance's unified 4K image generation and editing model with professional-grade text rendering and commercial photography quality.

$0.025 / generation
nano-banana-2-new
Google's next-gen image model powered by Gemini 3.1 Flash with native 2K/4K resolution, chain-of-thought reasoning, precise multi-language text rendering, and up to 14 reference images.

$0.025 / generation
seedream-5.0-lite
ByteDance's efficient multimodal image model with native 2K/3K resolution, deep thinking, bilingual text rendering, and smart editing.

$0.021 / generation
wan-2.7-image
Alibaba's unified Wan 2.7 image family for text-to-image generation and reference-based image editing with standard and pro quality tiers.

$0.025 / generation
nano-banana-2-lite
Google's fastest and most cost-efficient Gemini Image model, built for rapid ideation, high-throughput image generation, low-latency creative workflows, and scaled production pipelines.

$0.018 / generation
kling-o3-image
Kling's O3 image family supports prompt-only generation, multi-reference editing, single outputs, and connected series generation with 1K to 4K resolution tiers.

$0.018 / generation
kling-o1-image-edit
Kling O1 Image Edit focuses on reference-based image transformation with optional elements guidance, flexible aspect ratios, and 1K or 2K output tiers.

$0.02 / generation
flux-dev
Black Forest Labs image model for text-to-image generation and single-image editing with output-size based pricing.

$0.04 / generation
grok-imagine-image-quality
xAI's higher-fidelity Grok Imagine image model for polished text-to-image generation and reference-based image editing.
图像生成图像模型 API - 价格和模型适配
通过 PoYo.ai 调用 图像生成图像模型 API,查看透明的 API 价格、模型页面和统一的异步生成流程。
精确任务匹配
只有模型目录元数据包含 Image to Image 的模型才会出现在这里。
供应商对比
无需离开任务目录,即可查看不同供应商的模型系列。
API 就绪路径
每张模型卡都会跳转到 PoYo 的价格、示例、Playground 控件和 API 详情。
| 模型 | 供应商 | 任务类型 | 价格 |
|---|---|---|---|
| Nano Banana Pro nano-banana-pro | Text to Image, Image to Image | $0.04次生成 | |
| GPT Image 2 gpt-image-2 | OpenAI | Text to Image, Image to Image | $0.01次生成 |
| Qwen Image 3.0 qwen-image-3 | Alibaba | Text to Image, Image to Image, Uncensored | $0.024次生成 |
| Seedream 5.0 Pro seedream-5.0-pro | Seedream | Text to Image, Image to Image, Uncensored | $0.075次生成 |
| GPT Image 2.5 gpt-image-2.5-flare | OpenAI | Text to Image, Image to Image | $0.007次生成 |
| Grok Imagine Image 2.0 grok-imagine-image-2.0 | xAI | Text to Image, Image to Image | $0.04次生成 |
| Qwen Image 2.1 qwen-image-2.1 | Alibaba | Text to Image, Image to Image | $0.03次生成 |
| FLUX.2 flux-2-pro | Black Forest Labs | Text to Image, Image to Image | $0.03次生成 |
| GPT Image 1.5 gpt-image-1.5 | OpenAI | Text to Image, Image to Image | $0.01次生成 |
| Nano Banana nano-banana | Text to Image, Image to Image | $0.025次生成 | |
| Seedream 4 seedream-4 | Seedream | Text to Image, Image to Image | $0.025次生成 |
| Flux Kontext flux-kontext-pro | Black Forest Labs | Text to Image, Image to Image | $0.04次生成 |
| GPT-4o Image gpt-4o-image | OpenAI | Text to Image, Image to Image | $0.02次生成 |
| Seedream 4.5 seedream-4.5 | Seedream | Text to Image, Image to Image, Uncensored | $0.025次生成 |
| Nano Banana 2 nano-banana-2-new | Text to Image, Image to Image | $0.025次生成 | |
| Seedream 5.0 Lite seedream-5.0-lite | Seedream | Text to Image, Image to Image, Uncensored | $0.025次生成 |
| Wan 2.7 Image wan-2.7-image | Alibaba | Text to Image, Image to Image | $0.021次生成 |
| Nano Banana 2 Lite nano-banana-2-lite | Text to Image, Image to Image | $0.025次生成 | |
| Kling O3 Image kling-o3-image | Kling | Text to Image, Image to Image | $0.018次生成 |
| Kling O1 Image kling-o1-image-edit | Kling | Image to Image | $0.018次生成 |
| Flux Dev flux-dev | Black Forest Labs | Text to Image, Image to Image | $0.02次生成 |
| Grok Imagine Image Quality grok-imagine-image-quality | xAI | Text to Image, Image to Image | $0.04次生成 |
常见问题
PoYo.ai 上的图像生成图像页面包含什么?
+
这个页面汇总任务类型包含 Image to Image 的模型,覆盖参考图生成、图像编辑、风格变体和多图输入等工作流。
图像生成图像模型更适合哪些场景?
+
它更适合已有源图的场景,例如产品图改版、人物形象调整、海报变体、风格迁移、参考图扩写和视觉一致性生成。
接入图像生成图像 API 前应该先看什么?
+
建议先确认模型支持的输入图片数量、编辑方式、输出尺寸、价格和异步任务流程,再用同一组参考图对比模型效果。