
Văn bản cho mô hình hình ảnh
Mô hình AI chuyển văn bản thành hình ảnh
Tạo ra các khái niệm sản phẩm, tài sản tiếp thị, bản vẽ minh họa và hình ảnh sẵn sàng API trực tiếp từ các prompt bằng văn bản.
Các mô hình phù hợp
Các mô hình 26

$0.04 / generation
nano-banana-pro
Google's advanced Gemini 3 Pro image model featuring 4K resolution, enhanced reasoning, real-time data integration, and multi-image composition capabilities.

$0.01 / generation
gpt-image-2
OpenAI's premium image generation family for text-to-image creation and single-image editing with simple prompt-driven controls.

$0.024 / generation
qwen-image-3
Alibaba's third-generation image model for long prompts, information-dense layouts, fine text rendering, multilingual visuals, and knowledge-rich image generation.

$0.075 / generation
seedream-5.0-pro
ByteDance's flagship image model for precise prompt following, dense layouts, multilingual text rendering, and region-precise editing.

$0.007 / generation
gpt-image-2.5-flare
OpenAI image generation and editing with high-fidelity reference images, precise inpainting, multi-turn consistency, and transparent backgrounds. Explore Flare and Sunburst and compare GPT Image 2.5 with GPT Image 2 on PoYo.

$0.04 / generation
grok-imagine-image-2.0
xAI image generation and editing with up to three input images, five supported aspect ratios, 1K or 2K resolution, selectable quality, and up to four outputs.

$0.03 / generation
qwen-image-2.1
Alibaba's 7B open-weight unified text-to-image and editing model with native RGBA transparency, up to 10 reference images, and 2K resolution.

$0.02 / generation
nano-banana-2.1
Generate and edit 1K, 2K and 4K images with Google's Nano Banana 2.1. Combine up to 10 reference images for product visuals, posters and consistent creative assets. From 4 credits ($0.02) per image.

$0.02 / generation
seedream-5.0-flash
ByteDance's fast, cost-efficient image model for 1K/1.5K/2K generation, multi-reference editing, and multilingual visual content. 4 credits per image.

$0.03 / generation
flux-2-pro
Black Forest Labs' production-grade model combining 4MP image generation and editing with multi-reference support, precise typography, and hex color control.

$0.01 / generation
gpt-image-1.5
OpenAI's latest image model with 4x speed, precision editing, and superior text rendering.

$0.03 / generation
grok-imagine-image
xAI's Aurora-powered visual AI for image generation and video creation with Fun, Normal, and Spicy creative modes.

$0.025 / generation
nano-banana
Google's leaderboard-topping image model (Gemini 2.5 Flash) excelling in natural language editing, character consistency, and multi-image blending.

$0.025 / generation
seedream-4
ByteDance's multimodal image model for native 1K to 4K generation, reference-guided editing, and batch outputs up to 15 images.

$0.01 / generation
z-image
Alibaba's efficient 6B-parameter image model with sub-second generation and exceptional Chinese-English bilingual text rendering capabilities.

$0.04 / generation
flux-kontext-pro
Black Forest Labs' in-context image generation and editing family for scene-preserving edits, text replacement, and prompt-driven visual changes.

$0.02 / generation
gpt-4o-image
OpenAI's native multimodal image generator with exceptional text rendering, precise prompt following, and conversational editing capabilities.

$0.025 / generation
seedream-4.5
ByteDance's unified 4K image generation and editing model with professional-grade text rendering and commercial photography quality.

$0.025 / generation
nano-banana-2-new
Google's next-gen image model powered by Gemini 3.1 Flash with native 2K/4K resolution, chain-of-thought reasoning, precise multi-language text rendering, and up to 14 reference images.

$0.025 / generation
seedream-5.0-lite
ByteDance's efficient multimodal image model with native 2K/3K resolution, deep thinking, bilingual text rendering, and smart editing.

$0.021 / generation
wan-2.7-image
Alibaba's unified Wan 2.7 image family for text-to-image generation and reference-based image editing with standard and pro quality tiers.

$0.025 / generation
nano-banana-2-lite
Google's fastest and most cost-efficient Gemini Image model, built for rapid ideation, high-throughput image generation, low-latency creative workflows, and scaled production pipelines.

$0.018 / generation
kling-o3-image
Kling's O3 image family supports prompt-only generation, multi-reference editing, single outputs, and connected series generation with 1K to 4K resolution tiers.

$0.02 / generation
flux-dev
Black Forest Labs image model for text-to-image generation and single-image editing with output-size based pricing.

$0.002 / generation
flux-schnell
Black Forest Labs fast text-to-image model for low-cost prompt-based image generation with output-size based pricing.

$0.04 / generation
grok-imagine-image-quality
xAI's higher-fidelity Grok Imagine image model for polished text-to-image generation and reference-based image editing.
Văn bản sang hình ảnh Model API - Giá cả và mô hình phù hợp
Gọi đến các API mô hình Văn bản sang hình ảnh thông qua PoYo.ai để xem xét giá cả minh bạch API, trang mô hình và một dòng phát sinh bất đồng bộ thống nhất.
Đúng phù hợp với nhiệm vụ
Các mô hình chỉ xuất hiện ở đây khi các metadata danh mục của họ bao gồm Text to Image.
So sánh nhà cung cấp
Xem xét các gia đình mô hình trên các nhà cung cấp mà không rời khỏi thư mục nhiệm vụ.
Các đường dẫn sẵn sàng cho API
Mỗi thẻ mô hình liên kết đến giá PoYo, ví dụ, điều khiển sân chơi và chi tiết API.
| Mô hình | Nhà cung cấp | Các loại nhiệm vụ | Giá |
|---|---|---|---|
| Nano Banana Pro nano-banana-pro | Text to Image, Image to Image | $0.04lần tạo | |
| GPT Image 2 gpt-image-2 | OpenAI | Text to Image, Image to Image | $0.01lần tạo |
| Qwen Image 3.0 qwen-image-3 | Alibaba | Text to Image, Image to Image, Uncensored | $0.024lần tạo |
| Seedream 5.0 Pro seedream-5.0-pro | Seedream | Text to Image, Image to Image, Uncensored | $0.075lần tạo |
| GPT Image 2.5 gpt-image-2.5-flare | OpenAI | Text to Image, Image to Image | $0.007lần tạo |
| Grok Imagine Image 2.0 grok-imagine-image-2.0 | xAI | Text to Image, Image to Image | $0.04lần tạo |
| Qwen Image 2.1 qwen-image-2.1 | Alibaba | Text to Image, Image to Image | $0.03lần tạo |
| Nano Banana 2.1 nano-banana-2.1 | Text to Image, Image to Image | $0.02lần tạo | |
| Seedream 5.0 Flash seedream-5.0-flash | Seedream | Text to Image, Image to Image | $0.02lần tạo |
| FLUX.2 flux-2-pro | Black Forest Labs | Text to Image, Image to Image | $0.03lần tạo |
| GPT Image 1.5 gpt-image-1.5 | OpenAI | Text to Image, Image to Image | $0.01lần tạo |
| Grok Imagine grok-imagine-image | xAI | Text to Image, Image to Video, Uncensored | $0.03lần tạo |
| Nano Banana nano-banana | Text to Image, Image to Image | $0.025lần tạo | |
| Seedream 4 seedream-4 | Seedream | Text to Image, Image to Image | $0.025lần tạo |
| Z-Image z-image | Alibaba | Text to Image, Uncensored | $0.01lần tạo |
| Flux Kontext flux-kontext-pro | Black Forest Labs | Text to Image, Image to Image | $0.04lần tạo |
| GPT-4o Image gpt-4o-image | OpenAI | Text to Image, Image to Image | $0.02lần tạo |
| Seedream 4.5 seedream-4.5 | Seedream | Text to Image, Image to Image, Uncensored | $0.025lần tạo |
| Nano Banana 2 nano-banana-2-new | Text to Image, Image to Image | $0.025lần tạo | |
| Seedream 5.0 Lite seedream-5.0-lite | Seedream | Text to Image, Image to Image, Uncensored | $0.025lần tạo |
| Wan 2.7 Image wan-2.7-image | Alibaba | Text to Image, Image to Image | $0.021lần tạo |
| Nano Banana 2 Lite nano-banana-2-lite | Text to Image, Image to Image | $0.025lần tạo | |
| Kling O3 Image kling-o3-image | Kling | Text to Image, Image to Image | $0.018lần tạo |
| Flux Dev flux-dev | Black Forest Labs | Text to Image, Image to Image | $0.02lần tạo |
| Flux Schnell flux-schnell | Black Forest Labs | Text to Image | $0.002lần tạo |
| Grok Imagine Image Quality grok-imagine-image-quality | xAI | Text to Image, Image to Image | $0.04lần tạo |
Những câu hỏi thường được hỏi
Trang văn bản-đối với hình ảnh PoYo.ai bao gồm những gì?
+
Trang này liệt kê các mô hình có metadata nhiệm vụ bao gồm Text to Image, với thẻ mô hình, giá khởi điểm, nhà cung cấp, nhãn nhiệm vụ và liên kết đến các trang mô hình đầy đủ.
Các quy trình làm việc nào là mô hình văn bản-được hình ảnh tốt nhất cho?
+
Chúng là tốt nhất để tạo ra hình ảnh mới từ các mô tả bằng văn bản, bao gồm quảng cáo, khái niệm sản phẩm, minh họa, bìa, hình ảnh nhỏ và bản thảo hình ảnh khối lượng cao.
Tôi nên kiểm tra điều gì trước khi tích hợp một văn bản-to-photos API?
+
So sánh giá cả, tham số đầu vào, kích thước đầu ra, chất lượng theo dõi nhanh và các ví dụ trước, sau đó mở trang mô hình cho các tài liệu API và dòng phát triển không đồng bộ.
Khám phá tất cả các mô hình AI
Mở thị trường mô hình đầy đủ để so sánh hình ảnh, video, âm thanh, trò chuyện, 3D và mô hình công cụ.
Xem tất cả các mô hìnhXây dựng bằng API
Sử dụng tài liệu PoYo API khi bạn sẵn sàng gửi các công việc, tình trạng cuộc thăm dò hoặc nhận được cuộc gọi trở lại.
Các tài liệu API