Đặc điểm:
- Aesthetic quality tốt nhất (đặc biệt art, illustration)
- Chạy qua Discord bot hoặc Web UI
- Closed-source, subscription-based
- Hỗ trợ: text-to-image, image-to-image, vary, upscale
Pricing (2026):
- Basic: $10/month (~200 images)
- Standard: $30/month (~900 images)
- Pro: $60/month (unlimited relaxed)
中途參數
/imagine prompt: a dragon flying over mountains --ar 16:9 --v 6 --stylize 750
Parameters:
--ar 16:9 → Aspect ratio
--v 6 → Version
--stylize 750 → Creativity level (0-1000)
--chaos 50 → Variation (0-100)
--quality 2 → Detail level
--no text → Exclude elements
--tile → Seamless pattern
--seed 12345 → Reproducibility
2. 通量
# Flux: open-source từ Black Forest Labs (ex-Stability AI team)
# Kiến trúc: DiT (Diffusion Transformer) + T5 text encoder
# Quality ngang DALL-E 3, open-source
from diffusers import FluxPipeline
import torch
pipe = FluxPipeline.from_pretrained(
"black-forest-labs/FLUX.1-dev",
torch_dtype=torch.bfloat16,
)
pipe.to("cuda")
image = pipe(
prompt="A cat holding a sign that says 'Hello World'",
num_inference_steps=30,
guidance_scale=3.5,
height=1024,
width=1024,
).images[0]
通量變體
型號
許可證
品質
速度
Flux.1 專業版
僅限 API
最佳
快速
Flux.1 開發
非商業
太棒了
中等
Flux.1 施內爾
阿帕契2.0
好
非常快(4步)
3. 谷歌圖片 3
# Google Imagen 3 via Vertex AI
from google.cloud import aiplatform
from vertexai.preview.vision_models import ImageGenerationModel
model = ImageGenerationModel.from_pretrained("imagen-3.0-generate-001")
response = model.generate_images(
prompt="A peaceful Japanese garden with cherry blossoms",
number_of_images=4,
aspect_ratio="16:9",
safety_filter_level="block_some",
person_generation="dont_allow",
)
response[0].save("imagen_output.png")
4.Adobe Firefly API
# Adobe Firefly: trained on licensed content → commercially safe
import requests
headers = {
"Authorization": f"Bearer {FIREFLY_TOKEN}",
"Content-Type": "application/json",
}
response = requests.post(
"https://firefly-api.adobe.io/v2/images/generate",
headers=headers,
json={
"prompt": "A modern office space with plants",
"contentClass": "photo", # photo, art
"size": {"width": 2048, "height": 2048},
"n": 1,
"styles": {"presets": ["photo"]},
}
)
5. 比較平台
特色
標清/標清XL
達爾-E 3
中途
助焊劑
圖 3
螢火蟲
品質
太棒了
優秀
最佳(藝術)
優秀
優秀
好
圖片中的文字
可憐
太棒了
好
最佳
好
好
提示關注
好
優秀
好
優秀
優秀
好
速度
快速(本地)
〜10 秒
〜30秒
中
〜10 秒
〜15秒
成本/圖片
免費(GPU)
0.04-0.08 美元
~$0.03
免費/API
~$0.04
~$0.04
開源
是的
沒有
沒有
部分
沒有
沒有
商業用途
是的
是的
是(付費)
變化
是的
是的
微調
是的
沒有
沒有
是的
沒有
沒有
自架
是的
沒有
沒有
是的
沒有
沒有
6. 多模型編排
class ImageGeneratorOrchestrator:
"""Route to best model based on use case"""
def __init__(self):
self.dalle = OpenAI()
self.sd_pipe = StableDiffusionXLPipeline.from_pretrained(...)
async def generate(self, prompt, use_case="general"):
if use_case == "text_in_image":
# Flux/DALL-E best for text rendering
return await self._dalle_generate(prompt)
elif use_case == "artistic":
# Local SD with LoRA for custom styles
return self._sd_generate(prompt)
elif use_case == "commercial":
# Firefly for copyright-safe content
return await self._firefly_generate(prompt)
elif use_case == "batch":
# Local SD for cost efficiency
return self._sd_generate(prompt)
else:
return await self._dalle_generate(prompt)
async def _dalle_generate(self, prompt):
response = self.dalle.images.generate(
model="dall-e-3", prompt=prompt, size="1024x1024"
)
return response.data[0].url
def _sd_generate(self, prompt):
return self.sd_pipe(prompt, num_inference_steps=25).images[0]