Input
Output
Result valid for 24 hours
Qwen Image 2.0 - 强大的 AI 图像生成
AiBox 提供实惠且可靠的 Qwen Image 2.0 访问。使用标准版和 Pro 版模型生成精美图像。承诺 99.9% 在线率。
获取 API 密钥Qwen Image 2.0 核心功能
为 AI 图像生成工作流打造的高级特性
标准版与 Pro 版模型
选择 qwen-image-2.0 实现快速生成,或选择 qwen-image-2.0-pro 获取更高的图像质量和细节。
灵活的宽高比
支持多种宽高比,包括 16:9、1:1、9:16 等。可输出 PNG 或 JPG 格式。
图生图支持
上传参考图引导 Qwen 生成,支持风格迁移和图像变换工作流。
异步处理
提交任务并轮询结果,实现大规模生产环境中的无缝集成。
快速生成
Qwen 优化的基础设施提供极速图像生成,低延迟满足生产级工作负载。
全面的技术支持
获取详细的文档、技术支持和社区资源,轻松完成集成。
简单的接口集成
采用 OpenAI 兼容格式,可轻松集成到任何编程语言中。
高质量输出
采用最新图像生成技术,实现照片级真实感和艺术风格的图像创作。
Qwen Image 2.0 快速入门
按照以下简单步骤开始生成图像
注册
创建免费 AiBox 账户开始使用。
充值
为账户充值以开始使用服务。
获取密钥
在控制面板中创建密钥。
定价详情
透明定价,无隐藏费用。按量付费,用多少付多少。
| 规格 | 当前价格 | 官方价格 | 为你节省 |
|---|---|---|---|
| 默认 | 0.2 Credits/次 ~$0.02 | 0.25 Credits/次 ~$0.025 | 20% |
* 实际费用以最终输出为准。
常见问题
相关模型
gemini-2.5-flash-image-preview
Nano Banana (gemini-2.5-flash-image-preview) is Google DeepMind's fast, conversational image generation and editing model. It delivers natural-language text-to-image, precise multi-turn edits, strong character consistency, and multi-image fusion at low latency.
doubao-seedream-5-0-pro
Seedream 5.0 Pro (doubao-seedream-5-0-pro) is ByteDance's quality-first text-to-image model. It produces cinematic 1K and 2K images with best-in-class text rendering, supports up to 10 reference images for style and character consistency, and unifies generation and editing in a single API call.
midjourney
Midjourney is a leading text-to-image AI model renowned for its painterly aesthetics, strong composition, and cinematic lighting. Generate high-quality images from text prompts or reference images, with fine-grained control over style, aspect ratio, and creative parameters — accessible programmatically through AiBox's unified API, no Discord required.
wan2.7-image
wan‑2‑7‑image is an advanced image generation model in Alibaba’s Wan 2.7 series designed to create high‑quality visuals from text prompts. It excels at generating detailed, realistic images with accurate object representation and rich composition. The model supports multimodal input (e.g., combining text with reference images) to influence style and structure, and it’s well‑suited for creative workflows such as marketing assets, product visuals, social media graphics, and artistic content production. With strong semantic understanding and prompt adherence, wan‑2‑7‑image delivers both fidelity and expressive visual output.