AI Models and API Pricing

Category
Vendor
Capabilities

92 models · sorted by recommendation

NEW

claude-sonnet-5-5

Anthropic

A next-generation cost-efficient flagship workhorse featuring 30%+ faster output, optimized for everyday coding and professional workflows.

TOKEN1M ctx
Input$1.70/1M tokens
Output$8.50/1M tokens
Cache read$0.170/1M tokens
Cache write$2.12/1M tokens Save 15%
NEW

gpt-6.1-sol

OpenAI

A cost-efficient frontier model offering near-Astra intelligence, specializing in complex coding, computer use, and scientific research.

TOKEN922K ctx
Input$1.70 /1M tokens
Output$8.50 /1M tokens
Cache read$0.085 /1M tokensSave 15%
NEW

qwen-audio-3.0-tts-plus

阿里通义

Qwen-Audio-3.0-TTS-Plus is a high-performance large-scale text-to-speech model designed for high-quality speech generation scenarios.

PKGPer 万字符
Per 万字符$0.144 Save 30%
NEW

seed-tts-2.0

字节火山引擎

A next-generation high-fidelity speech synthesis model featuring lifelike natural prosody, rich emotional depth, and instant voice cloning.

Voice
PKGPer 万字符
Per 万字符$0.265 Save 40%

gpt-6-luna

OpenAI

A next-generation empathetic and creative foundation model specializing in deep resonance, artistic synthesis, and nuanced interaction.

TOKEN922K ctx
Input$0.085 /1M tokens
Output$0.425 /1M tokens
Cache read$0.0085 /1M tokensSave 15%

claude-opus-5-5

Anthropic

A top-tier cognitive and deep-reasoning model specializing in complex logic analysis, full-stack architecture, and extended thinking.

TOKEN1M ctx
Input$3.40/1M tokens
Output$17.00/1M tokens
Cache read$0.170/1M tokens
Cache write$4.25/1M tokens Save 15%

gpt-6-sol

OpenAI

A next-generation frontier foundation model featuring top-tier scientific insight, long-horizon planning, and autonomous problem-solving capabilities.

TOKEN922K ctx
Input$1.70 /1M tokens
Output$8.50 /1M tokens
Cache read$0.170 /1M tokensSave 15%
Dynamic billing

deepseek-v4-1-flash-260910

深度求索

A next-generation lightweight, high-efficiency reasoning model specializing in code architecture, mathematical derivation, and rapid response.

文本生成
TOKEN128K ctx
Input$0.118 /1M tokens
Output$0.471 /1M tokens
Cache read$0.0024 /1M tokensSave 20%
Picked

gemini-3.8-flash

Google

Google's most intelligent Flash model, designed for long-running software engineering, autonomous agents, and complex enterprise workflows.

文本生成VisionTools
TOKEN1M ctx
Input$0.637/1M tokens
Output$3.19/1M tokens
Cache read$0.064/1M tokens Save 15%

gpt-image-2.5-flare

OpenAI

A lightweight, next-generation image model combining ultra-fast rendering with lifelike, detailed visual generation.

Image generation
PKGPer 张
Per 张$0.050 Save 15%

gpt-image-2.5-sunburst

OpenAI

A flagship high-fidelity image generation model featuring superior lighting rendering, dynamic colors, and ultra-fine details.

Image generation
PKGPer 张
Per 张$0.050 Save 15%
Picked

claude-fable-5-1

Anthropic

Built for complex reasoning and long-running tasks, excelling at coding, multistep research, and document work.

Long contextCode文本生成
TOKEN1M ctx
Input$8.50/1M tokens
Output$42.50/1M tokens
Cache read$0.212/1M tokens
Cache write$10.62/1M tokens Save 15%
Picked

doubao-seedream-5-0-260128

字节火山引擎

Doubao Seedream 5.0 is Doubao’s flagship image-generation model, offering stronger knowledge grounding, reference consistency, prompt understanding, and professional visual creation for diverse commercial and creative scenarios across modern production workflows.

Image generation
PKGPer 张
Per 张$0.032
Picked

doubao-seed-2-1-pro-260628

字节火山引擎

Doubao Seed 2.1 Pro is a flagship multimodal reasoning model built for complex analysis, coding, long-context understanding, document processing, and reliable agentic workflows across demanding professional applications and business scenarios.

Code文本生成Long context
TOKENToken based
Input$0.882/1M tokens
Output$4.41/1M tokens
Cache read$0.176/1M tokens
Cache write$0.0025/1M tokens
Picked

gpt-6-astra

OpenAI

GPT-6 Astra is OpenAI’s flagship model for demanding end-to-end tasks. It excels at complex reasoning, software engineering, computer use, research, and professional document creation. The model supports text and image input, with a context window of approximately 1.05 million tokens and up to 128,000 output tokens.

文本生成VisionTools
TOKEN100K ctx
Input$8.50 /1M tokens
Output$42.50 /1M tokens
Cache read$0.850 /1M tokensSave 15%
Picked

wan3.0-video

阿里通义

Wan 3.0 Video is Alibaba Cloud’s all-in-one multimodal video generation model. It accepts text, images, video, audio, documents, and web links for text-to-video, first-and-last-frame generation, and reference-based video creation. It can natively generate dialogue, background music, and sound effects, with output up to 1080P, 30 FPS, and 30 seconds.

Video generation
PKGPer 秒
Per 秒$0.035 Save 20%
Picked

wan3.0-video-prime

阿里通义

阿里通义 Video

Video generation
PKGPer 秒
Per 秒$0.053 Save 20%
Picked

grok-4.6

xAI

Grok 4.6 is xAI’s flagship model for coding, agentic tasks, and knowledge work. It offers advanced reasoning, code generation, tool use, and long-running task execution. The model supports text and image input, text output, and a 500K-token context window, making it suitable for software development, research, complex analysis, and automated workflows.

Code文本生成Search
TOKENToken based
Input$3.71/1M tokens
Output$11.12/1M tokens
Cache read$0.926/1M tokens Save 10%

grok-video-3

xAI

Grok Imagine Video 3 is xAI’s video generation model for creating videos from text, images, and reference media, with optional preset voices. It is suitable for advertising, product visualization, creative production, and social media content.

Video generation
PKGPer 秒
Per 秒$0.013 Save 10%

glm-5.3

智谱

A next-generation bilingual foundation model featuring outstanding long-context analysis, code architecture, and deep-reasoning capabilities.

TOKENToken based
Input$1.18/1M tokens
Output$4.12/1M tokens
Cache read$0.294/1M tokens
Picked

Qwen 3.8-max

阿里通义

2.4万亿参数MoE旗舰,编程与办公能力全面跃升,可自主编程十数天交付完整项目

文本生成VisionTools
TOKEN1M ctx
Input$1.68/1M tokens
Output$5.03/1M tokens
Cache read$0.140/1M tokens
Cache write$2.10/1M tokens Save 5%
Picked

qwen-image-3.0-pro

阿里通义

Qwen-Image 3.0 Pro is the flagship model for professional image generation and editing. It offers stronger complex-instruction understanding, realistic textures, fine details, semantic alignment, and text rendering for advertising, branding, infographics, and information-dense layouts.

Image generation
PKGPer 张
Per 张$0.059 Save 20%

kling-v3-omni-video

可灵

Kling Video V3 Omni is a professional multimodal video model focused on advanced reference generation and exceptional consistency. It understands text, images, video, and audio while preserving character appearance, voice traits, objects, and visual style across newly generated scenes.

TOKENToken based
720p$0.106 without video input /1M tokens
720p$0.106 with video input /1M tokensSave 20%
Picked

MiniMax-H3

MiniMax

MiniMax Video

Video generation
PKGPer 秒
Per 秒$0.074

mimo-v2.5-pro

XiaoMi

Xiaomi’s flagship trillion-parameter MoE model for complex software engineering, tool use, and long-horizon agentic tasks.

Code文本生成Long context
TOKEN1M ctx
Input$0.772 /1M tokens
Output$2.32 /1M tokens
Cache read$0.154 /1M tokensSave 25%
Recommended during free time
Picked

DeepSeek-V4-Pro-0813

深度求索

A high-performance model for complex reasoning and agentic tasks.

文本生成ToolsJSON
TOKEN1M ctx
Input$0.596 /1M tokens
Output$1.79 /1M tokens
Cache read$0.060 /1M tokensSave 10%
Picked

Gemini 3.7 Flash

Google

A next-generation hybrid-reasoning multimodal model, combining tunable deep thinking with lightning-fast responsiveness.

文本生成VisionTools
TOKEN1M ctx
Input$0.637/1M tokens
Output$3.19/1M tokens
Cache read$0.064/1M tokens Save 15%
Picked

Kimi K3

月之暗面

Explore Kimi K3 for long-horizon coding, knowledge work, deep reasoning, visual understanding, and a 1-million-token context window.

文本生成VisionTools
TOKEN1M ctx
Input$2.79/1M tokens
Output$13.97/1M tokens
Cache read$0.279/1M tokens
Cache write$0.279/1M tokens Save 5%
Picked

Seedance-2-5

字节火山引擎

ByteDance's Seedance 2.5 is a next-generation AI video generation model that supports both text-to-video and image-to-video generation.

Video generation
TOKENToken based
480p$10.29 without video input /1M tokens
480p$6.18 with video input /1M tokens

Dreamina-Seedance-2-5

字节火山引擎

Seedance 2.5 is ByteDance’s multimodal video generation model for international markets. It accepts text, image, video, and audio inputs and can generate videos up to 30 seconds long, with reference-based creation, editing, and extension capabilities.

Video generation
TOKENToken based
480p$10.70 without video input /1M tokens
480p$6.40 with video input /1M tokens

MiniMax M2.7

MiniMax

MiniMax旗舰模型,原生深度集成多模态能力,高效处理百万字长文档

文本生成VisionTools
TOKEN1M ctx
Input$0.247/1M tokens
Output$0.988/1M tokens
Cache read$0.049/1M tokens
Cache write$0.309/1M tokens Save 20%
Picked

claude-fable-5

Anthropic

Claude Fable 5 是 Anthropic 能力最强的广泛发布模型

文本生成VisionTools
TOKEN1M ctx
Input$8.50/1M tokens
Output$42.50/1M tokens
Cache read$0.850/1M tokens
Cache write$10.62/1M tokens Save 15%
Picked

gpt-5.6-sol

OpenAI

Sol(旗舰模型)具备顶级语言理解力,是复杂推理的首选引擎

文本生成VisionTools
TOKEN1M ctx
Input$3.40 /1M tokens
Output$17.00 /1M tokens
Cache read$0.340 /1M tokensSave 15%
Picked

GLM-5.2

智谱

智谱 GLM-5.2, MIT 开源旗舰全面释放代码潜能,树立自主可控大模型标杆

文本生成VisionTools
TOKEN1M ctx
Input$0.941/1M tokens
Output$3.29/1M tokens
Cache read$0.235/1M tokens Save 20%

seedream-5-0-lite-260128

字节火山引擎

Seedream 5.0 Lite is ByteDance’s next-generation image generation and editing model for international markets. It offers stronger multimodal understanding, complex instruction following, subject consistency, and detail preservation. The model supports text-to-image, single- and multi-image editing, image blending, and sequential image generation, with output resolutions from 2K to 4K.

Image generation
PKGPer 张
Per 张$0.035

seedream-4-5-251128

字节火山引擎

Seedream 4.5 is ByteDance’s high-quality image generation and editing model. Compared with Seedream 4.0, it improves editing consistency, preservation of subjects and lighting, portrait quality, and small-text rendering. It supports text-to-image, multi-image blending, image editing, and sequential image creation at 2K or 4K resolution.

Image generation
PKGPer 张
Per 张$0.040
Picked

qwen-image-3.0

阿里通义

Qwen-Image 3.0 is a general-purpose image generation and editing model that balances output quality and response speed. It supports multilingual text rendering and is suitable for posters, web pages, interfaces, product visuals, and everyday creative production.

Image generation
PKGPer 张
Per 张$0.021 Save 20%

kling-v3-video

可灵

Kling Video V3 is a professional video generation model supporting videos of up to 15 seconds with native synchronized audio. It delivers improved photorealism, complex motion, multi-shot storytelling, text preservation, and subject consistency for advertising, short films, e-commerce, and cinematic production.

Video generation
PKGPer 秒
Per 秒$0.106 Save 20%
Self-hosted

minimax-h3-base

MiniMax

The 5090 Self- deploying general-purpose video generation model balancing visual quality, motion coherence, instruction following, and generation efficiency. Ideal for polished videos, advertising assets, and creative content.

Video generation
PKGPer 秒
Per 秒$0.021

mimo-v2.5

XiaoMi

A native omnimodal agent model for long-context reasoning across text, images, video, and audio by XiaoMi.

文本生成Long contextCode
TOKEN1M ctx
Input$0.309 /1M tokens
Output$1.54 /1M tokens
Cache read$0.062 /1M tokensSave 25%
Recommended during free time
Picked

Deepseek V4 Flash-0731

深度求索

A fast, cost-efficient model with strong agentic capabilities.

文本生成ToolsJSON
TOKEN1M ctx
Input$0.199 /1M tokens
Output$0.596 /1M tokens
Cache read$0.020 /1M tokensSave 10%
Picked

Gemini 3.6 Flash

Google

A cost-efficient, ultra-responsive multimodal model featuring millisecond-level inference and robust cross-modal understanding.

文本生成VisionTools
TOKENToken based
Input$0.637/1M tokens
Output$3.19/1M tokens
Cache read$0.064/1M tokens Save 15%
Picked

claude-sonnet-5

Anthropic

Claude Sonnet 5 深度强化逻辑交互,输出结果更精准稳定

文本生成VisionTools
TOKEN1M ctx
Input$1.70/1M tokens
Output$8.50/1M tokens
Cache read$0.170/1M tokens
Cache write$2.12/1M tokens Save 15%
Picked

gpt-5.6-terra

OpenAI

Terra(均衡模型)算力与效率精准权衡,广泛适配多元业务流

文本生成VisionTools
TOKEN1M ctx
Input$1.70 /1M tokens
Output$10.20 /1M tokens
Cache read$0.170 /1M tokensSave 15%

kimi-k2.7

月之暗面

月之暗面 Kimi,专注Agent开发与代码生成,高效解析超长文本

文本生成ReasoningLong context
TOKEN262.1K ctx
Input$0.765/1M tokens
Output$3.18/1M tokens
Cache read$0.076/1M tokens
Cache write$0.956/1M tokens Save 20%
Picked

GLM-5.1

智谱

智谱旗舰, 依托MIT协议开放前沿算法,全力驱动产业智能化升级

文本生成VisionTools
TOKEN200K ctx
Input$0.706 /1M tokens
Output$2.82 /1M tokens
Cache read$0.071 /1M tokensSave 20%
Self-hosted

minimax-h3-base-fast

MiniMax

The 5090 Self-Deploying speed-optimized video generation model that reduces processing time while maintaining solid visual quality and motion consistency. Ideal for rapid previews, batch production, and frequent creative workflows.

Video generation
PKGPer 秒
Per 秒$0.018
Recommended during peak hours
Picked

DeepSeek-v4-pro-0813(高峰特惠)

深度求索

A high-performance model for complex reasoning and agentic tasks.

Reasoning文本生成Tools
TOKEN1M ctx
Input$1.06/1M tokens
Output$3.18/1M tokens
Cache read$0.035/1M tokens
Cache write$0.0020/1M tokens Save 20%
Picked

Claude-Opus-5

Anthropic

Claude Opus 5 是 Anthropic 推出的 Opus 级别旗舰模型

文本生成VisionTools
TOKEN100K ctx
Input$4.25/1M tokens
Output$21.25/1M tokens
Cache read$0.425/1M tokens
Cache write$5.31/1M tokens Save 15%
Picked

gpt-5.6-luna

OpenAI

Luna(轻量模型)主打低延迟高吞吐,专为高频简易调用场景定制

文本生成VisionTools
TOKEN1M ctx
Input$0.170 /1M tokens
Output$1.02 /1M tokens
Cache read$0.017 /1M tokensSave 15%
Picked

Qwen3.7 Max

阿里通义

阿里通义旗舰 Agent 基座, 200 万上下文, 强代码 / 工具调用

TOKEN1M ctx
Input$1.32/1M tokens
Output$3.96/1M tokens
Cache read$0.264/1M tokens Save 20%
Picked

Seedance 2.0

字节火山引擎

字节火山 Seedance 2.0 视频生成, 突破物理动态限制,缔造电影级高清画质视频

Video generation
TOKENToken based
480p$6.76 without video input /1M tokens
480p$4.12 with video input /1M tokens

Claude Opus 4.8

Anthropic

Anthropic 旗舰, 突破推理极限兼备,顶尖Agent规划与代码实力

文本生成VisionTools
TOKEN1M ctx
Input$4.25/1M tokens
Output$21.25/1M tokens
Cache read$0.425/1M tokens
Cache write$5.31/1M tokens Save 15%

Kimi K2.6

月之暗面

月之暗面 Kimi, 聚焦自动化智能体编排,解锁超长文档深度分析范式

文本生成ToolsJSON
TOKEN262.1K ctx
Input$0.765/1M tokens
Output$3.18/1M tokens
Cache read$0.076/1M tokens
Cache write$0.956/1M tokens Save 20%
Self-hosted

minimax-h3-mini

MiniMax

The 5090 self-deploying lightweight and cost-efficient video generation model optimized for rapid creative validation and basic video production. Ideal for short-form videos, draft previews, and high-concurrency workloads.

Video generation
PKGPer 秒
Per 秒$0.015
Recommended during peak hours
Picked

DeepSeek-v4-flash-0731(高峰特惠)

深度求索

A fast, cost-efficient model with strong agentic capabilities.

文本生成ToolsJSON
TOKEN1M ctx
Input$0.353/1M tokens
Output$1.06/1M tokens
Cache read$0.012/1M tokens
Cache write$0.0020/1M tokens Save 20%
Picked

Doubao Seedance 2.0 Fast

字节火山引擎

Seedance 2.0 快速版, 低延迟高性价比

Video generation
TOKEN1K ctx
480p$4.08 without video input /1M tokens
480p$2.43 with video input /1M tokensSave 25%
Picked

GPT-5.5

OpenAI

OpenAI 新一代旗舰, 强化逻辑推理与工具链,深度重构软件开发工作流

文本生成VisionTools
TOKEN1M ctx
Input$4.25 /1M tokens
Output$25.50 /1M tokens
Cache read$0.425 /1M tokensSave 15%

Claude Opus 4.7

Anthropic

Anthropic Claude Opus 4.7, 专注攻坚复杂编程,推导过程严谨可靠

文本生成VisionTools
TOKEN1M ctx
Input$4.25/1M tokens
Output$21.25/1M tokens
Cache read$0.425/1M tokens
Cache write$5.31/1M tokens Save 15%

Gemini 3.5 Flash

Google

高速多模态 Gemini, 精妙平衡生成质量与运算成本,助力业务规模落地

文本生成VisionTools
TOKEN1M ctx
Input$1.27/1M tokens
Output$7.65/1M tokens
Cache read$0.128/1M tokens Save 15%

Qwen3.7 Plus

阿里通义

通义 3.7 多模态 Agent 模型, 精准识别UI控件并实操,编程调试一体化无缝衔接

TOKEN1M ctx
Input$0.235 /1M tokens
Output$0.941 /1M tokens
Cache read$0.024 /1M tokensSave 20%

Doubao Seedance 2.0 mini

字节火山引擎

字节火山 Seedance 2.0 视频生成, 突破物理动态限制,缔造电影级高清画质视频

Video generation
TOKENToken based
720p$1.35 without video input /1M tokens
720p$0.824 with video input /1M tokensSave 60%
Picked

DeepSeek V4 Pro

深度求索

深度求索旗舰, 突破代码与数学极限,开放百万字长文本协同框架

文本生成ToolsJSON
TOKEN1M ctx
Input$1.41/1M tokens
Output$2.82/1M tokens
Cache read$0.118/1M tokens Save 20%
Picked

HappyHorse 1.1 文生视频

阿里通义

阿里 HappyHorse 1.1 文生视频, 实现外语对白口型丝滑同步,有效打破语种传播壁垒

Video generation
PKGPer 秒
Per 秒$0.053 Save 20%

Claude Opus 4.6

Anthropic

Anthropic Claude Opus 4.6, 持续深化深层推理机制,致力于打造企业级智能中枢

文本生成VisionTools
TOKEN1M ctx
Input$4.25/1M tokens
Output$21.25/1M tokens
Cache read$0.425/1M tokens
Cache write$5.31/1M tokens Save 15%

GPT-5.4 mini

OpenAI

GPT-5.4 轻量版, 大幅优化响应速度,以极高性价比满足日常开发需求

文本生成ToolsJSON
TOKEN400K ctx
Input$0.637/1M tokens
Output$3.83/1M tokens
Cache read$0.068/1M tokens Save 15%

Qwen3.6 Plus

阿里通义

通义前代 Agentic 旗舰, 百万上下文, 深度推动企业工作流智能化

TOKEN1M ctx
Input$0.235 /1M tokens
Output$1.41 /1M tokens
Cache read$0.024 /1M tokensSave 20%

veo-3.1-lite-generate-preview

Google

轻量级AI视频生成模型,低成本快速创作高效输出

Video generation
PKGPer 秒
Per 秒$0.043 Save 15%
Picked

DeepSeek V4 Flash

深度求索

DeepSeek 高性价比版, 大幅降低Token调用门槛,普惠普及长文本应用

文本生成ToolsJSON
TOKEN1M ctx
Input$0.118/1M tokens
Output$0.235/1M tokens
Cache read$0.024/1M tokens Save 20%
Picked

HappyHorse 1.1 图生视频

阿里通义

阿里 HappyHorse 1.1 图生视频, 基于首帧解析自动生成对口唇形与背景运镜,交互自然

Video generation
PKGPer 秒
Per 秒$0.053 Save 20%

Claude Sonnet 4.6

Anthropic

均衡型 Claude, 实现代码与长文本处理能力的最优平衡,通用首选

文本生成VisionTools
TOKEN1M ctx
Input$2.55/1M tokens
Output$12.75/1M tokens
Cache read$0.255/1M tokens
Cache write$3.19/1M tokens Save 15%

Dreamina Seedance 2.0

字节火山引擎

海外版 Seedance 2.0 针对性优化本地化特效渲染能力

Video generation
TOKENToken based
480p$7.00 without video input /1M tokens
480p$4.30 with video input /1M tokens

Qwen3.6 Flash

阿里通义

通义高性价比 flash 档, 低延迟高吞吐, 百万上下文

TOKEN1M ctx
Input$0.141 /1M tokens
Output$0.847 /1M tokens
Cache read$0.014 /1M tokensSave 20%

veo-3.1-fast-generate-preview

Google

高速AI视频生成模型,优化响应速度与实时创作体验

Video generation
PKGPer 秒
Per 秒$0.085 Save 15%
Picked

HappyHorse 1.1 参考生视频

阿里通义

阿里 HappyHorse 1.1 参考生视频, 通过多模态特征提取,精准锁定IP角色神态与动作

Video generation
PKGPer 秒
Per 秒$0.053 Save 20%

Dreamina Seedance 2.0 Fast

字节火山引擎

海外版 Seedance 2.0 快速版, 低延迟高性价比

Video generation
TOKENToken based
480p$4.20 without video input /1M tokens
480p$2.48 with video input /1M tokensSave 25%

Dreamina Seedance 2.0-mini

字节火山引擎

Seedance 2.0 is an advanced multimodal AI video generation model developed by gobal ByteDance's Seed team.

Video generation
TOKENToken based
480p$1.40 without video input /1M tokens
480p$0.840 with video input /1M tokensSave 60%

veo-3.1-generate-preview

Google

专业级AI视频生成模型,支持高真实感智能化创作

Video generation
PKGPer 秒
Per 秒$0.170 Save 15%

万相 2.7 文生视频

阿里通义

通义万相 2.7 文生视频, 支持剧本直驱成片,融合多机位运镜与高品质音频混音

PKGPer 秒
Per 秒$0.071 Save 20%

Gemini 3.1 Flash-Lite

Google

轻量多模态模型,支持视听图文输入,专为高并发API与轻量Agent打造

文本生成VisionTools
TOKENToken based
Input$0.212/1M tokens
Output$1.27/1M tokens
Cache read$0.021/1M tokens Save 15%

Nano Banana 2 Lite

Google

Gemini极速图像模型,极致压缩运营成本,专为大规模商业化部署打造

Image generationTools
TOKENToken based
Input$0.212/1M tokens
Output$25.50/1M tokens Save 15%

万相 2.7 图生视频

阿里通义

通义万相 2.7 图生视频, 静态画作秒变流畅影像,保留原作风格并新增声音维度

PKGPer 秒
Per 秒$0.071 Save 20%

万相 2.7 参考生视频

阿里通义

通义万相 2.7 参考生视频, 注入角色与道具约束,确保跨镜头叙事视觉绝对一致

PKGPer 秒
Per 秒$0.071 Save 20%

Nano Banana 2

Google

全能主力模型,兼顾高质量生成与广博知识库,擅长多参考图一致性编辑

Image generationTools
TOKENToken based
Input$0.425/1M tokens
Output$51.00/1M tokens Save 15%

GPT Image 2

OpenAI

OpenAI 图像生成, 输出电影级高清画质,支持精准局部重绘与风格化编辑

Image generationVision
PKGPer 张
Per 张$0.050 Save 15%

Qwen-Image 2.0

阿里通义

通义新一代图像模型, 7B / 原生 2K, 创意生成与版面精修二合一

PKGPer 张
Per 张$0.024 Save 20%

Nano Banana Pro

Google

Gemini 原生图像生成与编辑,覆盖从概念发散到精细打磨的一站式视觉生产

Image generationVision
TOKENToken based
Input$1.70/1M tokens
Output$102.00/1M tokens Save 15%

Qwen-Image 2.0 Pro

阿里通义

通义图像专业版, 专攻4K出版级排版,精准还原文字细节,杜绝排版错漏

PKGPer 张
Per 张$0.059 Save 20%

HappyHorse 1.0 文生视频

阿里通义

阿里 HappyHorse 1.0 文生视频, 摒弃迭代扩散模式,单次前向传播即刻生成视频

Video generation
PKGPer 秒
Per 秒$0.099 Save 20%

HappyHorse 1.0 图生视频

阿里通义

阿里 HappyHorse 1.0 图生视频, 采用极简拓扑网络,将静态海报瞬间转化动态短片

Video generation
PKGPer 秒
Per 秒$0.099 Save 20%

HappyHorse 1.0 参考生视频

阿里通义

阿里 HappyHorse 1.0 参考生视频, 支持最多五张草图并行约束,严密把控空间透视感

Video generation
PKGPer 秒
Per 秒$0.099 Save 20%

HappyHorse 1.0 视频编辑

阿里通义

阿里 HappyHorse 1.0 视频编辑, 仅需语音指令干预关键帧,免重渲染实现视频修改

Video generation
PKGPer 秒
Per 秒$0.099 Save 20%