Category

doubao-seed3d-2-0-260328

ByteDance’s next-generation, disruptive, high-precision, simulation-grade 3D asset generation large model
API
Image Processing
Pricing:
$11/1M Tokens

doubao-seed-2-1-turbo-260628

ByteDance’s new-generation, low-cost, low-latency, deep-thinking large model designed for large-scale production scenarios
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.43/1M tokens
Output:
$2.14/1M tokens

doubao-seed-2-1-pro-260628

ByteDance’s next-generation flagship model, designed for the era of Coding and Agents, featuring deep thinking capabilities.
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.86/1M tokens
Output:
$4.28/1M tokens

glm-5.2

Zhipu AI’s next-generation, self-developed, high-end flagship large model specifically designed for multimodal agents.
Model
LLM
Model capability: thinkingModel capability: function_call
Input:
$1.4/1M tokens
Output:
$4.4/1M tokens

kimi-k2.7-code

Kimi’s next-generation flagship large model for adaptive code engineering based on deep reinforcement learning
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.95/1M tokens
Output:
$4/1M tokens

qwen3.7-plus

Tongyi Qianwen 3.7: The Most Cost-Effective Multimodal Intelligent Agent Foundation in the Family
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.228/1M tokensstarting from
Output:
$0.92/1M tokensstarting from

MiniMax-M3

MiniMax’s new-generation trillion-parameter MoE multimodal flagship large model.
Model
LLM
Model capability: audioModel capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.6/1M tokensstarting from
Output:
$2.4/1M tokensstarting from

step-3.7-flash

[30-Day Limited-Time Free] StepStar’s Flagship Language Reasoning Model
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Pricing:
Limited Time Free

qwen3.7-max

Alibaba’s Qwen 3.7, a closed-source flagship large model designed specifically for the era of intelligent agents.
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$1.8/1M tokens
Output:
$5.3/1M tokens
LLM

doubao-seed-2-1-turbo-260628

ByteDance’s new-generation, low-cost, low-latency, deep-thinking large model designed for large-scale production scenarios
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.43/1M tokens
Output:
$2.14/1M tokens

doubao-seed-2-1-pro-260628

ByteDance’s next-generation flagship model, designed for the era of Coding and Agents, featuring deep thinking capabilities.
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.86/1M tokens
Output:
$4.28/1M tokens

glm-5.2

Zhipu AI’s next-generation, self-developed, high-end flagship large model specifically designed for multimodal agents.
Model
LLM
Model capability: thinkingModel capability: function_call
Input:
$1.4/1M tokens
Output:
$4.4/1M tokens

kimi-k2.7-code

Kimi’s next-generation flagship large model for adaptive code engineering based on deep reinforcement learning
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.95/1M tokens
Output:
$4/1M tokens

qwen3.7-plus

Tongyi Qianwen 3.7: The Most Cost-Effective Multimodal Intelligent Agent Foundation in the Family
Model
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Input:
$0.228/1M tokensstarting from
Output:
$0.92/1M tokensstarting from
Image Generations

wan2.7-image-pro

Alibaba’s Tongyi Lab has launched a professional-grade AI image-generation tool—a full-scene visual creation assistant that combines high aesthetics with strong controllability.
API
Image Generations
Pricing:
$0.08/image

wan2.7-image

Alibaba’s Tongyi Lab has launched a professional-grade AI image-generation tool—a full-scene visual creation assistant that combines high aesthetics with strong controllability.
API
Image Generations
Pricing:
$0.03/image

Kling Image O3

Kuaishou’s AI image generation model features two core capabilities: Text-to-Image and Image Edit.
API
Image Generations
Pricing:
$0.028/image

starting from

Z-Image

The latest image-generation model released by Tongyi Lab
API
Image Generations
Pricing:
$0.05/call
Image Processing

doubao-seed3d-2-0-260328

ByteDance’s next-generation, disruptive, high-precision, simulation-grade 3D asset generation large model
API
Image Processing
Pricing:
$11/1M Tokens

qwen-image-edit-plus-2025-12-15

The Tongyi Qianwen series Image Editing Plus model further optimizes inference performance and system stability based on the initial Edit model.
API
Image Processing
Pricing:
$0.03/image

Qwen-Image-Layered

A model that decomposes an image into multiple RGBA layers endows the image with inherent editability.
API
Image Processing
Pricing:
$0.05/call
Video Generation

happyhorse-1.0-r2v

Alibaba Group’s next-generation cutting-edge AI video generation model
API
Video Generation
Pricing:
$0.156/second

starting from

happyhorse-1.0-i2v

Alibaba Group’s next-generation cutting-edge AI video generation model
API
Video Generation
Pricing:
$0.156/second

starting from

happyhorse-1.0-t2v

Alibaba Group’s next-generation cutting-edge AI video generation model
API
Video Generation
Pricing:
$0.156/sec

starting from

wan2.7-videoedit

The video editing features of the 2.7 series consistently preserve detailed information such as the image subject, style, and text.
API
Video Generation
Pricing:
$0.1/second

starting from

Audio-Video Processing

GLM-ASR-2512

Zhipu's next-generation speech recognition model supports real-time conversion of speech into high-quality text.
API
Audio-Video Processing
Pricing:
$0.025/M tokens

GLM-TTS

GLM Speech Synthesis Model Combining Large Language Models and Diffusion Model Technologies
API
Audio-Video Processing
Pricing:
$0.03/1000 characters
Data Processing

GLM-OCR Layout analysis

The layout analysis model released by Zhipu is used to parse the layout of documents and images and extract text content.
API
Data Processing
Pricing:
$0.03/M Tokens
RAG-related

qwen3-rerank

Text ranking model trained based on the Qwen LLM foundation
API
Input:
$0.07/1M tokens
Output:
Free
Tools API

Anime to Real Person

Explore various creative approaches to anime-to-live-action adaptations
API
Tools API
Pricing:
Optimize tokens generated by prompts + image generation API costs

Clothing Flat Lay

Explore various creative ways to use clothing flat lays
API
Tools API
Pricing:
Optimize tokens generated by prompts + image generation API costs

3D Doll

Explore various creative ways to play with 3D dolls
API
Tools API
Pricing:
Optimize tokens generated by prompts + image generation API costs

City in Toy Box

Explore multiple creative ways to play with the city in the toy box
API
Tools API
Pricing:
Optimize tokens generated by prompts + image generation API costs