Every model below is callable right now with the same API key and the same request format.
DeepSeek V4 Flash
deepseek-ai/DeepSeek-V4-Flash
977K contextTool callingJSON modeReasoningContext caching
- Input
- $0.5280
- Output
- $1.5840
- Cached input
- $0.0170
DeepSeek V4 Pro
deepseek-ai/DeepSeek-V4-Pro
977K contextTool callingJSON modeReasoningContext caching
- Input
- $1.5840
- Output
- $4.7520
- Cached input
- $0.0530
DeepSeek V4 Flash Vision
deepseek-ai/DeepSeek-V4-Flash-Vision
977K contextTool callingJSON modeReasoningVision
- Input
- $0.5280
- Output
- $1.5840
- Cached input
- $0.0170
Tool callingJSON modeReasoningContext caching
- Input
- $1.3720
- Output
- $4.8000
- Cached input
- $0.3430
GLM-5.3 Flash
zai-org/GLM-5.3-Flash
977K contextTool callingJSON modeReasoningVision
- Input
- $0.1370
- Output
- $0.4800
- Cached input
- $0.0400
Tool callingJSON modeReasoningContext caching
- Input
- $0.6850
- Output
- $2.7430
- Cached input
- $0.1370
Kimi K2.7 Code
moonshotai/Kimi-K2.7-Code
256K contextTool callingJSON modeContext caching
- Input
- $1.1150
- Output
- $4.6290
- Cached input
- $0.2230
Kimi K2.6
moonshotai/Kimi-K2.6
256K contextTool callingJSON modeVisionContext caching
- Input
- $1.1150
- Output
- $4.6290
- Cached input
- $0.1880
Prices are in USD per million tokens. Billing is per request, metered to the token.