Tokenify

Models and API prices

5 models on one endpoint, priced per million tokens with no subscription and no minimum. Prices come from the live catalogue at build time; the models their vendors price by the clock carry an off-peak rate as well.

deepseek

DeepSeek V4.1 Flash

deepseek/deepseek-v4.1-flash
50% off

Fast multimodal model. Reads text, images and video; cheapest per token we sell.

Input
$0.300$0.150/M
$0.090 off-peak
Output
$1.200$0.600/M
$0.360 off-peak
Cache read
$0.006/M
$0.003 off-peak

Peak is Mon-Fri 09:00-12:00 and 14:00-18:00 (UTC+8); every other hour is off-peak, priced by when the request arrives.

DeepSeek V4.1 Flash API details →

DeepSeek V4 Pro

deepseek/deepseek-v4-pro
50% off

Frontier reasoning model. Best for agents, long-context analysis and code generation.

Input
$1.320$0.660/M
Output
$3.960$1.980/M
Cache read
$0.044/M
DeepSeek V4 Pro API details →

DeepSeek V4 Flash

deepseek/deepseek-v4-flash
50% off

High-efficiency workhorse. Classification, extraction, chat and high-volume pipelines.

Input
$0.440$0.220/M
Output
$1.320$0.660/M
Cache read
$0.014/M
DeepSeek V4 Flash API details →

glm

GLM 5.3 Flash

z-ai/glm-5.3-flash

Fast multimodal model with enforced structured output. Reads text, images and video.

Input
$0.150/M
Output
$0.500/M
Cache read
$0.030/M
GLM 5.3 Flash API details →

GLM 5.2

z-ai/glm-5.2
50% off

Flagship model for long-horizon agent tasks.

Input
$1.400$0.700/M
Output
$4.400$2.200/M
Cache read
$0.260/M
GLM 5.2 API details →

Where to go next

Last updated 2026-10-01. Prices come from the live catalogue at the time this page was built.