DeepSeek V4 Flash
$0.00894% offper 1M tokens$0.1494% off
An efficiency-tuned mixture-of-experts model from DeepSeek, with 13B of 284B parameters active. It is built for fast inference.
1.05MContext
$0.008Input / 1M
$0.016Output / 1M
94%Saved vs retail
client.chat.completions.create( model="deepseek-v4-flash", messages=[ {"role": "user", "content": "Hello"}, ],)
Cost per 1M input tokens
Claw Hunter$0.008
Retail$0.14