engyLive Kimi K3 and DeepSeek V4 Flash are now live on engy.ai!

Pricing

Per-token, pay as you go. No subscriptions and no minimums; you pay only for the tokens you use.

modelinput $/1Moutput $/1Mcached input $/1M
deepseek-v4-flash-0731new$0.045$0.09$0.009
kimi-k3new$1.50$7.50$0.15
glm-5.2$0.68$1.50$0.18
qwen3.6-35b-a3b$0.045$0.30$0.015

Prompt-cache hits bill at the cached rate automatically, with no config and no cache_control markers. Agentic workloads (coding assistants, multi-turn tools) typically hit 90%+ cache on repeated prefixes, so effective input cost is usually far below the headline rate.

Prices are live from the billing engine. The API reports the same numbers at https://api.engy.ai/v1/models.