Pricing
Per-token, pay as you go. No subscriptions and no minimums; you pay only for the tokens you use.
| model | modalities | input $/1M | output $/1M | cached input $/1M |
|---|---|---|---|---|
| qwen3.8-27bnew | text, imagetext | $0.045 | $0.32 | $0.015 |
| deepseek-v4-flash-0731new | texttext | $0.045 | $0.09 | $0.009 |
| glm-5.2 | texttext | $0.68 | $1.50 | $0.18 |
| kimi-k3 | texttext | $1.50 | $7.50 | $0.15 |
| qwen3.6-35b-a3b | texttext | $0.045 | $0.30 | $0.015 |
Prompt-cache hits bill at the cached rate automatically, with no config and no cache_control markers. Agentic workloads (coding assistants, multi-turn tools) typically hit 90%+ cache on repeated prefixes, so effective input cost is usually far below the headline rate.
Prices are live from the billing engine. The API reports the same numbers at https://api.engy.ai/v1/models.