Skip to content

Pricing

CompactifAI offers simple, transparent pay-as-you-go pricing for our API services.

Our pricing is based on tokens processed through the API. We charge separately for:

  • Input tokens: Text sent to the API
  • Output tokens: Text generated by the API

A token represents approximately 4 characters or 0.75 words in English.

ModelInput (per 1M tokens)Output (per 1M tokens)
cai-mistral-small-3-1-slim$0.05$0.08
mistral-small-3-1$0.11$0.17
nemotron-3-nano-omni$0.20$0.80
gpt-oss-120b$0.05$0.23
hypernova-60b$0.04$0.14
carina-60b$0.05$0.20
qwen-3-6-27b$0.15$0.90
glm-5-1$0.95$3.15
cai-glm-5-1$1.00$3.00
glm-5-2$1.10$3.50
quasar-438b$0.60$1.80
ModelPrice (per audio minute)Billing Granularity
cai-whisper-large-v3-turbo-slim$0,000134Per-second billing with 1-minute minimum
  • Pay only for what you use
  • No monthly commitments or minimum fees
  • Billing based on actual token usage
Usage: 5M input tokens + 2M output tokens with hypernova-60b
Cost calculation:
- 5M input tokens × $0.04 = $0.20
- 2M output tokens × $0.14 = $0.28
Total cost: $0.48
  • Use the CompactifAI Dashboard to manage API tokens, monitor usage, view account settings, and review billing.

How do I estimate my token usage? A token represents approximately 4 characters or 0.75 words in English. For more specific estimations, contact our support team.

How often will I be billed? Billing occurs monthly based on your actual usage.

Is there a minimum spend requirement? No, you only pay for what you use.

For questions about pricing or to request a custom enterprise plan, please contact our sales team.