Cerebras API Pricing

Llama and Qwen models on dedicated wafer-scale silicon, offering ultra-low latency inference.

Prices are per 1 million tokens and updated daily from the official Cerebras pricing page. View full dashboard →