> Cerebras
Cerebras Inference is powered by wafer-scale WSE chips, offering world-leading inference speeds. It provides an OpenAI-compatible API with $5 in free credits and generous rate limits for developers. Details were checked against the provider's own page at www.cerebras.ai.
// pros
- First-party operator (not a relay), so no risk of stolen upstream keys; relatively reliable stability and compliance
- 1M tokens/day free quota is generous for its class and requires no credit card
- OpenAI-compatible endpoint; just change base_url to https://api.cerebras.ai/v1 to use existing SDKs
- Extremely fast inference (about 2,600 tokens/sec measured on Llama 4 Scout)
// cons
- Free-tier model list has narrowed over time; by mid-2026 the public free tier is mainly gpt-oss-120b and a few others, with many Llama/Qwen variants moved to paid Dedicated Endpoints
- Free tier has context-length and rate limits, unsuitable for large-context or high-concurrency production
- Essentially a free developer layer of a paid cloud; once quota is exhausted it becomes paid
Recorded as first-party-free: the free developer layer of a major vendor's paid cloud. Good for developers to claim 1M tokens/day for prototyping and light calls; note the free model list changes and production use should assess rate and context limits. Informational only, not an endorsement.
community_reputation
The community and several AI-credit directory sites generally rate the Cerebras free tier positively: large quota, fast speed, no credit card. As a chip maker it is a legitimate first-party vendor with no runaway or stolen-key discussions; the main gripe is the shrinking list of free models over time.
imported analysis — independent third party, not Uprouter editorial · verify before relying on it
How we estimated this: the free allowance is converted to a dollar value using the provider’s own pay-as-you-go rates for the models a typical developer would reach for. When the allowance is metered in requests, tokens or minutes rather than dollars, we assume median usage of the cheapest capable model. Because assumptions are involved, the figure carries the est. flag — treat it as an order of magnitude, not a quote.
note: Free dev tier: 1M tokens/day (daily reset, no card, ~30 RPM / 60K TPM; some report 5 RPM) → ≈$30/mo blended at $1/1M; beyond that paid PAYG (developer tier from $10).
| Plan | Type | Monthly | Input /1M | Output /1M | Note |
|---|---|---|---|---|---|
| Free tier (source-reported) | free | — | — | — | Free developer tier: 1 million tokens per day (resets daily, not a one-time credit), no credit card, rate limits around 30 RPM / 60K TPM, with the free-tier context once capped near 8K; register at cloud.cerebras.ai to generate an API key. Beyond that it becomes paid pay-as-you-go. |
| Free tier | free | — | Free | Free | 1M tokens/day, no card |
| Subscription | subscription | — | $0.850 | $1.20 | llama-3.3-70b; higher limits |
| GPT OSS 120B | payg | — | $0.350 | $0.750 | — |
Pricing normalized from public sources — always verify with the provider.
// model prices
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| gpt-oss-120b | — | — | 131k |
| Qwen3 32B | — | — | 131k |
| Llama 4 Scout | — | — | 1.3M |
| zai-glm-4.7 | — | — | 205k |
$ Does Cerebras have a free tier?
Yes — Cerebras offers a free tier that we value at roughly $30.00 in free credits (estimate — derived from rate-limited usage at pay-as-you-go prices). Always confirm current limits on Cerebras's official pricing page — free tiers change without notice.
$ How much does Cerebras cost per 1M tokens?
Cerebras does not publish simple per-1M-token pricing (it may bill per request, per output, or via a subscription plan). Check their official pricing page for exact figures.
$ Is Cerebras safe to use?
Uprouter rates Cerebras as low risk. Rated low risk. an official provider with a free tier with pricing, authentication and terms documented on its own site. The model coverage is consistent with a real upstream, so data and reliability risk are modest. Risk ratings are editorial, evidence-based, and never influenced by affiliate relationships.
$ Can I use Cerebras through Uprouter Connect?
Yes — Cerebras is Connect-compatible. Add your Cerebras API key in Uprouter Connect (encrypted at rest) and route requests to it through one OpenAI- and Claude-compatible API, optionally behind failover aliases. Uprouter meters the traffic in compute at the provider's listed rate.
Raw community sentiment, shown unblended. Votes never feed into objective data fields like price, risk, or status — and never into sort order.
No published reviews yet — be the first.
// more_informationprice history · live status · risk & confidence · code snippets
price_history
No recorded changes yet.
live_status
current: upNo probes recorded yet.
Uprouter runs its own probes; best-effort, not a guarantee. Probe results depend on our network location and cadence — for production, always use the provider’s own status page.
risk_&_confidence
Rated low risk. an official provider with a free tier with pricing, authentication and terms documented on its own site. The model coverage is consistent with a real upstream, so data and reliability risk are modest.
Confidence reflects how much verifiable evidence (official docs, probe history, community reports) backs this entry. Risk is our editorial assessment of operator reliability and terms — not financial advice.
code_snippets
curl https://uprouter.online/api/connect/v1/chat/completions \ -H "Authorization: Bearer upr_live_YOUR_KEY" \ -H "Content-Type: application/json" \ -d '{
"model": "gpt-oss-120b",
"messages": [{ "role": "user", "content": "Hello via Cerebras" }]
}'Metered in Uprouter compute — the usage log shows which upstream served each request and its exact cost.
Ranked by neutral similarity signals only — live status, Connect compatibility, free-tier shape and free-credit proximity. Commercial relationships never influence this list.