Cerebras Inference

Wafer-scale speed for open models.

Inference API running on wafer-scale engines, offering thousands of tokens per second on popular open models.

Vendor
Cerebras
Category
Inference & model hosting
Pricing
Freemium
Open source
No
Platforms
API
Launched
2024
Website
cerebras.ai

Features

Best for

More inference & model hosting

Links