Operational
Cerebras
Ultra-fast inference on wafer-scale hardware.
Models
llama-3.3-70b
llama-3.1-8b
Capabilities
✓ Text✓ Streaming
Configuration
Environment variable: CEREBRAS_API_KEY=... Or configure via dashboard: Dashboard → Provider Keys → Add Key → cerebras