Technology1h ago
Cerebras Doubles CS-4 Inference Throughput Without New Chip
The chip maker extracts twice the token-per-second performance from its existing WSE-3 wafer through clock speed increases and architectural improvements, letting customers double AI inference revenue on flat hardware budgets.