OpenAI announced initial performance results of their own Jalapeño chip for AI inference.
> 1.5–1.9× more AI work per watt
> 1.7–3.6× lower end-to-end latency
> 2.1–4.1× higher performance on highly interactive workloads
Post #9073
1.01K

- ❤ 7
- 👍 3