OpenAI announced initial performance results of their own Jalapeño chip for AI inference.
1.5–1.9× more AI work per watt 1.7–3.6× lower end-to-end latency 2.1–4.1× higher performance on highly interactive workloads
Looking forward to seeing this deployed!
some investors have been asking me lately about the wave of inference chip startups. My answer: we need real performance data before forming a view. So here’s o...