Crazy: Artificial Analysis currently ranks Celeris-1 #1 for output speed at roughly 2,086 tokens per second @ 75.9% on MMLU-Pro!
Celeris says it achieves this on off-the-shelf GPUs, without custom inference silicon.
In Celeris' own same-harness evaluation, the model scored 75.9% on MMLU-Pro, compared with 78.5% for GPT-5 mini and 81.9% for GPT-5, while responding 13-16× faster.