Top 3 Fastest AI Models for Text Generation

While specific rankings vary by benchmark and task, Celeris-1 has reported impressive output speeds of 1,644 tokens/second, and Gemini 3.6 Flash achieves 238 tokens/second with a high quality score. For lowest latency to first answer, LiquidAI's LFM2-24B-A2B recorded 0.42 seconds.

Sources

Open the full topic