The more I use frontier models, the less I believe in one model to rule them all.
• Opus 5: best at teaching me with generated html, but verbose and sloppy. • GPT-5.6 Sol: great backend, weak frontend. • Kimi K3 / GLM-5.2: cheap and efficient for 90% of day-to-day coding, but weaker at AI research tasks like writing kernels.
Every lab is training for AGI. But today’s LLMs are still specialists.