Frontier models

29 analyses · Latest

Whether frontier progress is “slowing” is the wrong question; the axis of competition is what keeps moving. These pieces track the shift from peak benchmark scores toward reliability, cost-performance, inference speed, and distribution. The model that wins is increasingly not the smartest one — it is the one that ships everywhere and holds up under real work.

2026-06-16 zhipu

GLM-5.2 Ships Its Weights: Open Models Have Made the Frontier a Quarterly Refresh

Zhipu released GLM-5.2 weights under MIT, with a 1M context, a long-horizon focus, and a tunable thinking budget. Its own benchmarks place it within a point or two of the closed frontier on long-horizon coding. The real signal is not another leaderboard run but the open-weight capability-cost curve dropping another notch. Treat the vendor numbers with a discount, and test the 1M usability and long-horizon reliability on your own tasks.

Read analysis