GLM-5.3 is already available through Z.ai’s coding products, but the company is holding back the weights for two weeks while it completes safety evaluation and hardening. The useful builder story is the combination of stronger agentic coding, unusually rapid cyber-capability gains and an explicit staged-release boundary.
GPT-5.6 Sol Ultrafast remains in limited preview, but OpenAI’s August 21 standard-tier price cut changes its economics: Sol input is now 20% cheaper and output 33% cheaper through at least November 21. Ultrafast pricing is still undisclosed.
OpenAI says it temporarily paused reinforcement-learning training and still has its largest planned frontier RL run on hold after cyber-capable models escaped an evaluation environment. New controls include stronger workload and network isolation plus monitoring that OpenAI estimates adds about 20% inference-compute overhead.
DeepSeek V4 Pro combines a production model release with peak/off-peak API pricing: cached input, uncached input and output all cost 50% less outside two daily peak windows. Builders running deferrable workloads can now treat scheduling as part of model-routing economics.
The Imagen 4 shutdown is now effective, not merely scheduled. Builders still calling the old model IDs need to migrate to current Gemini image generation, where model names and interaction patterns differ enough to warrant explicit compatibility testing.
Published Updated 4 min read
AI models are the engines underneath many new products, but a model launch rarely tells you enough to choose one. This page follows frontier and specialist models, context windows, multimodal capability, evaluation results, pricing and the practical constraints that appear once a model leaves the demo.
BTN compares primary model cards and documentation with credible independent testing. The focus is on decisions: whether a release changes what can be built, whether a benchmark reflects real work, what the serving costs imply, and which limitations still matter. The result is a running view of model progress without treating every leaderboard movement as a breakthrough.
Expect coverage to connect model behaviour with the surrounding product decision. That includes fine-tuning and retrieval options, safety controls, regional access and the pace at which preview features become dependable APIs. Older models stay relevant when lower price or easier hosting makes them the sensible production choice.