GLM-5.3
z.aitool
Z.ai's next model reuses the GLM-5.2 base and puts every gain into post-training — weights held back roughly two weeks this time, after training produced cybersecurity capabilities the team didn't plan for.
z.aitool
Z.ai's next model reuses the GLM-5.2 base and puts every gain into post-training — weights held back roughly two weeks this time, after training produced cybersecurity capabilities the team didn't plan for.
blog.googletool
A second Flash release just three weeks after 3.6 — same base model, algorithmic improvements only, and double-digit gains on coding and agent benchmarks.
seed.bytedance.comtool
ByteDance's low-latency tier of Seed 2.1, on Volcano Engine — pitched as beating Claude Opus 4.6 at up to 80% lower cost for high-frequency agent workloads.
huggingface.cotool
DeepSeek retrains its flash tier of V4 for coding, agents and tool use — 284B total parameters, 13B active, and a 1M-token context window.
help.aliyun.comtool
Alibaba's cheap, fast vision-language tier of Qwen3.7 — $0.03/M input tokens, a 991K context window, and no accompanying technical report.
huggingface.cotool
Moonshot open-sources a 2.8-trillion-parameter, 1M-context model — billed as the largest open-weight model released so far, with only 104B parameters active per token.
www.anthropic.comtool
Anthropic's flagship rolls out across providers — close to Fable 5's frontier intelligence at half the price, and the new state of the art on coding and knowledge-work evals.
blog.googletool
Google ships three models in one day — Gemini 3.6 Flash, 3.5 Flash-Lite, and a cybersecurity-tuned 3.5 Flash Cyber — with Gemini 3.5 Pro still nowhere in sight.
www.anthropic.comtool
Anthropic's mid-tier model closes most of the gap to Opus while keeping Sonnet's speed and pricing — $2/$10 per million tokens, and the new default for everyday agentic work.