AI Models & LLM News: Releases, Benchmarks, Comparisons
New models, benchmarks and our tests
qwen drops 3.7 plus – the model that reproduces apps on its own
5 releases, +0.3%: open-weight coding hit a wall
opus 4.8 beats mythos on agentic computer use
how tokenspeed made 397b qwen3.5 run at 580 tps?
xiaomi follows deepseek's playbook: mimo-v2.5-pro api now matches deepseek-v4-pro pricing to the cent
minimax m3: sparse attention hits 15.6x decoding speedup at 1m tokens
grok foundation model v9-medium training completed – looks like it's grok 4.5
deepseek decision makes it the cheapest among the smart models
5 things buried in the qwen 3.7 release that nobody's talking about
Gemini 3.5 Flash: 284 tok/s and IQ 55 — Fastest Smart Model