AI Models & LLM News: Releases, Benchmarks, Comparisons
New models, benchmarks and our tests
rio 3.5: a brazilian city hall just shipped a 397b open-weight model
kimi-k2.7-code vs minimax m3 on a frontend task
minimax m3 surges 198% – still priciest in a chinese-dominated top 4
openai is about to join the price war started by deepseek
anthropic's hle lead hits 16 points – and it's still accelerating
frontiercode: first bench to test if ai code is actually mergeable
two open agent models in one day: one codes, one lives
gemma 4 e2b qat: google squeezes an ai model to 0.84 gb
liquid ai just dropped a tiny model that reads images and pulls out exactly the data you need
minimax m3 vs kimi k2.6: when a 1-point gap is the whole story