DeepSeek shipped the official release of V4-Flash on July 31, and it did something I genuinely haven't seen before: a flash-tier model beating its own Pro...
Every video model shipped this year has chased the same upgrade: longer duration, higher resolution, smoother motion. MiniMax went somewhere else entirely.
Every major AI music generator is in court right now — or waiting for the phone to ring. Suno settled with UMG in March.
Moonshot AI released the open weights for Kimi K3 on July 27 — a 2.
A security researcher pointed Kimi K3 at a Redis 8.8.
DeepSWE v1.1 is the benchmark nobody talks about because almost nobody passes it.
NVIDIA's research team took Nemotron-3-Nano, a 30-billion-parameter hybrid model, literally cloned it into two halves, froze one, trained the other to...
Stability AI spent four years iterating on the same convolutional backbone. Then, on April 6, they killed it.
Moonshot AI's Kimi K3 weights went live at midnight UTC — 594 gigabytes of MXFP4 safetensors, the largest open-weight release in AI history.
Black Forest Labs was the image generation company. FLUX 1 gave developers an open-weight alternative to Midjourney.
Moonshot AI just released the largest open-weight model anyone has ever published — and then immediately told everyone it's worse than the competition in...
Alibaba announced Qwen 3.8-Max on July 19 at the World AI Conference in Shanghai.
Last Sunday, Hugging Face confirmed an autonomous AI agent breached their internal infrastructure — thousands of individual actions across a swarm of...
Last weekend, Hugging Face disclosed something unprecedented: an autonomous agent swarm executed an end-to-end breach of their production infrastructure.
Mira Murati's Thinking Machines Lab dropped Inkling last week with a line you almost never hear from a model release: "Inkling is not the strongest...
Hugging Face disclosed a breach this week.
For two months, a model called "Owl Alpha" sat near the top of OpenRouter's leaderboards. First place on Hermes Agent by call volume.
Moonshot's latest model just took the #1 spot on LMArena's Frontend Code benchmark, beating Claude Fable 5 in blind developer testing. It clocked 93.
Last Tuesday I needed to put the same matte black speaker on twelve different surfaces — marble counter, oak desk, concrete ledge, backyard grass.
Every large language model you use right now — GPT-5.6, Claude, Gemini, Llama — generates text the same way: one token, then the next, then the next.