Most people running LoRA fine-tunes in August 2026 are using the same config they copied from a tutorial in 2024.
Z.ai just shipped GLM-5.
Z.ai shipped GLM-5.
Nvidia just dropped $7 billion on Poolside, a coding AI startup most developers have never heard of. And the weird part?
Moonshot uploaded 1.56 terabytes of model weights to Hugging Face on July 27.
Zhipu quietly shipped something two months ago that should have gotten louder. GLM-5.
Forty-eight of Qwen 3.8's 64 layers have never heard of softmax attention.
Alibaba's flagship demo for Qwen Image 3.0 is a single image.
Alibaba dropped two Qwen3.8 models this week.
Stability AI spent four years refining the U-Net. SD 1.
NVIDIA dropped Nemotron 3.5 Lightning on August 11 with a premise that sounds backwards: your agent is too smart.
Nvidia shipped its first open-weight model this week, and the interesting part is what they didn't optimize for.
On August 1, DeepSeek's API processed 8 trillion tokens in a single day — 5 trillion from free-tier users alone.
Mark Zuckerberg published a 6,500-word essay on Sunday arguing that the biggest risk in AI isn't the technology itself — it's concentrating that...
Yesterday Meta Superintelligence Labs dropped Muse Glimmer, a 30-billion-parameter model purpose-built for agentic workloads, under Apache 2.0.
Last week a GitHub discussion thread casually dropped a result that deserves more attention: Qwen3.6-27B-FP8 reached 90.
Alibaba is about to release the weights for a 2.4-trillion-parameter model — the largest open-weight release from any lab, period.
Moonshot published Kimi K2.7 Code on June 12 with a SWE-bench Verified score of 60.
Every agent framework in 2026 has the same dirty secret: memory is a retrieval hack.