Forty-eight of Qwen 3.8's 64 layers have never heard of softmax attention.
Alibaba dropped two Qwen3.8 models this week.
Stability AI spent four years refining the U-Net. SD 1.
Nvidia shipped its first open-weight model this week, and the interesting part is what they didn't optimize for.
Two days ago Meta Superintelligence Labs dropped Muse Glimmer, a 30B dense model distilled from their larger Muse Spark system.
Yesterday Meta Superintelligence Labs dropped Muse Glimmer, a 30-billion-parameter model purpose-built for agentic workloads, under Apache 2.0.
Last week a GitHub discussion thread casually dropped a result that deserves more attention: Qwen3.6-27B-FP8 reached 90.
Alibaba is about to release the weights for a 2.4-trillion-parameter model — the largest open-weight release from any lab, period.
I deleted 47 LoRAs last week. Character models I'd trained on curated datasets of 30–50 images each, some taking three hours on an A100.
Every major AI music generator is in court right now — or waiting for the phone to ring. Suno settled with UMG in March.
Every week someone on r/LocalLLaMA posts the same question: "I quantized my 70B model to Q4_K_M, it fits in VRAM, but when I set the context to 128K it...
Stability AI spent four years iterating on the same convolutional backbone. Then, on April 6, they killed it.
Four months ago you typed ollama run qwen3 and got a chat window.
GDDR6 spot prices tripled since autumn 2025. From roughly 2.
Ollama 0.31 dropped this week with a single headline number: Gemma 4 generates tokens nearly 90% faster on Apple Silicon.
Alibaba's Qwen team has a running habit of shipping models that make you question your assumptions about parameter counts.
Alibaba's own 397B MoE flagship just got outscored by a model thirteen times smaller from the same team. Qwen 3.
Poolside raised $626 million, hired a team of ex-DeepMind and ex-Meta researchers, and then went quiet for nearly three years.
You need a training video dubbed into Tagalog by Friday. ElevenLabs supports Tagalog — barely.
Every time a new model drops, the ritual plays out the same way.