Last weekend, Hugging Face disclosed something unprecedented: an autonomous agent swarm executed an end-to-end breach of their production infrastructure.
Mira Murati's Thinking Machines Lab dropped Inkling last week with a line you almost never hear from a model release: "Inkling is not the strongest...
Hugging Face disclosed a breach this week.
For two months, a model called "Owl Alpha" sat near the top of OpenRouter's leaderboards. First place on Hermes Agent by call volume.
Moonshot's latest model just took the #1 spot on LMArena's Frontend Code benchmark, beating Claude Fable 5 in blind developer testing. It clocked 93.
Last Tuesday I needed to put the same matte black speaker on twelve different surfaces — marble counter, oak desk, concrete ledge, backyard grass.
Every large language model you use right now — GPT-5.6, Claude, Gemini, Llama — generates text the same way: one token, then the next, then the next.
Moonshot AI dropped K2.7 Code on June 12 and it landed with a strange flex: on Kimi Code Bench v2, it scores 62.
Twelve months ago, Chinese-made language models handled less than 2% of the tokens flowing through OpenRouter from US-based developer accounts.
Imagine shipping a feature that depends on a model you can see in benchmarks but can't call from your API key.
Krea dropped the weights for their 12-billion-parameter image model on June 22. Within hours, quantized variants and ComfyUI nodes were live.
Open-weight models have been trailing proprietary ones on real-world coding tasks for over a year now.
Ideogram just dropped a 9.3-billion-parameter image model that outperforms rivals three to eight times its size on text rendering benchmarks.
Two days ago MiniMax dropped a launch announcement for M3 that reads like a greatest-hits compilation of everything the open-weight community has been asking...
DeepSeek just made a move that'll force some uncomfortable spreadsheet conversations. The 75% promotional discount on V4 Pro?
Stability AI dropped four audio models this week and none of them will generate a human voice. No lyrics, no vocal harmonies, no Auto-Tuned hooks.
Every mainstream diffusion model follows the same three-part recipe: a text encoder tokenizes your prompt, a diffusion backbone denoises in latent space, and a...
Two days ago, Cursor shipped Composer 2.5.
Mistral just did something no other open-weight lab has pulled off: they shipped a 128B model that scores 77.