Mira Murati's Thinking Machines Lab dropped Inkling last week with a line you almost never hear from a model release: "Inkling is not the strongest...
Mistral didn't set out to build a bug finder. They built Leanstral 1.
For two months, a model called "Owl Alpha" sat near the top of OpenRouter's leaderboards. First place on Hermes Agent by call volume.
Moonshot's latest model just took the #1 spot on LMArena's Frontend Code benchmark, beating Claude Fable 5 in blind developer testing. It clocked 93.
The largest open-weight model in history just shipped, and the number that matters most isn't 2.8 trillion.
Mira Murati left OpenAI, raised $2 billion at seed, and the first thing her lab ships is a model that openly admits it isn't the best at everything.
Open your VS Code model picker today and you'll see a new name sitting next to Claude Sonnet 5 and GPT-5.6 Terra: Kimi K2.
Tencent dropped Hy3 on July 6 with a claim that sounded like a typo: a 295B MoE model with 21B active parameters beating DeepSeek V4 Pro — a 1.
Moonshot AI dropped K2.7 Code on June 12 and it landed with a strange flex: on Kimi Code Bench v2, it scores 62.
SpaceXAI's newest model landed on Tuesday, and the discourse immediately split into two camps that aren't talking to each other.
Alibaba's Qwen team has a running habit of shipping models that make you question your assumptions about parameter counts.
For two months, a model called "Owl Alpha" quietly climbed the OpenRouter charts. No launch event, no press tour, no Twitter hype cycle.
Two days ago at Build, Microsoft did something it's never done before: announced a full family of homegrown AI models that compete directly with the...
Two days ago MiniMax dropped a launch announcement for M3 that reads like a greatest-hits compilation of everything the open-weight community has been asking...
The biggest announcement from Microsoft Build 2026 wasn't a new Azure service or another Windows agent feature. It was a quiet decoupling.
Poolside raised $626 million, hired a team of ex-DeepMind and ex-Meta researchers, and then went quiet for nearly three years.
Zyphra just dropped a model that makes me question everything I assumed about parameter efficiency. ZAYA1-8B is an 8.
Every large language model you've touched — GPT, Claude, Llama, Qwen, Gemma — generates text the same fundamental way. Predict the next token.
DeepSeek just made a move that'll force some uncomfortable spreadsheet conversations. The 75% promotional discount on V4 Pro?