A new foundation model — the brain.
Shifts what's possible. New ceiling on reasoning, coding, context.
latestClaude Opus 5Anthropic
What just shipped in AI that's actually worth knowing about.
A new foundation model — the brain.
Shifts what's possible. New ceiling on reasoning, coding, context.
latestClaude Opus 5Anthropic
A consumer app or feature.
How non-builders actually feel AI in their day.
watchingnone in feed yet — when one ships, it lands here
A developer API, SDK, or capability.
What you can build with this week — the new primitives.
latestBuzzBlock
Hardware, platforms, the backbone.
Who runs what at what price — the rails everyone shares.
watchingnone in feed yet — when one ships, it lands here
Most launches in this feed come from the same handful of labs. Here's who they are, where they're based, and what they're actually known for — so the org chips below stop being acronym soup.
Set the pace for the modern wave — ChatGPT made conversational AI a household tool overnight.
GPT-5 · ChatGPT · Sora · Realtime API
Founded by ex-OpenAI safety researchers; Claude is the model coding tools quietly standardize on.
Claude Sonnet · Claude Opus · MCP · Skills
DeepMind + Brain merged into the Gemini team; the only lab with the chips, the data, and the products.
Gemini 3 · AI Studio · Vertex · Veo
Open-weights leader in the West — Llama set the floor every other open lab is measured against.
Llama · Movie Gen · Ray-Ban Display
Musk's lab. Massive Memphis training cluster, edgier brand voice, integrated tightly with X.
Grok · Colossus supercluster
Europe's flagship lab — punches above its weight on small, efficient, mostly open-weights models.
Mistral Large · Mixtral · Codestral
Sparked the 2025 cost-collapse — frontier-class reasoning at a fraction of the price, weights and all.
DeepSeek-R1 · V3 · open weights
Qwen team ships open-weights models at every size with ruthless cadence — the workhorse of Chinese AI.
Qwen 3 · Qwen-VL · Qwen-Coder
Sells the picks and shovels. Every other lab on this list trains on its chips — H100, B200, GB200, and counting.
Blackwell · CUDA · NeMo · Nemotron
Plenty of others ship too — Cursor, GitHub, Cloudflare, Midjourney, Hugging Face, Runway, ByteDance, MiniMax, Z.ai. They show up in the chips below as their launches land.
Most launches are real. Some are framing. These are the six tells that separate "actually ships" from "press release with a roadmap" — handy when a headline beats the feed below to your inbox.
watch forCharts that only show the benchmarks they win. Comparisons against last year's model, not this week's.
what it meansReal on a narrow slice — often weaker on the harder evals (GPQA, ARC-AGI, Humanity's Last Exam) that didn't make the slide.
watch for"Rolling out over the coming weeks." Waitlists. A blog post but no API page, no pricing, no model ID.
what it meansMarketing landed today; the product might land later. Treat the date as when it became real to journalists, not to you.
watch forHeadlines built on "10× cheaper" or "3× faster" with no matching jump on capability evals.
what it meansBig win if you're spending on inference. Not a new ceiling — same model territory at a better price.
watch forHighlight reels, narrator voiceover, "hand-picked examples," only logged-out demos.
what it meansStage version ≠ API version. The cherry-picked scene is the ceiling, not the median run you'll get.
watch for"Llama community license," commercial caps ("700M monthly users"), custom acceptable-use clauses.
what it meansFree to download. Not free to ship. Read the license before betting a product on a model.
watch forPerformance claims with no methodology. "Internal evals" footnotes. No system card on launch day.
what it meansIf they didn't publish the eval setup, the numbers are vibes. Wait a week for independent runs (Artificial Analysis, LMSYS).
None of these mean a launch is fake — just that the spin is doing work. The feed below tries to flag the spin in the summary when it spots it.
Anthropic — Get near-Fable-5 quality at half the price, now the default on Claude Pro and Max. A low/medium/high effort toggle lets you dial cost against capability per task, with sharper autonomous coding. Live in the API today.
read more →Poolside — Download a free, open-weight coding model that punches above rivals 10x its size on SWE-Bench. It's a 118B mixture-of-experts activating just 8B per token, with a 1M-token context. Grab it on Hugging Face or run it via OpenRouter.
Google — Code and run agents on a cheaper workhorse model — $1.50 per million input tokens, output priced well under the old Flash tier, and 17% fewer output tokens per job. Live in the API and GitHub Copilot now.
Block — Run a free, open-source team chat that treats AI agents as real members with their own identity and permissions. Includes code hosting and automations, and plugs into Claude Code, Codex, or goose. Mac, Windows, Linux.
A hand-curated log of AI launches that actually move the field — new flagship models, agentic products, developer APIs, and infrastructure shifts. No press releases. No minor patches. No prerelease beta noise.