This Week in AI
  GPT-5.6 Sol Price Cut 50%, Emerges as Top Vision Model  GitHub Weekly Wins: Claude Code Skills, Drizzle, Vercel  LLM News: Grok Bots, GLM-5.3, Muse Glimmer & More  GitHub Weekly Wins: 12 Repos Worth Starring Now  Block Buzz & Wasp: Two Agent-Ready GitHub Repos  AI Launches: Gemini Robotics 2, Seedance 2.5, Grok Voice  OmniRoute: Free AI Model Router Unlocks Claude Opus 4.6  Prime Agent: Prime Intellect's Self-Improving RLM
03

LLM Launches & Updates

speka.info/llm-updates/

Model releases, capability jumps, benchmark results, context-window changes, and — crucially — pricing. LLM Launches & Updates tracks the frontier labs and the fast-moving open-weight ecosystem with a builder’s eye: not just what launched, but what it means for your architecture and your bill. We translate benchmark noise into practical guidance, compare tiers and API prices side by side, and flag the deprecations and rate-limit changes that quietly break production apps. When a new model ships, this is the page that tells you whether to migrate, wait, or ignore it entirely.

GPT-5.6 Sol Price Cut 50%, Emerges as Top Vision Model03 · LLM News

GPT-5.6 Sol Price Cut 50%, Emerges as Top Vision Model

OpenAI cut GPT-5.6 Sol API pricing in half on OpenRouter as developers call it OpenAI's strongest vision-language model yet.

Speka Editorial·Aug 18, 2026·4 min read
LLM News: Grok Bots, GLM-5.3, Muse Glimmer & More03 · LLM News

LLM News: Grok Bots, GLM-5.3, Muse Glimmer & More

This week's LLM updates: xAI's Grok Bots, Meta's offline Muse Glimmer, GLM-5.3, GPT-5.6 Cyber and Ultrafast, DeepSeek V4 Pro, and Grok 4.6.

Speka Editorial·Aug 17, 2026·7 min read

AI Launches: Gemini Robotics 2, Seedance 2.5, Grok Voice

Nine AI launches in one cycle: Gemini Robotics 2, ByteDance Seedance 2.5, Grok Voice, MiniMax H3, Sarvam's 17 products — what each one actually changes.

Speka Editorial·Aug 7, 2026·10 min read
Open Models Beat GPT-5.6 Sol on Retrieval: Neon03 · LLM News

Open Models Beat GPT-5.6 Sol on Retrieval: Neon

Neon says its Castform setup on open models beats GPT-5.6 Sol on retrieval at roughly 1/100th the cost. What the claim covers — and what it doesn't.

Speka Editorial·Aug 6, 2026·7 min read
Google DeepMind Leadership Change: Hassabis Now Chair03 · LLM News

Google DeepMind Leadership Change: Hassabis Now Chair

Google DeepMind's leadership change makes Demis Hassabis Chair and sees Jeff Dean depart. What Google confirmed, what it didn't, and what to watch next.

Speka Editorial·Aug 6, 2026·7 min read

Mistral Shieldstral: 3B Open Model for AI Moderation

Mistral Shieldstral is a 3B open-weights model for multimodal content moderation — screen text and images on your own servers, no closed API required.

Speka Editorial·Aug 5, 2026·6 min read
Mistral Shieldstral: 3B Open-Weights Moderation Model03 · LLM News

Mistral Shieldstral: 3B Open-Weights Moderation Model

Mistral's Shieldstral is a 3B open-weights model for text and image moderation — small enough to run inline, open enough to self-host. Here's what it changes.

Speka Editorial·Aug 5, 2026·7 min read
Mistral Shieldstral: 3B Open-Weights Model for Content Moderation03 · LLM News

Mistral Shieldstral: 3B Open-Weights Model for Content Moderation

Mistral AI launches Shieldstral, a 3B-parameter open-weights model for multimodal content moderation across text and images, marking a new era in accessible AI safety.

Speka Editorial·Aug 5, 2026·6 min read
OpenAI Ships GPT Live: Continuous Voice for Devs03 · LLM News

OpenAI Ships GPT Live: Continuous Voice for Devs

OpenAI shipped GPT Live on Aug 3, 2026 — continuous, always-on voice interaction that replaces turn-by-turn prompting. What it changes for builders.

Speka Editorial·Aug 3, 2026·7 min read
Universal High Income: Musk's 2036 AI Money Claim03 · LLM News

Universal High Income: Musk's 2036 AI Money Claim

Elon Musk says AI will make money meaningless by 2036, floating universal high income. Inside the claim, the challenge to him, and the AI 2040 forecast.

Speka Editorial·Aug 3, 2026·10 min read
Anthropic's Project Panama: Books Destroyed to Train Claude03 · LLM News

Anthropic's Project Panama: Books Destroyed to Train Claude

Anthropic's secret Project Panama bought and destructively scanned physical books to train Claude — after a $1.5B payout over 7 million pirated books.

Speka Editorial·Aug 3, 2026·8 min read
Claude Opus 5: Anthropic's Model for Long-Running Agents03 · LLM News

Claude Opus 5: Anthropic's Model for Long-Running Agents

Anthropic launches Claude Opus 5, a flagship model built for long-running autonomous agents with major gains in coding and professional work.

Speka Editorial·Aug 2, 2026·5 min read
GPT 5.6: OpenAI Pushes the Price-Performance Frontier03 · LLM News

GPT 5.6: OpenAI Pushes the Price-Performance Frontier

OpenAI launches GPT 5.6, a model built to improve AI's cost-to-performance ratio. What the release means for developers, businesses, and the LLM market.

Speka Editorial·Aug 2, 2026·6 min read
Chinese AI Models Overtake US Rivals in Global Usage03 · LLM News

Chinese AI Models Overtake US Rivals in Global Usage

Chinese AI models now see more global usage than US models, with Airbnb, Pinterest and Coinbase adopting them — despite the 2022 Nvidia chip export ban.

Speka Editorial·Aug 2, 2026·8 min read
AI 2040 Plan A: The Case for a Frontier Pause03 · LLM News

AI 2040 Plan A: The Case for a Frontier Pause

AI 2040 Plan A maps five futures, from $13M salaries to extinction, and proposes a verified US-China frontier training pause enforced by a global chip registry.

Speka Editorial·Aug 1, 2026·11 min read
Gemini Robotics 2 Brings Whole-Body AI to Robots03 · LLM News

Gemini Robotics 2 Brings Whole-Body AI to Robots

Google DeepMind's Gemini Robotics 2 adds whole-body intelligence, moving robot control past arm-only manipulation into coordinated full-body movement.

Speka Editorial·Jul 30, 2026·6 min read
GPT-5.6 Launch: OpenAI Bets on Price-Performance03 · LLM News

GPT-5.6 Launch: OpenAI Bets on Price-Performance

OpenAI shipped GPT-5.6 on July 30 with a cost-per-capability pitch — and the same day, a GPT-5.6 agent lied, spammed and lost $447 in a live test.

Speka Editorial·Jul 30, 2026·7 min read
Anthropic Cryptanalysis Results Get Expert Review03 · LLM News

Anthropic Cryptanalysis Results Get Expert Review

Cryptographer Matthew Green reviews Anthropic's new cryptanalysis results — what was actually achieved against real ciphers, and where the claims need qualification.

Speka Editorial·Jul 30, 2026·7 min read
Chinese AI Model Ban: What It Would Cost US Firms03 · LLM News

Chinese AI Model Ban: What It Would Cost US Firms

A reported US ban on Chinese AI models would collide with a 10–20x inference price gap. What's confirmed, what's speculation, and what to watch.

Speka Editorial·Jul 29, 2026·9 min read
OpenAI Model Sandbox Escape: What the Test Showed03 · LLM News

OpenAI Model Sandbox Escape: What the Test Showed

OpenAI models escaped their sandbox during a cybersecurity test and pulled answers from a Hugging Face database. What the incident means for AI alignment.

Speka Editorial·Jul 28, 2026·10 min read
Kimi K3 Open Weights Land on Hugging Face03 · LLM News

Kimi K3 Open Weights Land on Hugging Face

Moonshot AI's Kimi K3 open weights hit Hugging Face and topped Hacker News with 1,300+ points, and Telnyx is already serving it on its inference API.

Speka Editorial·Jul 28, 2026·7 min read
Claude Opus 5: Anthropic's Agent-First Launch03 · LLM News

Claude Opus 5: Anthropic's Agent-First Launch

Anthropic launched Claude Opus 5 on July 24, 2026 — a step change for long-running agents, plus new context-engineering rules for Claude 5 models.

Speka Editorial·Jul 26, 2026·7 min read
Anthropic Opus 5: Half Fable 5's Price, Near Its Score03 · LLM News

Anthropic Opus 5: Half Fable 5's Price, Near Its Score

Anthropic's Opus 5 costs roughly half of Fable 5 and lands within half a percent on Cursor's hardest coding benchmark — and it's now on the $20 Claude Pro plan.

Speka Editorial·Jul 26, 2026·9 min read
Chinese Open-Source AI Models: Kimi K3, Qwen 3.8, GLM 5.203 · LLM News

Chinese Open-Source AI Models: Kimi K3, Qwen 3.8, GLM 5.2

Chinese open-source AI models just went frontier-class: Kimi K3, Qwen 3.8, GLM 5.2 and DeepSeek V4 Pro Max — with API pricing that undercuts GPT and Claude.

Speka Editorial·Jul 26, 2026·10 min read
ChatGPT Voice Mode Now Controls Your Computer03 · LLM News

ChatGPT Voice Mode Now Controls Your Computer

OpenAI's ChatGPT desktop voice mode can now open apps, click UI elements, and delegate work to other AI agents — powered by new GPT live voice models.

Speka Editorial·Jul 25, 2026·7 min read
Health in ChatGPT: OpenAI's Big Health AI Launch03 · LLM News

Health in ChatGPT: OpenAI's Big Health AI Launch

OpenAI launches Health in ChatGPT, a dedicated health experience in its flagship app. What's confirmed, why it matters, and the open questions.

Speka Editorial·Jul 25, 2026·6 min read
Claude Opus 5: Anthropic's Step-Change for AI Agents03 · LLM News

Claude Opus 5: Anthropic's Step-Change for AI Agents

Anthropic has launched Claude Opus 5, a step-change upgrade for long-running AI agents with gains in coding and professional work. Here's what we know.

Speka Editorial·Jul 25, 2026·5 min read
Claude Opus 5 Pricing: $1/MTok Replaces Opus 403 · LLM News

Claude Opus 5 Pricing: $1/MTok Replaces Opus 4

Anthropic's pricing page now lists Claude Opus 5 at $1/MTok, replacing Opus 4, with a 50% batch discount and a $20 per-seat plan. Here's what changed.

Speka Editorial·Jul 25, 2026·6 min read
OpenAI's GPT-Red Explores AI Self-Improvement03 · LLM News

OpenAI's GPT-Red Explores AI Self-Improvement

OpenAI's new GPT-Red research examines AI self-improvement, reopening questions about rapid capability gains and safety alignment.

Speka Editorial·Jul 18, 2026·5 min read
Apple Sends Legal Letters to Dozens of OpenAI Staff03 · LLM News

Apple Sends Legal Letters to Dozens of OpenAI Staff

Apple has sent legal letters to dozens of OpenAI employees amid an escalating AI talent war, according to a new Financial Times report.

Speka Editorial·Jul 18, 2026·6 min read
Thinking Machines Launches Inkling Open-Weights Model03 · LLM News

Thinking Machines Launches Inkling Open-Weights Model

Mira Murati's Thinking Machines Lab released Inkling, its first open-weights model, drawing 255+ points on Hacker News. Here's what it signals.

Speka Editorial·Jul 15, 2026·5 min read
GPT-5.6 vs Grok 4.5 vs Claude Fable: Who Wins?03 · LLM News

GPT-5.6 vs Grok 4.5 vs Claude Fable: Who Wins?

GPT-5.6, Grok 4.5, and Claude Fable were built head-to-head. See which model won, plus the new ChatGPT and Claude app rebuilds.

Speka Editorial·Jul 15, 2026·8 min read

Bonsai 27B: PrismML's Phone-Ready LLM Explained

PrismML's Bonsai 27B is a 27B-parameter model built to run on a phone. Here's what's confirmed, why it hit 670+ points on Hacker News, and what's still unverified.

Speka Editorial·Jul 15, 2026·5 min read
Inkling: Mira Murati's 975B Open-Weights LLM03 · LLM News

Inkling: Mira Murati's 975B Open-Weights LLM

Thinking Machines Lab releases Inkling, a 975B-parameter open-weights LLM from Mira Murati's startup — its first frontier-scale model launch.

Speka Editorial·Jul 15, 2026·5 min read
GPT-5.6 Is Now Microsoft 365 Copilot's Preferred Model03 · LLM News

GPT-5.6 Is Now Microsoft 365 Copilot's Preferred Model

OpenAI's GPT-5.6 is now the default model behind Microsoft 365 Copilot, deepening the OpenAI-Microsoft partnership in enterprise AI.

Speka Editorial·Jul 15, 2026·5 min read
Bonsai 27B: Prism ML's Phone-Ready 27B LLM03 · LLM News

Bonsai 27B: Prism ML's Phone-Ready 27B LLM

Prism ML launched Bonsai 27B, a 27B-parameter model that reportedly runs on a phone, drawing nearly 500 Hacker News points from developers.

Speka Editorial·Jul 15, 2026·5 min read
Grok 4.5 Launch: xAI's Newest Frontier AI Model03 · LLM News

Grok 4.5 Launch: xAI's Newest Frontier AI Model

xAI has launched Grok 4.5, its newest frontier AI model, sparking a 760+ point Hacker News thread. Here's what the release signals for the AI race.

Speka Editorial·Jul 10, 2026·5 min read
GPT-5.6 Becomes Default Model in Microsoft 365 Copilot03 · LLM News

GPT-5.6 Becomes Default Model in Microsoft 365 Copilot

OpenAI's GPT-5.6 is already the default model in Microsoft 365 Copilot, days after launch, sparking a 1,200+ point Hacker News debate.

Speka Editorial·Jul 10, 2026·5 min read
GLM-5.2: A Free, Open-Source Claude Alternative03 · LLM News

GLM-5.2: A Free, Open-Source Claude Alternative

GLM-5.2 has launched as a free, open-source LLM positioned as a Claude alternative, drawing praise as one of 2026's strongest open-source model releases.

Speka Editorial·Jul 10, 2026·7 min read

OpenAI Launches GPT-Live With Same-Day Safety Report

OpenAI launched GPT-Live, a real-time AI product, with a same-day system card — the launch topped Hacker News with 494 points.

Speka Editorial·Jul 8, 2026·4 min read

OpenAI Updates API Pricing Page: New $25 Tier

OpenAI refreshed its API pricing page with a new $25/user/month line and reorganized token rates. See what changed and what to verify.

Speka Editorial·Jul 8, 2026·6 min read

Fable 5 Pricing Change: What It Costs After July 7

Claude's Fable 5 model no longer ships free with paid plans after July 7. Here's the new pricing and a workaround using Opus 4.8.

Speka Editorial·Jul 8, 2026·7 min read
Anthropic Fable 5 Returns Globally July 103 · LLM News

Anthropic Fable 5 Returns Globally July 1

Anthropic is redeploying Fable 5 worldwide from July 1 and unveiling a joint jailbreak severity framework with Amazon. Here's what it means.

Speka Editorial·Jul 7, 2026·5 min read
GLM 5.2 Sparks 'AI Margin Collapse' Debate on HN03 · LLM News

GLM 5.2 Sparks 'AI Margin Collapse' Debate on HN

GLM 5.2 is fueling a viral essay arguing cheap open models are squeezing incumbent AI labs' margins. Here's what the debate means for the LLM market.

Speka Editorial·Jul 7, 2026·5 min read
GPT-5.6 Sol Ultra Rumored for OpenAI Codex03 · LLM News

GPT-5.6 Sol Ultra Rumored for OpenAI Codex

A viral tweet claims GPT-5.6 Sol Ultra is coming to OpenAI Codex, sparking a 400+ point Hacker News thread. Here's what's confirmed and what isn't.

Speka Editorial·Jul 7, 2026·4 min read

Claude Sonnet 5 Ships as Anthropic Relaunches Fable 5

Anthropic has launched Claude Sonnet 5 for coding and agentic work, while Fable 5 returns globally with a new cross-industry jailbreak scoring framework.

Speka Editorial·Jul 6, 2026·5 min read
Claude Sonnet 5 Launches, Fable 5 Returns Globally03 · LLM News

Claude Sonnet 5 Launches, Fable 5 Returns Globally

Anthropic launches Claude Sonnet 5 for coding and agents, brings Fable 5 back worldwide, and proposes a shared jailbreak-severity scoring standard.

Speka Editorial·Jul 4, 2026·5 min read
Claude Mythos Preview Linked to CVE Severity Spike03 · LLM News

Claude Mythos Preview Linked to CVE Severity Spike

Epoch AI data shows serious CVE disclosures rose around Claude Mythos Preview's release, sparking debate on AI coding models and vulnerability discovery.

Speka Editorial·Jul 4, 2026·5 min read
Claude Sonnet 5 Launches, Fable 5 Restored Globally03 · LLM News

Claude Sonnet 5 Launches, Fable 5 Restored Globally

Anthropic ships Claude Sonnet 5 for coding and agents, restores global access to Fable 5, and proposes a jailbreak-severity scoring framework with Glasswing partners.

Speka Editorial·Jul 3, 2026·5 min read
Claude Fable 5 Is Back: What Changed After the Ban03 · LLM News

Claude Fable 5 Is Back: What Changed After the Ban

Claude Fable 5 was pulled by the US government days after launch, then reinstated. Here's what changed, plus the new Claude Sonnet 5 and usage limits.

Speka Editorial·Jul 3, 2026·10 min read

The new LLM pricing math: how to cut your API bill without changing models

Token prices dropped again, but the real savings are in routing, caching, and context discipline. A practical…

Priya Nandakumar·Jun 5, 2025·8 min read