the week in AI, briefly. then briefly again.
Artificial Intelligence — briefly, then briefly again · tbb.ceo
listen tothe weekly #005 0:00 –:––
#005 ·19 JUN 2026 ·FRIDAY ·6 MIN READ ·10 STORIES + 18 EXTRAS

Cold Feet, Hot Models

No. 05 for June 13-19, 2026: Anthropic hits undo on its own billing change, OpenAI files for a trillion-dollar IPO while GPT-5.6 slips out the side door, Gemini 3.5 Pro idles on the runway, and Grok learns to draw -- the week AI couldn't decide whether to ship or stall.

01 / The Ten

The week, ranked

10

Anthropic yanks its own billing overhaul on launch day

Anthropic paused its planned June 15 Claude Agent SDK / claude -p billing change on the very day it was due to take effect, telling developers "nothing changes for now." The scrapped plan would have moved Agent SDK and third-party-app usage off subscription pools onto tiered $20-$200 monthly credits plus API billing.

Why it mattersA rare public retreat: with OpenAI reportedly weighing steep API cuts and an IPO looming, there's no room to raise effective costs on your own developer base mid-price-war.

OpenAI files confidentially for a $1T+ IPO as GPT-5.6 ships

OpenAI filed confidentially with the SEC on June 8 for an IPO targeting a $1T+ valuation, Goldman Sachs and Morgan Stanley leading, with a possible December 2026 listing. In parallel it is rolling out GPT-5.6 this month, a "meaningful improvement" that cuts token usage per task another 10-15% over GPT-5.5.

Why it mattersAnnualized revenue has blown past $25B; the filing turns the AI-bubble debate into a public-market referendum, and cheaper-per-task models keep squeezing rivals on price.

Gemini 3.5 Pro on the runway: 2M-token context, Deep Think

Google's Gemini 3.5 Pro, unveiled at I/O on May 19 with a 2M-token context window and a "Deep Think" reasoning mode, is in internal and limited enterprise testing as of mid-June, with general availability slated for "June." Prediction markets put "released by June 30" at ~50-55%.

Why it mattersPro is the model Google needs to answer GPT-5.6 and Claude's Mythos line; a 2M-token window reframes how much codebase or corpus an agent can hold at once.

Anthropic overhauls Claude Design with code sync and brand controls

On June 17 Anthropic shipped a major Claude Design update adding design-system imports, tighter Claude Code integration, direct canvas editing, and more export formats. The beta is available to Pro, Max, Team, and Enterprise customers.

Why it mattersIt pushes Claude further into the designer-to-code pipeline, exactly where Figma-style tools and coding agents are now colliding.

Grok Imagine 1.5 lands with a faster image-to-video sibling

xAI logged Grok Imagine 1.5 Preview in its release notes on June 11, built on the Aurora-2 engine, alongside Grok Imagine Video 1.5, an image-to-video model now in wide release that renders 720p clips in ~25 seconds with sharper physics.

Why it mattersxAI is racing to match Veo and Kling on generative media while piping Grok Imagine straight into X's billion-user feed.

xAI's next Grok finishes training at 1.5 trillion parameters

xAI's new model (reported as V9-Medium) completed training at 1.5T parameters, triple its current production model, with a coding focus and a mid-June release target. Flagship Grok 5 is still training on Colossus 2, with Polymarket pricing a pre-June-30 launch at just 12-33%.

Why it mattersMusk is betting raw scale plus a terminal-native coding agent can leapfrog Claude and Codex; the gap between "training done" and "shipped" is where xAI keeps slipping.

Kling v3 tops the video leaderboard as Seedance 2.0 steals buzz

Kling v3 leads the text-to-video arena at 2031, ahead of LTX-2 Fast (1920) and Seedance 2.0 Fast (1851). Seedance 2.0 is the model creators keep naming in blind tests and image-to-video work, while OpenAI has discontinued Sora 2 for new projects.

Why it mattersThe video-gen crown changes hands monthly now; OpenAI quietly retiring Sora 2 shows even frontier labs can't keep pace with open and Chinese-lab challengers.

NotebookLM goes agentic, builds your source library for you

Google's June 8 NotebookLM update moves to Gemini 3.5 as default and adds agentic research that suggests and gathers sources via Google Search, in-notebook code execution, and editable PDF/PowerPoint exports, rolling out to AI Ultra and Workspace business users.

Why it mattersNotebookLM is morphing from a study aid into a research agent, a template for how "chat with your docs" tools absorb autonomous web research.

AI matches or beats ER doctors on diagnosis and case management

A Harvard-led study in Science found OpenAI reasoning models conducting real-world ER triage, ordering appropriate tests, and managing cases at or above well-trained physicians; one system (MIRA) hit 87.8% diagnostic accuracy versus 78.1% for doctors.

Why it mattersResearchers say clinical reasoning is "approaching the ceiling," but caution that matching a benchmark isn't the same as replacing context, uncertainty, and patient judgment.

EU AI Act's high-risk deadline looms August 2 amid a delay fight

The EU's "Digital Omnibus," the first amendments to the AI Act since 2024, is nearing final adoption, but the August 2, 2026 compliance deadline for high-risk systems stands, with a contested push to delay it to December 2027 unresolved. EU regulators issued 50 fines totalling EUR250M in Q1 2026, mostly for GPAI non-compliance.

Why it mattersCompanies face a hard date under shifting rules; how Brussels handles the delay signals whether AI regulation bends to industry pressure.
02 / Also

Worth knowing

18
Copilot goes usage-based
GitHub Copilot switched to usage-based billing with AI Credits on June 1, adding a Copilot Max upgrade tier on top of existing subscriptions.
developersdigest.tech ↗
Fable 5 hits Copilot
Anthropic's Mythos-class Claude Fable 5 (~$10/$50 per Mtok) reached general availability inside GitHub Copilot Pro+, Max, Business, and Enterprise on June 9.
developersdigest.tech ↗
Gemini 3.1 Flash-Lite
Google shipped Gemini 3.1 Flash-Lite at $0.25 per million input tokens with 2.5x faster responses and 45% faster output than prior Gemini versions.
codersera.com ↗
Microsoft's own models
Microsoft unveiled MAI-Code-1-Flash and a reasoning model, MAI-Thinking-1, to cut its reliance on OpenAI and lower developer costs.
cnbc.com ↗
Grok 4.3 on Bedrock
xAI's Grok 4.3 reached general availability on Amazon Bedrock with a 1M-token context, configurable reasoning, and the lowest hallucination rate among frontier models.
releasebot.io ↗
Grok Build opens up
xAI's terminal-native coding agent Grok Build expanded to all SuperGrok and X Premium+ subscribers, with grok-build-0.1 priced at $0.20/$1.50 per Mtok.
eweek.com ↗
AlphaSense raises $350M
Market-intelligence platform AlphaSense closed a $350M round at a $7.5B valuation after topping $600M ARR in Q1.
techstartups.com ↗
Oxford Quantum's Series C
Oxford Quantum Circuits raised GBP260M in a Series C round in early June.
techstartups.com ↗
JPMorgan makes AI core
JPMorgan reclassified its AI spend from experimental R&D to core infrastructure, backing a ~$19.8B 2026 tech budget and 2,000 dedicated AI staff.
blog.mean.ceo ↗
The revenue race
OpenAI has pushed past $25B in annualized revenue while Anthropic approaches $19B.
sacra.com ↗
MCP becomes plumbing
The Model Context Protocol crossed 10,000+ servers and 97M monthly SDK downloads, cementing it as the default standard for connecting agents to tools.
firecrawl.dev ↗
Q1 funding record
Q1 2026 venture funding hit ~$300B with AI taking 80%; OpenAI, Anthropic, xAI and Waymo alone raised $188B.
news.crunchbase.com ↗
GPT Image 2 leads
OpenAI's GPT Image 2 topped a 2026 image-gen comparison at 9.6/10 for realism, prompt fidelity, and text handling.
diyai.io ↗
Claude Code usage attribution
Claude Code added per-skill/agent/plugin/MCP usage breakdowns and managed-model allowlists for enterprise admins.
releasebot.io ↗
Compiling agents
A "compiling agentic workflows" technique distills planner/researcher/writer/reviewer pipelines into a single fine-tuned forward pass.
requesty.ai ↗
Agent teams over solos
Coordinated teams of specialized agents working in parallel are displacing single-agent loops to beat context-window limits.
firecrawl.dev ↗
Enterprise-managed MCP
Anthropic added enterprise-managed MCP connector access starting with Okta, with zero-touch provisioning across Claude chat, Code, and Cowork.
releasebot.io ↗
New open weights
Qwen3.7 Plus and MiniMax M3 landed this stretch, keeping the open-weight frontier a largely Chinese-labs story.
llm-stats.com ↗
Be subscriber #012 one email a week · no spam · unsubscribe anytime
Back issues

The Archive

Every Friday · 10:00

Get the brief

One email a week. The ten things in AI that mattered, and why. Choose your channel.