Happy Sunday, !
Welcome back to your weekly skimmable AI news roundup.
In case you missed it, here’s this week’s Thursday post:
Don’t Let AI Just Guess What You Said
TL;DR Voice dictation can get things wrong, and AI doesn’t challenge its mistakes. This instruction makes AI stop and check first.
Note: If you’re consistently missing out on my emails, remember to check your “Promotions” tab and mark whytryai@substack.com as a “Safe Sender.”
Here’s what happened in AI last week:
👩💻 AI releases
Anthropic released Claude Fable 5.1, its best model for coding, knowledge work, and research, at roughly 25% lower token cost than Fable 5.
fal launched fal.live, a platform for interactive AI livestreams where viewers direct what happens next via text prompts in real time.
Google news:
Gemini 3.8 Flash takes extra reasoning steps on complex tasks and is especially well-suited for autonomous work and long-horizon coding.
Gemini Spark has come to Google Photos, letting users search, edit, organize, share, and automate photo workflows with a single prompt.
Gemini Voice now lets you co-write in Docs, search your Gmail, and organize your notes in Keep using conversational voice commands.
Google Pics is a new image tool that integrates into Docs, Drive, and Slides, with collaborative editing, in-image text editing, and object isolation features.
Lyria 3.5 music model is now available in the Gemini app and lets users generate full songs with control over genres, lyrics, and vocals.
Meta news:
Muse Spark 1.3 performs better on long-running agentic and coding tasks while using fewer tool calls and tokens.
Muse Voice Transcribe is a real-time speech-to-text model that can identify 20+ speakers in a single conversation and was trained on 70+ languages.
OpenAI introduced GPT-6 Astra, its new flagship model for computer use, coding, research, cybersecurity, and complex professional work.
OpenClaw released OpenClaw 2.0 with easier setup, a rebuilt browser app, and shared sessions that let teams collaborate without losing context.
Perplexity launched Hybrid Compute on Mac, which splits AI tasks between cloud and local models so sensitive data never leaves your device.
🔬 AI research
Google previewed Agentic Video Understanding, which searches for video segments via audio, frames, and transcripts rather than processing entire videos.
Runway research:
GWM Worlds 2 generates interactive 720p AI environments with spatial audio in real time, which respond to your inputs as you explore.
Solaris can create working apps and websites in real time as users interact with them, frame by frame, without any underlying code.
World Labs previewed Atlas, a world model that can generate, reconstruct, and simulate 3D-consistent worlds with precise camera control.
📖 AI resources
“PII-TRACE” [BENCHMARK]: measures how well AI can detect personal information in long conversations.
🔀 AI random
Anthropic tightened its security practices with stronger sandboxing and real-time monitoring after its models took unauthorized actions during evaluations.
New York City introduced an AI classroom moratorium banning student AI use through eighth grade (and companion chatbots across all grades).
Senator Sanders introduced a bill to pause AI development after recent reports of AI agents circumventing restrictions and gaining unauthorized access.
🤦♂️ AI fail of the week
The prompt for this was simply “An insane pool trickshot,” and, I mean, fair enough.


