Claude Design, Slides, and Docs Move Into Claude Code
Claude Design, Claude Slides, and Claude Docs now work inside Claude Code. Developers can ask for a design review deck or a UI mockup and point it at the actual files and RFCs in a repository, edit the result themselves, or keep iterating in the conversation and share a link when it is ready.
Mistral and Mozilla Bring Private AI to the Browser
Mistral and Mozilla announced a partnership to bring privacy, control, and choice to people using AI to browse online. The deal promises open, private, multilingual AI in the browser itself, though the specific product form and launch timing were not disclosed.
the main thing i was excited about launching this week will be next week instead, but imo worth the wait!
Sam Altman, OpenAI
Vercel's Default Safety Reviewer Switches to Jev
Vercel's fx now defaults to auto mode, with a safety reviewer analyzing every command. That reviewer runs on GPT Luna today, but the new Jev is up to 18x faster at p95 and more accurate — and it is coming to Vercel AI Gateway, likely as the new default.
TypeSafe AI Ships Its 'System One' Model, Jev
Jev forgoes text generation entirely. It is built for fast decisions inside software: unstructured data in, type-safe structured results with calibrated probabilities out — a category the company, founded by a ChatGPT co-inventor, calls a "System One" model.
DeepMind Opens an Institute for AGI's Economic and Social Questions
Demis Hassabis and Shane Legg are expanding interdisciplinary research on how AGI could reshape the economy, science, and society. The new DeepMind Institute is meant to spur the discussions they believe are needed to get the next steps right, drawing on two decades of shared work.
Meta Makes the Case for Personal Superintelligence
Meta argues that superintelligence should broadly benefit individuals rather than a few institutions. The company is building personal agents and creation tools — emphasizing privacy and the balancing of power — in a vision modeled on the personal computer and the internet rather than on centralized control.
Runway's Fall Collection Lands
Runway's biggest model drop yet adds Fish Audio S2.1 Pro, MiniMax H3 Max, Cartesia Sonic 3.6, and Flux video upscaling and editing alongside the best models from frontier labs.
Vidu S2 Adds Real-Time Style Editing
S2-Editing switches styles, outfits, subjects, and backgrounds in real time — edit instantly and see it happen instantly.
StepFun Debuts StepAudio 3 Music
StepFun and ACE Studio turn a prompt and lyrics into a complete song, controlling genre, mood, vocal character, instruments, key, BPM, and structure.
Wan3.0 Called a Generational Leap
Alibaba positions Wan3.0 as more than an incremental update, pointing to motion reference, consistency, creative control, and resolution gains for creators.
vLLM Runs Kimi K2.x on W4A16 MoE Kernels
vLLM's Humming backend runs Novita's open-source Chord kernel, reaching 1.33x on H200 TP8 and 2.15x decode on B300 EP8.
SGLang Serves a Trillion Tokens a Day
A production deep dive shows SGLang HiCache and Mooncake Store serving trillion-parameter models while meeting strict inference SLOs.
Sakana AI's Year of Shipping
Sakana Chat, Namazu, Translate, Marlin, and the Fugu line — a Tokyo research lab proves it can ship products, not just papers.
MiniMax H3 Family Arrives on Monid
Claimed to be 10x faster than Seedance 2.5, with pay-as-you-go pricing and no subscription.
Doubao 2.1 Pro Update Goes Live
The upgrade focuses on agent task delivery, multimodal coding, multimodal understanding, and inference cost, fully available on Volcano Ark.
StepAudio 3 Realtime Technical Report
An audio-language foundation model built on a hear-speak-think-act loop, with "Think-While-Speaking" parallelizing reasoning and speech output.
Combined Anchors Lift Long-Term Memory 28x
Combining data, function, and weight anchors with merged LoRA raises final retention on 100 tasks from 1.2% to 34.9%.
Extropic Reveals Z1T for Sparse Hardware
The first family of transformer-like models built for Z1 sparse probabilistic hardware, claiming meaningful acceleration.
HuggingChat Defaults to DeepSeek-V4.1-Flash
HuggingChat updated its default model to DeepSeek-V4.1-Flash, a fresh measure of how far open AI has come.
Vidu S2 Enters Beta for Streaming Video
Built for real-time interaction, it improves motion, emotion, intent understanding, voice stability, and reference editing for livestreams, companions, and game NPCs.
Future AI is a cause for fear & Current AI is rapidly becoming part of life.
Ethan Mollick
Perplexity Ships a Search SDK for Coding Agents
Parallel retrieval, official-doc filtering, and snippet extraction help agents build sourced dependency migrations before touching code.
AI Saves Scientists Seven Hours a Week
Google research finds acceleration alongside shifts toward verification work and, possibly, safer research topics.
The Benchmark Crisis Is Real
Mollick warns that famous benchmarks are maxed out while the rest are riddled with errors that underestimate AI's abilities.
Arcee AI Hits a $1B Valuation
A Series B round values the company above $1 billion and accelerates its next-generation models and products.
G5 Labs Raises $14M in Seed
Emerging from stealth, it is building a new abstraction layer for AI-native software.
Hugging Face CEO: Cyber Laws Lag AI Attacks
After an AI-led cyberattack on Hugging Face, Clement Delangue says existing cyber laws are likely outdated.
Unity Ships an Official Codex Plugin
A first-party integration brings OpenAI Codex directly into Unity's editor workflow.
Large-Scale RL Resources Go Public
Researchers praise one of the coolest at-scale reinforcement learning resources released to date.
The Under-Recognized AI Flood
Overwhelming communication channels is a major AI harm with dozens of distinct forms and a heavy time cost.
LM Studio Bionic Adds Multilingual UI
Bionic 1.1.3 adds Japanese, German, Spanish, and French interfaces.
Muse Code Runs Natively on Windows
No WSL required; PowerShell-fluent and sandboxed by default.
Custom Connectors Reach Beta
Admins configure HTTPS REST APIs and API keys beyond first-party connectors.
Interactive Avatar API Launches
Real-time avatars embed via LiveKit with your own LLM and knowledge base.
Model Vault Goes Confidential
Confidential computing claims no one — not even Cohere — can access runtime workloads.
Databricks Deploys Astra Wall-to-Wall
Engineers company-wide now run the coding agent in daily work.
Everyone Races to Be a General Agent
OpenAI renaming Codex to ChatGPT echoes the rush for the agent position.
GPT-6 Sol Reportedly Slips a Week
Posters tie the delayed launch to next week's window.
Formal Proof Search Is Not General Intelligence
LeCun argues beating humans at theorem proofs does not equal AGI.
Recraft V4 Styles Turns One Image Into Many
Characters, objects, and scenes all build from a single reference image.