Qwen-Image-2.1-Turbo: eight steps, open weights
Alibaba's accelerated image checkpoint brings near-instant generation and editing on the same 7B visual-generation architecture.
Built on Qwen-Image-2.1, Turbo is an accelerated checkpoint that needs only eight denoising steps to generate and edit images. The team insists that fewer steps do not mean lower quality, and ships the weights openly. The release continues a steady cadence of open image models from the Qwen family, aimed at developers who want fast, self-hosted visual generation without sacrificing fidelity.

Codex predicts your next message
OpenAI's composer predictions beta in Codex reads the conversation and how a user talks to it, then suggests the next message. Available now for Pro users.
Anthropic opens a new reporting cadence
Beyond system cards and risk reports, Anthropic will publish model-behavior reports more frequently. The first describes four categories of behavior observed during evaluations and internal use.
SGLang gets tuned for NVIDIA's Vera Rubin
Working with NVIDIA on early-access Rubin hardware, the SGLang team optimized attention, MoE, and speculative-verification kernels, and used the stack to accelerate Kimi K3 inference. FP8 MLA runs up to 20% faster at batch 1 with 128K context.
If you want to really understand AI RSI, science should be your reference point.
— François Chollet
TRL v1.15 cuts peak VRAM up to 82%
Fused LM head is now on by default, slashing peak VRAM and supporting roughly 7x longer sequences.
Codex for Windows gets a new sandbox
Built on Microsoft Execution Containers, it brings faster setup, stronger network enforcement, and granular file access controls.
Claude Code Projects opens the waitlist
Every Pro and Max user on the list is now in, with a four-minute walkthrough covering the basics.
Tencent Hunyuan ships ExplorationBench
A benchmark for rule discovery in environments that demand active exploration: 55 goals, 140 held-out tasks.
Rillet turns requests into PRs with agents
eve agents run on Vercel; features ship in as little as two hours and merged PRs tripled.
Replit desktop enters private preview
With Microsoft and NVIDIA, Windows builds run in isolated MXC-based sandboxes; the waitlist is open.
Semantic routing Decision 2.0 lands
Multiple questions about one input, answered in a single forward pass with per-option probabilities.
Cognition is Rubin's first production customer
Hosted by CoreWeave and serving SWE-2 since September: 4.8x per GPU over GB200 at matched interactivity.
clef-omni takes every modality
Audio, video, image, and text in one model, while clef-flash gets another price cut.

OpenAI math problems become open RL environments
The problems are repackaged as open-source reinforcement-learning environments for direct training and reproduction.
Agents are starting to buy infrastructure
The Vercel CEO says agents buy cloud products and services via CLI more than expected; the capability now extends to domains.
LeWAM: JEPA imagines state and action
A JEPA for the real world that predicts both what happens next and what to do, with 32.3x faster training.
Frontier models beat human forecasters
New results claim frontier AI models outperform human experts on financial-prediction tasks.
GR00T robot model hits Hugging Face
Agile One S SSD Pick is a GR00T-based deployment model for SSD carrying tasks.
Rubin inference profit per watt up 3.2x
NVIDIA's vLLM Rubin inference boosts profit per gigawatt 3.2x and per-dollar performance up to 10x.
Neel Nanda questions OpenAI's account
The interpretability researcher says the firing story "doesn't add up," laying out two possibilities.

AI tutoring raises scores; ghostwriting fades
A randomized GPT-4o trial found tutoring gains persist a week later, while having AI write for students does not.

Older Gemini matches doctors on ER advice
Gemini 2.5 Pro and Flash advice was rated comparable to doctors in urgent care, with no safety issues spotted.
Decision-1: a model that only chooses
A Qwen3.5-9B-based model that judges given options and returns probability scores, live on Microsoft's platform.
AlphaProtein Novo designs enzymes
A pipeline for de novo enzyme design, built with Frances Arnold's lab.
ChatGPT dots now start on mobile
You can create your dot straight from the ChatGPT app on iOS and Android.

CEOs' AI questions turn organizational
Kai-Fu Lee says the conversation has shifted from technical questions to how AI changes the economics of a business, what happens to work and leadership, and how a company must reorganize.

The smartest models interact with the world
In a Modal talk, Sara Hooker argues we are in the era of interaction: the strongest model will be one that interacts with the world and rapidly incorporates new knowledge.
Qwen3.8-Max free for a week
Qwen3.8-Max, Qwen3.8-Flash, and Wan3.0 are free on GMI Cloud for one week.
Product Shots reuses one image
Place the same product shot into new settings without rebuilding from scratch.
Moses shot on stage with AI
Jon Erwin used real actors plus AI to carry them into any world the story needed.
v0 for teams ships
See what teammates are working on, jump into their chats, and build together.
LlamaParse untangles nested tables
18 values correctly assigned under the right headers in Micron's earnings deck.
Animation from code, not video gen
Claude plus Higgsfield Katana replaces After Effects and Blender — pure code.
Ideogram 4.5 claims precision
The team calls 4.5 the most precise image-edit model available.
40 edits, head to head
Nano Banana 2.1 and Ideogram 4.5 each iterated on their own last image, same prompt.
Bending the Curve of Discovery
Hassabis and Manyika publish a new essay on AI and science.
pplx-decider tops Decision Bench
94.5% accuracy at the lowest cost on the leaderboard.
93 hours of explainers
Computer made explainer videos for all 722 OpenAI math preprints.
Whistle: 16.9MB speech-to-text
A tiny on-device model that rivals Whisper base.
MedDecider goes open-weight
Decision models for medicine reach leading benchmark performance.
Open weights net-boost AI demand
Compressed model-layer margins, but expanded usage means more infrastructure.
Busabase: memory for agents
An MIT-licensed database so agent output stops getting trapped in chat history.
ARC-AGI-3 hits 59.17%
Yi-Chia Chen leads the ARC Prize 2026 board, ahead of Tufa Labs.
Provider token ranking sparks debate
A top-five list by daily tokens draws scrutiny over methodology.
Efficiency beats breakthroughs
Natolambert bets on efficiency gains as the bigger lever for diffusion.
Datology's curation playbooks
Practical reports on filtering data across pre, post-training, and evals.
PoolDINO drops 4-16x tokens
RAE reaches comparable quality with far fewer tokens.
A tenth of Opus's cost?
Stas Bekman weighs openai-gpt-6.1-sol for cheaper code quality.
TermGrade for terminal agents
Open-source reinforcement-learning environments, fully open.
A blog feature, built by voice
Simon Willison used Codex Desktop voice mode while cooking dinner.
Google's test after Gemini 4
Mollick: a single interface plus orchestrated agents, not fragmented products.
More proofs aren't progress
New goals are needed for what math is trying to do.
Startup perks paused after 3 days
Free Claude Team plans and $1,000 API credits paused as demand exploded.