September 17, 2026 · Thursday


the main thing i was excited about launching this week will be next week instead, but imo worth the wait!

Sam Altman, OpenAI

Product

Runway's Fall Collection Lands

Runway's biggest model drop yet adds Fish Audio S2.1 Pro, MiniMax H3 Max, Cartesia Sonic 3.6, and Flux video upscaling and editing alongside the best models from frontier labs.

Model Release

Vidu S2 Adds Real-Time Style Editing

S2-Editing switches styles, outfits, subjects, and backgrounds in real time — edit instantly and see it happen instantly.

Model Release

StepFun Debuts StepAudio 3 Music

StepFun and ACE Studio turn a prompt and lyrics into a complete song, controlling genre, mood, vocal character, instruments, key, BPM, and structure.

Model Release

Wan3.0 Called a Generational Leap

Alibaba positions Wan3.0 as more than an incremental update, pointing to motion reference, consistency, creative control, and resolution gains for creators.

Inference

vLLM Runs Kimi K2.x on W4A16 MoE Kernels

vLLM's Humming backend runs Novita's open-source Chord kernel, reaching 1.33x on H200 TP8 and 2.15x decode on B300 EP8.

Inference

SGLang Serves a Trillion Tokens a Day

A production deep dive shows SGLang HiCache and Mooncake Store serving trillion-parameter models while meeting strict inference SLOs.

Product

Sakana AI's Year of Shipping

Sakana Chat, Namazu, Translate, Marlin, and the Fugu line — a Tokyo research lab proves it can ship products, not just papers.

Model Release

MiniMax H3 Family Arrives on Monid

Claimed to be 10x faster than Seedance 2.5, with pay-as-you-go pricing and no subscription.

Model Release

Doubao 2.1 Pro Update Goes Live

The upgrade focuses on agent task delivery, multimodal coding, multimodal understanding, and inference cost, fully available on Volcano Ark.

Research

StepAudio 3 Realtime Technical Report

An audio-language foundation model built on a hear-speak-think-act loop, with "Think-While-Speaking" parallelizing reasoning and speech output.

Research

Combined Anchors Lift Long-Term Memory 28x

Combining data, function, and weight anchors with merged LoRA raises final retention on 100 tasks from 1.2% to 34.9%.

Model Release

Extropic Reveals Z1T for Sparse Hardware

The first family of transformer-like models built for Z1 sparse probabilistic hardware, claiming meaningful acceleration.

Product

HuggingChat Defaults to DeepSeek-V4.1-Flash

HuggingChat updated its default model to DeepSeek-V4.1-Flash, a fresh measure of how far open AI has come.

Model Release

Vidu S2 Enters Beta for Streaming Video

Built for real-time interaction, it improves motion, emotion, intent understanding, voice stability, and reference editing for livestreams, companions, and game NPCs.


Future AI is a cause for fear & Current AI is rapidly becoming part of life.

Ethan Mollick
Industry & ResearchLabs

BriefsTools & Apps

© 2026 FAV0 · AI Daily