September 29, 2026 · Tuesday



The purpose of the AI industry should be to produce tools that, in the human hand, improve human prosperity and welfare — not to create a "successor species" to the human race.

— François Chollet


Models & ReleasesAI · DAILY
Video Generation

Kling 4.0 Lands in October, Flash Goes Live First

The company says 4.0 advances visual realism, creative control, and narrative completeness, with upgraded audio-visual quality and stable motion; Ultra yearly subscribers can already use the Flash version.

Meta

Meta Ships the Muse Series Over Six Months

Meta recaps six months of work — Muse Spark, Image, Video, Code, the Meta Model API, and Muse Glimmer — and says it is just getting started.

Speech

ElevenLabs v4 Launches on Runway

v4 is more expressive and natural in narration and character dialogue, and is now available on Runway alongside its image and video models.

Avatars

Synthesia Releases Avatar Model Express-3

Synthesia calls it its strongest digital human model yet — faster, sharper, and more expressive, with custom styles available on all plans.

xAI

Grok 4.7 Launches on Amazon Bedrock

Grok 4.7 is now available through AWS Bedrock, further entering the enterprise cloud ecosystem.

xAI

Grok Bot Adds Team Collaboration

Grok's Bot now works in teams, letting multiple members share the same bot and its accessible materials.

Developer Ecosystem

OpenAI Names the WebMCP Challenge Top 10

The ten winning projects show how humans and agents can collaborate when websites expose structured tools to agents.

Image Editing

Qwen Image 2.1 Frame-Lock LoRA Released

The LoRA locks edits to the original frame, producing zero image shift under the same prompt edits for stable, consistent changes.

Tooling

Replit Ships Multiple Updates at Once

It adds Meta Quest VR app building, app creation through Muse, Airwallex payments, new models, and conversational data insights.


Benchmarks

GPT-6 Sol's Measured Results on ARC-AGI

On ARC-AGI-3, the standard harness scored 4.6% (about $5,600), while a provider adapter harness reached 23% (about $8,700) — a notable gap in both score and cost.

Datasets

YODAS v3 Releases 1.1 Million Hours of Speech

The dataset is now on Hugging Face with 1.1 million hours, making it one of the largest audio training datasets available.

Research

LeCun Lab Doubles Navigation Success With a Brain Trick

The work borrows techniques from the human brain to more than double an AI's success rate in finding target paths.

Open Source

K3-Node: A Native Keras 3 GNN Library

Its public API fully aligns with PyG; models run on JAX, torch, and TF, with hardware acceleration for Apple Silicon, TPU, and more.

Books

Neuroevolution Textbook Goes to Print

Authored by Ha, Risi, Tang, and Miikkulainen and published by MIT Press, with a free online edition also available.

Paper

FuseReg Narrows the Reconstruction-Generation Gap

One decoder handles full, sparse, and single-layer inputs without retraining; swapping only the decoder on ImageNet-256 cuts gFID by 27%.

Paper

A Full Overview of Nemotron Post-Training

The article covers SFT, Cascade RL, RLVR, PivotRL, agent training, and MOPD, and reviews the evolution from Llama-Nemotron to Nemotron 3 Ultra.

Paper

OpenAgentSafety Paper Clashes With NVIDIA's Name

The framework covers eight risk types and 350+ multi-turn tasks; testing five mainstream LLMs found unsafe behavior in up to 51% of vulnerable tasks.

Evaluation Science

What's Still Worth Measuring After Saturation

Using CORE-Bench Hard as a case study, the paper argues that after saturation, measurement should shift to construct validity, out-of-distribution generalization, efficiency, reliability, and human-AI collaboration gains.

Datasets

Largest Open Human Video Preference Dataset

datapoint released the largest open human video preference dataset to date and doubled its data funding to $2 million.

Paper

HalluWorld Builds a Controllable World for Hallucinations

The team argues reality is too messy to measure hallucinations well, so it built a controlled environment for more precise evaluation.

Security

Agents Can Tamper With Their Own Traces

A new paper shows Claude Code, Codex, and Antigravity can modify their own traces, challenging observability and evaluation credibility.

Paper

Linear Mode Connectivity Is Underrated

The author revisits a classic result from four years ago and argues linear mode connectivity is one of the most overlooked properties in neural network training.

Infrastructure

xLLM Makes Training Rework Costs Manageable

IFM releases xLLM for unavoidable rework in large-scale LLM pretraining and fine-tuning, aiming to lower rerun costs and improve efficiency.

Economics

Still No Evidence AI Is Hitting New-Grad Jobs

A new NBER paper takes a cautious stance and finds no evidence that AI has raised unemployment among recent college graduates.

Commentary

Chollet: I No Longer Write Code, I Direct LRMs

He explains this is not because model code quality is good enough or instructions are always perfectly executed, but because large reasoning models have changed coding itself.

Product

Manus Launches 2.0 With Video and Game Environments

Its first major update after leaving Meta brings a new underlying framework, video editing and game development environments, and a personal agent app called Cue.

Safety

Perplexity to Open-Source Its Agent Sandbox

Aravind Srinivas says Perplexity will work with NVIDIA to build a guarded safe agent sandbox and plans to open-source all results.


SignalsAI · DAILY

© 2026 FAV0 · AI Daily