vLLM-Omni Ships Day-0 Support for Qwen-Image-2.1

The vLLM Recipes blog details day-0 support: a 7.1B single-stream DiT paired with a Qwen3-VL-8B text encoder and a 16x RGBA autoencoder, using block-causal attention and a cross-step prefix KV cache to handle generation, editing and transparent-image output in a single model.
SGLang Runs Qwen-Image 2.1 on a Single RTX 4090
SGLang-Diffusion also announced day-0 support, running full-precision inference on a single RTX 4090 24GB with CPU offload — 18.7 seconds for 1024×1024 generation and 21.7 seconds for editing at 22.7 GiB peak memory. On an RTX PRO 6000 96GB, generation drops to 8.0 seconds.
ICLR 2027 Submissions Beat Every Prior Year Combined

Denny Zhou notes that ICLR 2027 has received more submissions than 2013 through 2026 combined — a vivid illustration of how explosively AI paper output has grown, and a sharp reminder of how hard peer review is becoming to scale.
Politico Reconstructs the White House's 85-Day Fight With Anthropic
A Politico investigation interviewing more than a dozen participants reconstructs the 85-day battle between the Trump administration and Anthropic over AI safety regulation from April to July this year. The story largely defined the current US frontier-model regulatory framework — even as newer models race past it.
Qwen-Image-2.1 Now Supported in ComfyUI
Qwen announced the generation and editing model can now be called directly from ComfyUI workflows.
Two New OCR Models Land on Hugging Face
Tencent's WeVisDoc (2B and 4B, Apache 2.0) and jinaai/jina-ocr-v1 (non-commercial) arrived the same day.
Lambert: China's Top Labs Shift Inference to Huawei
At scale for RL, China's leading AI labs are using Huawei heavily for inference and Nvidia for training, a trend set to accelerate with agent swarms.
rasbt: Jev's Breakthrough Is Its Generalization
Jev shouldn't be dismissed as "just a classifier" — the key is that it generalizes well where encoder models usually stay narrow.
Researchers Call for Canceling ICLR 2027
Hieu Huynh suggests authors and organizers cancel the conference and draft policies that restore top venues' prestige.
Report Corrected: Waymo Is Safer Than NYC Commercial Vehicles
An earlier "Waymo is more dangerous" analysis proved wrong; the updated report finds robotaxis clearly safer and urges data transparency.
GPT-6 Astra Tops Terminal-bench Science
GPT-6 Astra leads at 65.7%, a 31.4-point jump over the previous leader.
Jev Founder: LLMs Took the RLHF Detour
Diogo Almeida argues Jev aimed at automation from the start, while current LLMs wandered down an epic detour called RLHF.

Recraft V4 Styles Locks a Style From a Single Image
One reference image keeps a consistent look across photography, illustration, 3D, icons and posters.

Vidu S2-Editing Enables Real-Time Style and Outfit Changes
Swap style, clothing, subject and background in real time and see results as you edit.

Developer Builds a Real-Time 3D Scene Generator With JEV
Hundreds of concurrent judgments over 100+ prefab assets handle shading, lighting and state to assemble a matching interior in one second.
Grok Imagine Image 2.0 Rises to No. 4 in Text-to-Image
The model climbed to fourth on the Artificial Analysis text-to-image leaderboard, a clear jump in capability.
Bending Spoons Self-Hosts Open Models for 99% of Requests
About 99% of AI requests and tokens go to self-hosted open-weight models; only ~1% call frontier closed models.
LeCun Reiterates Autoregressive LLMs Won't Reach Human-Level AI
LeCun stresses his wording: autoregressive LLMs alone will not bring human-level AI, as he keeps sparring with the Hinton camp.
Nikkei Profiles the Head of Google DeepMind Tokyo
hardmaru shares the feature and recalls co-founding the Google Brain Tokyo team with him in 2018.

Qwen-Image-2.1 Launches a Hugging Face Space
Generate from short prompts or edit existing images with text, with prompts rewritten first to improve results.
Researcher Proposes 'AI Safety Level' Facilities
Physical isolation inside a Faraday cage, with levels graded by the parameter count or compute of the models inside.
Mollick: Long-Horizon Agents Show Language Degradation
The most annoying failure in long agentic tasks is no longer code errors but language drift that worsens as the run progresses.
Mollick: Claude's Missing Image Generation Is a Knowledge-Work Gap
Google's and OpenAI's image capabilities offer more options for slides, prototypes and infographics.
7B Open-Weight Image Model Claimed to Beat Nano Banana
The open-weight model also supports RGB-to-RGBA for images with a transparency channel.
GPT-6 Astra Links Unreal and Blender
A developer shares a new way to chain the engine and modeling tool with GPT-6 Astra.
Wan 3.0 Adds Peel-Off Sticker Animation
A single photo drops you into any scene, animated by the Peel-Off Sticker skill.
LeCun Criticizes Hinton for Fueling AI Lockdown Talk
The open-versus-regulated research debate keeps simmering.
Lambert Charts the History of Open vs. Closed Models
Long-horizon data on the two models' share of releases.
Jev Puts Encoder Models Back at the Center of Classification
Commentary argues classification-style models are returning to the spotlight.


