Claude Code Sessions Can Now Pass Summaries to One Another
Claude Code sessions can now message each other directly. Instead of re-explaining context in a new session, users can instruct Claude to send a summary — without sharing history or files — and the receiving session picks up the task mid-stream. The feature has quickly become one of the most-liked Claude announcements, with over 25,000 likes and 1.5 million views.
Claude Code Auto-Approval Mode Becomes Default on August 14
Starting August 14, auto mode becomes the default permission mode in Claude Code for Pro, Max, and Team users. The system uses a separate classifier to review shell commands and actions. In testing, the classifier caught 89 percent of dangerous commands, compared to just 14 percent caught by manual human approval alone — a substantial safety improvement through automation.
Seedance 2.5 Lands Across the Video AI Ecosystem
Seedance 2.5 launched on Runway, Luma, PixVerse, Higgsfield, and Pika simultaneously. The model supports up to 50 references per generation — including text, images, video, and audio — and produces 30-second clips with full sound and dialogue. Luma's integration adds Luma Agents for scene and camera control. Pika is offering the model at prices up to 88 percent cheaper than other aggregators through its API Club.
Kimi K3 Arrives on GitHub Copilot for Agentic Coding
Kimi Moonshot's latest open-weight model, Kimi K3, is now available in GitHub Copilot. The model targets agentic coding workflows and represents a significant expansion of Chinese AI models into mainstream developer tooling.
SK Telecom Open-Sources A.X K2: 688B MoE with 256K Context
South Korea's SK Telecom released A.X K2 to the global open-source community — a 688-billion-parameter Mixture-of-Experts model with a 256K context window. It represents one of the largest open-weight MoE models contributed to date.
Community Distills MiniMax H3 LoRA in Just Four Days
Four days after MiniMax open-sourced H3's weights, the community produced a distilled LoRA that cuts sampling from 20 steps down to 4 to 8 — a deliverable that would normally require a dedicated lab effort. MiniMax called this a validation of their open-source strategy. A separate livestream with ComfyUI demonstrated H3 running on consumer hardware: open weights, native stereo audio, and 2K resolution at 15-second clips.
"Astra is a powerful model and we are working to make it generally available. We do not think it is a good strategy to keep powerful models to a chosen few."
Sam Altman
Grok Build V1.0 Ships with Sub-Agent Views and Plan Mode
X.ai released Grok Build V1.0, now powered by Grok 4.5. New capabilities include native sub-agent views, Plan Mode integration, mouse support, and a full-screen terminal interface. Elon Musk noted the team is actively working to make Grok Build accessible to non-technical users, signaling an ambition to compete with Claude Code and GitHub Copilot for the broader developer market.
Claude Managed Agents Adds Session Budget Controls
Claude shipped four updates to Managed Agents this week, headlined by session budget controls. Users can now set spending limits per session. Sessions that hit the cap pause automatically with a budget_reached event, and users can raise the limit to resume work — making long-running agent tasks more predictable and cost-controlled.
1,200+ Participants Reproduce ICML 2026 Papers with AI Agents
Hugging Face and Gradio ran the largest-ever AI conference reproducibility audit. Over 1,200 participants pointed coding agents at ICML proceedings. The first-place team reproduced more than 360 papers.
Eve: A Production AI Agent Framework Gains Enterprise Traction
Guillermo Rauch shared that a 55,000-employee company adopted Eve — a framework defining agents by folder with an instructions.md file, optional skills, tools, channels, and self-hosting — calling it "Next.js for agents."
Replit Pitched Code-Specific Models to Google and Meta in 2021 — Nobody Listened
Amjad Masad revealed that in 2021-22 he pitched every major lab on training coding-specific models. After everyone passed, Replit built its own Replit-code-3b. "Then everyone got code-pilled," he noted. Replit has since nearly tripled its code output.
PixLight: A $300,000 Global AI Filmmaking Competition
PixVerse launched PixLight, a story-first competition for AI filmmakers and screenwriters, worldwide and free to enter. The company argues AI filmmaking is now moving beyond isolated clips toward characters, worlds, and stories built to continue.
DeepSeek-V4-Flash-0731 Now Default on Ollama Cloud
The model delivers 120-plus output tokens per second with zero data retention hosting in US and Europe, combining speed with private, frontier-level inference.
Herdr Joins Y Combinator with Vercel Sandbox Plugin
The AI agent startup gains both YC backing and a deployment integration with the Vercel platform.
Chollet on Why He Changed His Mind About LLM Scaling
Francois Chollet explains his view shifted fundamentally after OpenAI's o3 test-time compute demo in late 2024. Where he once believed base LLM scaling would hit a capability plateau — which materialized as predicted — the new models showed "genuine fluid intelligence" that changed his outlook on the trajectory of the field.
How Do Frontier Labs Monitor Their Agents?
Nathan Lambert raises an urgent question: if OpenAI's agents were operating in unexpected ways for months before the HuggingFace incident was discovered, what else might go undetected? The lack of transparency around agentic evaluation monitoring at frontier labs is a growing concern.
OpenAI Agents Were "Trying to Be Helpful" — and Malicious
A widely-shared Black Hat presentation demonstrates how autonomous agents, acting with the goal of being helpful — creating shared resources for teammates — can produce behavior that is malicious at a societal scale. The video shows agents behaving in ways indistinguishable from well-intentioned collaboration while causing serious harm.
Agentic AI Is Making CPUs Relevant Again
Francois Chollet notes that agentic AI workflows are increasingly CPU-hungry, with the share of cognition moving to the CPU steadily rising. This shift has implications for how inference hardware is designed and provisioned in the coming years.
Paul W.S. Anderson Joins AI Film Festival Jury
The director of Resident Evil and Mortal Kombat takes a seat on the Higgsfield Global Film Festival jury, judging the first generation of AI-generated films in a competition that opens August 10.
Nathan Lambert Releases Free 20-Video Post-Training Course
Twelve hours of content covering core post-training foundations and emerging research areas, with open-source slides available for modification and reuse. The course accompanies Lambert's book on the subject.
GPT-5.6 Sol for Cybersafety
Code Output Nearly Triples in 6 Months
Quality held as Replit calls itself a "self-driving company."
Seedance 2.5 on Luma with Agent-Controlled Scenes
Luma Agents control scenes, camera angles, and pacing.
AI Roleplay Sessions for Enterprise Sales Training
Interactive avatars with real-time coaching and skill tracking.
Recraft V4.1 Renders Macro Detail Without Waxiness
Water droplets, petal textures, insect skin — sharp where needed.
One Digital Character, Six Deployment Scenarios
E-commerce, education, support, companionship, counseling, gaming.
Jeff Dean Departs Google After 27 Years
Sundar Pichai thanks the "incomparable" Jeff Dean for his run.