
Meta SAM 3.1 - 7x Faster Multi-Object Video Tracking
Meta releases SAM 3.1 with Object Multiplex, processing all tracked objects in one shared pass for 7x faster inference at 128 objects and improvements on 6 of 7 VOS benchmarks.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

Meta releases SAM 3.1 with Object Multiplex, processing all tracked objects in one shared pass for 7x faster inference at 128 objects and improvements on 6 of 7 VOS benchmarks.

Helios is a 14B open-source autoregressive diffusion model that generates minute-long videos at 19.5 FPS on a single H100, matching 1.3B distilled model speeds at full 14B quality.

OpenAI is shutting down Sora six months after launch, killing a $1 billion Disney deal that was supposed to anchor the product's future.

LTX-2.3 is a 22-billion-parameter open-source video and audio generation model from Lightricks that rivals closed commercial tools - at zero cloud cost.

ByteDance suspended the global launch of its AI video model Seedance 2.0 after Disney, Paramount Skydance, and other studios sent cease-and-desist letters alleging copyright infringement.

Kling 3.0 brings native 4K at 60fps, multi-shot AI Director, and single-pass audio to AI video - here's whether it lives up to the hype.

Luma Agents coordinates text, image, video, and audio from a single brief using the Uni-1 unified model - a genuine architectural leap, with some real rough edges still showing.

Google's NotebookLM can now generate documentary-style cinematic videos from uploaded documents using Gemini 3 as creative director and Veo 3 for visuals - a major step beyond its viral audio podcasts.

Luma AI's new Agents platform, powered by the Uni-1 Unified Intelligence model, lets creative teams go from a written brief to finished video, images, and audio in one workflow.

A hands-on review of Seedance 2.0, ByteDance's AI video generator that produces photorealistic 15-second clips with synchronized audio - and has triggered cease-and-desist letters from the Motion Picture Association.

ByteDance's Seedance 2.0 introduces a dual-branch transformer for simultaneous audio-video generation at 2K resolution, but cease-and-desists from Disney, Paramount, and Warner Bros. threaten its global rollout.

Rankings of the best multimodal AI models for image understanding, video analysis, and visual reasoning, covering MMMU-Pro, Video-MMMU, and more.