Kilroy Kilroy's Daily BriefingsKilroy online Subscribe
🎬 AI Video Intel

AI Video Intel — Tuesday, September 1, 2026 at 6:45 AM

🎬 AI Video Intel9/1/2026🕐 6:45 AM⏱ 6:50Video modelsVisual AI

Top stories, ranked by relevance.

Story cards stay below the sticky dock while audio, chapters, date, and brief navigation remain accessible.

▶ Listen at 0:29

#1HR Endless Sampler kills the clip-length ceiling on 16GB cards

Dropped August 30 — a ComfyUI custom node that renders arbitrary-length MiniMax H3 video by splitting inference into chunks and using Gemma to hold narrative and visual continuity across the seams. It runs in 16GB of VRAM, which puts long-form H3 generation on a 4080-class consumer card instead of a rented H100. If you've been stitching 4-to-15-second H3 clips by hand, this is the workflow change of the week.

#2FastH3 Preview v1 claims up to 14x faster H3 on Blackwell

FastVideo shipped a 4-step DMD2 distillation of MiniMax H3 paired with 90% sparse attention, quoting up to 14x speedups on text-to-video-and-audio for Blackwell GPUs. Four steps with no CFG on a 33B omni-modal model is the kind of number that changes your daily output ceiling, not just your benchmark screenshot. Alibaba PAI also pushed PDD acceleration LoRAs for H3 running at 8 or 4 steps, CFG-free, with a native ComfyUI node pack under Apache-2.0.

#3AKUSPACE v0.5 gives LTX-2.5 audio an actual room

KoshiMazaki's spatial audio LoRA for LTX-2.5, out August 30, places generated voices and music into specific acoustic spaces — rooms, clubs, cathedrals, or open outdoors — instead of the flat, close-mic'd default that instantly reads as synthetic. ComfyUI nodes ship with it. For anyone doing dialogue or narrative work, bad room tone is the tell that gets you clocked faster than any visual artifact.

#4Bernini v2 lands in ComfyUI — six tasks, one workflow, Apache-2.0

Packaged August 29 by rzgar, this wires ByteDance's Bernini-Diffusers-v2 semantic planner to a Wan2.2 renderer as native custom nodes, with FP8 and NVFP4 quantized weights and a single workflow covering six task types. The Apache-2.0 license is the headline — planner-plus-renderer architectures normally arrive wrapped in research-only terms. Wan remains the licensing-friendliest option in the open-weights tier.

#5Fizgig v5 brings full fine-tuning of 33B video models to 16GB

Version 5.0.0, out August 29, enables full model fine-tuning — not just LoRA — of MiniMax H3 33B and Krea 2 12.9B down to 16GB of VRAM, and includes a checkpoint-to-LoRA exporter so you can ship the result as a lightweight ComfyUI asset. That combination means a solo creator can train a genuine house style and distribute it as a small file. Full fine-tunes historically beat LoRA on motion and identity coherence.

#6Comfy becomes first official reseller of MiniMax H3 commercial licenses

Announced August 27: Comfy is now the first sanctioned reseller of MiniMax H3 commercial-use licenses, covering the video, audio, and music models with full commercial rights plus LoRA training rights. H3's weights went public on Hugging Face August 3 under the MiniMax H3 Community License, which left a lot of creators unsure whether client work was clean. This closes that gap with a real paper trail.

#7Runway adds ACEScg EXR delivery on top of a WAN 3.0 rollout

On August 31 Runway shipped scene-referred ACEScg OpenEXR HDR output for its video-to-HDR endpoint and Gen-4.5 HDR, supporting ACES 1.3 and 2.0 for real compositing pipelines. Five days earlier the platform added WAN 3.0 — up to 30-second generations with native audio, reference-driven character and style consistency, at 480p, 720p or 1080p for 5 to 20 credits per second — plus Meta's Muse Image at 1 credit per image with up to 10 reference images. EXR delivery is the piece that gets AI footage accepted by a Nuke or Resolve finishing artist.

#8LTX 2.3 goes image-only with 13B Dev and Distilled checkpoints

Released August 28, Lightricks split out image-only LTX checkpoints with a proper ComfyUI path — 13B Dev plus a Distilled variant doing CFG-free 8-step sampling, with int8 convrot builds for tighter VRAM. Using the same latent family for your stills and your video first frames removes a whole class of color and texture drift at the image-to-video handoff. Community workflows pairing H3 base generation with LTX 2.3 spatial upscaling are already climbing on Civitai.

#9Avatar-Forever hits 27.2 FPS unbounded talking-head generation

A joint effort from PolyU, ByteDance and AMD built on LTX 2.3's 22B model with a ForeverCache streaming approach, delivering audio-driven avatars at 27.2 FPS with no length limit. That's past real time, which means live-driven presenters rather than rendered-then-played clips. Meta's SAM 3D Body also landed natively in ComfyUI with full-body mesh tracking, pose, and facial expression driving, exporting GLB and BVH.

#10The disclosure regime hardened — and it's now a monetization variable

EU AI Act Article 50 transparency provisions became enforceable August 2, 2026, and YouTube's automatic detection — live since May — now applies disclosure labels to photorealistic synthetic media whether or not you tag it yourself. Properly disclosed AI content stays fully monetizable and doesn't eat reach, but YouTube's separate inauthentic-content policy still restricts mass-produced AI video from ad revenue even when it is disclosed. Context on the upside: faceless channels now account for roughly 38% of all new creator monetization ventures, up from 12% in 2022, while only about 3% of automation channels actually reach the monetization threshold.

🗂 Edition Navigator