EditVid Framework Enables Diverse Video Editing Without Retraining
September 2, 2026
EditVid is a training-free framework for instruction and subject-guided video editing using sparse causal memory and post-attention token injection. It achieves 78.16 FiVE-Acc on the FiVE benchmark, outperforming existing training-free baselines in style transfer, object insertion, and subject replacement.
HOW THIS AFFECTS YOU
●
builderYou can implement high-quality, identity-preserving video edits without the overhead of fine-tuning models.
●
designerThis offers a unified toolset for complex video manipulations like part-level editing and subject replacement via text.