ConsistI2V
★ 261Official code for ConsistI2V (TMLR 2024), a diffusion method that turns a still image plus text prompt into a consistent video using first-frame spatiotemporal attention and low-frequency noise initialization. Includes inference and training scripts, a Gradio demo, and Hugging Face weights.
AI Tools | Python · image-to-video · video-generation
View Project →