TencentARC/SCoPE · Hugging Face

SCoPE: Sightline-Coordinate Positional Encoding for Video Diffusion Transformers SCoPE adds camera sightlines as positional coordinates to a pretrained video diffusion transformer. Given a first frame, a text prompt, and a camera trajectory, it generates a video that follows the requested camera motion while preserving the original image-to-video prior. This repository is a self-contained release for Wan2.2-I2V-A14B : it contains everything required for inference, so a separate Wan2.2 checkpoint download is not needed. arXiv : arxiv.org/abs/2606.27345 PDF : arxiv.org/pdf/2606.27345 GitHub : github.com/TencentARC/SCoPE Project : visual-ai.github.io/scope Demo : huggingface.co/spaces/TencentARC/scope-camera-video-generation

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论