TencentARC/SCoPE · Hugging Face
SCoPE: Sightline-Coordinate Positional Encoding for Video Diffusion Transformers SCoPE adds camera sightlines as positional coordinates to a pretrained video diffusion transformer. Given a first frame, a text prompt, and a camera trajectory, it generates a video that follows the requested camera motion while preserving the original image-to-video prior. This repository is a self-contained release for Wan2.2-I2V-A14B : it contains everything required for inference, so a separate Wan2.2 checkpoint download is not needed. arXiv : arxiv.org/abs/2606.27345 PDF : arxiv.org/pdf/2606.27345 GitHub : github.com/TencentARC/SCoPE Project : visual-ai.github.io/scope Demo : huggingface.co/spaces/TencentARC/scope-camera-video-generation
评论
?
参与讨论