Preprint2024
Snap video: Scaled spatiotemporal transformers for text-to-video synthesis
Unknown
This paper introduces Snap Video, a scaled spatiotemporal transformer model for text-to-video synthesis that generates high-quality videos from text prompts.
0Jan 1, 2024Transformers