Show-1
by showlab
[IJCV] Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
AI summary
Video generator library
This project enables text-to-video generation using a combination of pixel and latent diffusion models.
- stars
- 1.1K
- forks
- 62
- watching
- 39
Similar projects
Found by comparing what the projects do, not just their names.
showlab/vlog545
Video doc generator
Transforms video content into a long document containing visual and audio information that can be used for chat or other applications.
Text-Video Embedding Toolkit
Provides code and tools for learning joint text-video embeddings using the HowTo100M dataset
Video generator
Generates time-lapse videos from text inputs using deep learning models.
Video processor
An audio-visual language model designed to advance spatial-temporal modeling and audio understanding in video processing.
taoxugit/attngan1.3K
Text to image generator
Reproduces text-to-image generation with attentional generative adversarial networks.
Motion generator
Develops a unified model to generate high-quality motions and text descriptions from human motion data
Text-to-image generator
Develops a PyTorch model for 4K text-to-image generation using diffusion transformer
Text image generator
A text-to-image tool using CLIP and FFT/DWT parameters to generate detailed images from user-provided text prompts.
Diagram generator
A Python library to generate diagrams in various formats from structured data
Mask generator
Automatically generates masks for image inpainting using natural language input
Video preview generator
Generates image strips or GIFs from video files
Video generator
Converts music represented by a GNU LilyPond file into a video containing a horizontally scrolling music staff synchronized with audio rendering.
Motion generator
Generates human motion from text input using a diffusion model
Text generator
Provides tools and scripts for generating text using a pre-trained Chinese language model
Texture Generator
Generates synthetic digital images of visual textures based on mathematical models