IF

Image generator

A text-to-image synthesis model with a modular design, utilizing a frozen text encoder and cascaded pixel diffusion modules to generate photorealistic images.

GitHub

8k stars
84 watching
504 forks
Language: Python
last commit: over 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
compvis/stable-diffusionA text-to-image model trained on images and text prompts using a diffusion process68,750
lucidrains/dalle2-pytorchAn implementation of DALL-E 2's text-to-image synthesis neural network in PyTorch11,184
ai-forever/kandinsky-2A multilingual text2image latent diffusion model with improved aesthetics and controllability2,774
ashawkey/stable-dreamfusionGenerates 3D content from text using a combination of neural networks and image synthesis.8,351
doubiiu/dynamicrafterThis project generates animated videos from open-domain images by leveraging pre-trained video diffusion priors.2,668
stability-ai/stablediffusionA software project that enables high-resolution image synthesis through a specific type of generative model using latent diffusion processes.39,501
modelscope/diffsynth-studioA software framework for training and utilizing various types of diffusion models.6,641
openai/glide-text2imA diffusion-based text-conditional image synthesis model3,562
nvidia/pix2pixhdGenerates photorealistic images from conditional inputs using deep neural networks6,685
microsoft/deepspeedA deep learning optimization library that simplifies distributed training and inference on modern computing hardware.35,863
dmitryulyanov/deep-image-priorA project demonstrating image restoration using neural networks without learning7,920
luodian/otterA multi-modal AI model developed for improved instruction-following and in-context learning, utilizing large-scale architectures and various training datasets.3,570
xpixelgroup/diffbirThis project provides a deep learning-based pipeline for restoring degraded images3,445
doubiiu/tooncrafterGenerates cartoon-style videos from two images using pre-trained diffusion models5,447
openai/clipA neural network trained on image and text pairs to predict the most relevant text snippet given an image26,460