saharmor/dalle-playground
quality grade D, 44 out of 100A playground to generate images from any text prompt using Stable Diffusion (past: using DALL-E Mini)
- stars
- 2.7k
- stars gained this week
- +1this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Diffusion models, image editing, upscaling and the surrounding creative tooling.
Signals: stable-diffusion, diffusion-models, image-generation, text-to-image, generative-art, comfyui, controlnet, flux
879 results
A playground to generate images from any text prompt using Stable Diffusion (past: using DALL-E Mini)
Kandinsky 2 — multilingual text2image latent diffusion model
Simple command line tool for text to image generation using OpenAI's CLIP and Siren (Implicit neural representation network). Technique was originally created by https://twitter.com/advadnoun
Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
Awesome curated collection of images and prompts generated by GPT-4o and gpt-image-1. Explore AI generated visuals created with ChatGPT and Sora, showcasing OpenAI’s advanced image generation capabilities.
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis neural network, in Pytorch
Create and customize your AI influencer open-source
Paddle Multimodal Integration and eXploration, supporting mainstream multi-modal tasks, including end-to-end large-scale multi-modal pretrain models and diffusion model toolbox. Equipped with high performance and flexibility.
CLIP + FFT/DWT/RGB = text to image/video
Diffusion model papers, survey, and taxonomy
Implementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
[NeurIPS 2025 Spotlight] A Unified Tokenizer for Visual Generation and Understanding
Python library for designing and training your own Diffusion Models with PyTorch
Recurrent Video Restoration Transformer with Guided Deformable Attention (NeurlPS2022, official repository)
official code repo for paper "CogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers"
Generate images from texts. In Russian
Workflow-to-APP、ScreenShare&FloatingVideo、GPT & 3D、SpeechRecognition&TTS
Flow is a custom node designed to provide a user-friendly interface for ComfyUI.
Awesome diffusion Video-to-Video (V2V). A collection of paper on diffusion model-based video editing, aka. video-to-video (V2V) translation. And a video editing benchmark code.
Datamosh in the browser
DiffuEraser is a diffusion model for video inpainting, which performs great content completeness and temporal consistency while maintaining acceptable efficiency.
Official Pytorch Implementation for "TokenFlow: Consistent Diffusion Features for Consistent Video Editing" presenting "TokenFlow" (ICLR 2024)
[CVPR 2025 Highlight] X-Dyna: Expressive Dynamic Human Image Animation
24,535 repositories in the index in total.