haoningwu3639/StoryGen
quality grade D, 45 out of 100[CVPR 2024] Intelligent Grimm - Open-ended Visual Storytelling via Latent Diffusion Models
- stars
- 268
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Diffusion models, image editing, upscaling and the surrounding creative tooling.
Signals: stable-diffusion, diffusion-models, image-generation, text-to-image, generative-art, comfyui, controlnet, flux
879 results
[CVPR 2024] Intelligent Grimm - Open-ended Visual Storytelling via Latent Diffusion Models
Stochastic Adversarial Video Prediction
[NeurIPS 2024]OmniTokenizer: one model and one weight for image-video joint tokenization.
📚 Collection of awesome generation acceleration resources.
An open source code repository of driving world models, with training, inferencing, evaluation tools, and pretrained checkpoints.
Official Code for DiffMorpher: Unleashing the Capability of Diffusion Models for Image Morphing (CVPR 2024)
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
[ICLR 2025] Autoregressive Video Generation without Vector Quantization
A Collection of Papers and Codes for CVPR2026/CVPR2025/ICCV2025/CVPR2024/ECCV2026/ECCV2024 AIGC
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation
[ICML 2025] Official PyTorch Implementation of "History-Guided Video Diffusion"
Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025) , UltraViCo (ICLR 2026) and UltraImage
[ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
[ ICLR 2024 ] Official Codebase for "InstructCV: Instruction-Tuned Text-to-Image Diffusion Models as Vision Generalists"
Open-source SOTA multi-image editing model
This is a Go language version of the SDK based on stable-diffusion-webui. In your code, you can directly use the API interfaces of stable-diffusion-webui through object-oriented operations, instead of dealing with cumbersome JSON. Support extensions API !
DeOldify for Stable Diffusion WebUI:This is an extension for StableDiffusion's AUTOMATIC1111 web-ui that allows colorize of old photos and old video. It is based on deoldify.
online 3d openpose editor for stable diffusion and controlnet
Beautiful and Easy to use Stable Diffusion WebUI
Deforum extension for AUTOMATIC1111's Stable Diffusion webui
a self-hosted webui for 30+ generative ai
24,535 repositories in the index in total.