AssemblyAI-Community/MinImagen
quality grade D, 40 out of 100MinImagen: A minimal implementation of the Imagen text-to-image model
- stars
- 312
- stars gained this week
- -2
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Diffusion models, image editing, upscaling and the surrounding creative tooling.
Signals: stable-diffusion, diffusion-models, image-generation, text-to-image, generative-art, comfyui, controlnet, flux
880 results
MinImagen: A minimal implementation of the Imagen text-to-image model
Getting the latest versions of Disco Diffusion to work locally, instead of colab. Including how I run this on Windows, despite some Linux only dependencies ;)
🔥ICLR 2025 (Spotlight) One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to operate efficiently under stringent computational constraints.
Colab notebook for Stable Diffusion Hyper-SDXL.
AlignProp uses direct reward backpropogation for the alignment of large-scale text-to-image diffusion models. Our method is 25x more sample and compute efficient than reinforcement learning methods (PPO) for finetuning Stable Diffusion
Implementation of Encoder-based Domain Tuning for Fast Personalization of Text-to-Image Models
[CVPR2022 oral] A Simple and Effective Baseline for Text-to-Image Synthesis
FIBO is a SOTA, first open-source, JSON-native text-to-image model built for controllable, predictable, and legally safe image generation.
[ICLR 2025] Official Implementation of Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis
[Neurips 2023 & TPAMI] T2I-CompBench (++) for Compositional Text-to-image Generation Evaluation
Official Repository of the paper "Trajectory Consistency Distillation"
Official GitHub repository for FLUX.1 Krea [dev].
Just playing with getting CLIP Guided Diffusion running locally, rather than having to use colab.
[CVPR 2024] "MACE: Mass Concept Erasure in Diffusion Models" (Official Implementation)
End-to-end recipes for optimizing diffusion models with torchao and diffusers (inference and FP8 training).
Officail Implementation for "Cross-Image Attention for Zero-Shot Appearance Transfer"
Official implementation for "Stable Flow: Vital Layers for Training-Free Image Editing" [CVPR 2025]
An SDK/Python library for Automatic 1111 to run state-of-the-art diffusion models
attention map tools for huggingface/diffusers
[NeurIPS'23] "MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing".
[ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potential in Unified Multimodal Models through Self-supervised Learning.
Pytorch implementation of Generative Adversarial Text-to-Image Synthesis paper
Open-AI's DALL-E for large scale training in mesh-tensorflow.
24,523 repositories in the index in total.