Shilin-LU/TF-ICON
quality grade D, 47 out of 100[ICCV 2023] "TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition" (Official Implementation)
- stars
- 814
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Diffusion models, image editing, upscaling and the surrounding creative tooling.
Signals: stable-diffusion, diffusion-models, image-generation, text-to-image, generative-art, comfyui, controlnet, flux
880 results
[ICCV 2023] "TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition" (Official Implementation)
Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)
Text2Room generates textured 3D meshes from a given text prompt using 2D text-to-image models (ICCV2023).
CogView4, CogView3-Plus and CogView3(ECCV 2024)
A collection of resources on controllable generation with text-to-image diffusion models.
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
Turn any face into a video game character, pixel art, claymation, 3D or toy
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
[ECCV 2024] The official implementation of paper "BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion"
Text-to-Image generation. The repo for NeurIPS 2021 paper "CogView: Mastering Text-to-Image Generation via Transformers".
[ICML 2024] Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (RPG)
AI magics meet Infinite draw board.
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
A simple command line tool for text to image generation, using OpenAI's CLIP and a BigGAN. Technique was originally created by https://twitter.com/advadnoun
Just playing with getting VQGAN+CLIP running locally, rather than having to use colab.
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
A playground to generate images from any text prompt using Stable Diffusion (past: using DALL-E Mini)
Kandinsky 2 — multilingual text2image latent diffusion model
LTX-Video Support for ComfyUI
Simple command line tool for text to image generation using OpenAI's CLIP and Siren (Implicit neural representation network). Technique was originally created by https://twitter.com/advadnoun
Red Ink - A one-stop Xiaohongshu image-and-text generator based on the 🍌Nano Banana Pro🍌, "One Sentence, One Image: Generate Xiaohongshu Text and Images."
Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
24,523 repositories in the index in total.