garibida/cross-image-attention
quality grade D, 47 out of 100Officail Implementation for "Cross-Image Attention for Zero-Shot Appearance Transfer"
- stars
- 404
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Diffusion models, image editing, upscaling and the surrounding creative tooling.
Signals: stable-diffusion, diffusion-models, image-generation, text-to-image, generative-art, comfyui, controlnet, flux
879 results
Officail Implementation for "Cross-Image Attention for Zero-Shot Appearance Transfer"
Official implementation for "Stable Flow: Vital Layers for Training-Free Image Editing" [CVPR 2025]
An SDK/Python library for Automatic 1111 to run state-of-the-art diffusion models
[NeurIPS'23] "MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing".
[ICLR 2026] RecA: visual understanding help generation through self-supervised learning
Pytorch implementation of Generative Adversarial Text-to-Image Synthesis paper
Open-AI's DALL-E for large scale training in mesh-tensorflow.
🎨 精选 3000+ Gemini Nano Banana Pro 高质量提示词与生成案例 | 涵盖摄影、设计、艺术、营销等多领域 | 双语支持 | JSON 格式
An unified model that seamlessly integrates multimodal understanding, text-to-image generation, and image editing within a single powerful framework.
Official implementation of AsymFlow, pi-Flow, GMFlow
A CLI tool/python module for generating images from text using guided diffusion and CLIP from OpenAI.
StyleShot: A SnapShot on Any Style. 一款可以迁移任意风格到任意内容的模型,无需针对图片微调,即能生成高质量的个性风格化图片!
LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models (LLM-grounded Diffusion: LMD, TMLR 2024)
Papers and resources on Controllable Generation using Diffusion Models, including ControlNet, DreamBooth, IP-Adapter.
Implementation of Parti, Google's pure attention-based text-to-image neural network, in Pytorch
T2F: text to face generation using Deep Learning
Official implementation for "Blended Diffusion for Text-driven Editing of Natural Images" [CVPR 2022]
Official code for the CVPR 2025 paper "SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models."
Generative Adversarial Text to Image Synthesis / Please Star -->
(Accepted by IJCV) Liquid: Language Models are Scalable and Unified Multi-modal Generators
face-to-sticker
基于Stable Diffusion优化的AI绘画模型。支持输入中英文文本,可生成多种现代艺术风格的高质量图像。| An optimized text-to-image model based on Stable Diffusion. Both Chinese and English text inputs are available to generate images. The model can generate high-quality images in several modern art styles.
[ICCV 2023] A latent space for stochastic diffusion models
Flash Diffusion — accelerating conditional diffusion models (AAAI 2025 Oral)
24,535 repositories in the index in total.