Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Recent Transformer-based CV and related works.
| Date | Stars |
|---|---|
| 2026-07-24 | 1345 |
| 2026-07-25 | 1345 |
| 2026-07-28 | 1344 |
| 2026-07-30 | 1344 |
| 2026-08-06 | 1344 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# Transformer-in-Vision Recent Transformer-based CV and related works. Welcome to comment/contribute! The transformer is now a basic component, adopted in nearly all AI models. Keep updated --> updated irregularly. New Hope: [LLM-in-Vision](https://github.com/DirtyHarryLYL/LLM-in-Vision) ## Resource - **ChatGPT** for **Robotics**: Design Principles and Model Abilities, [[Paper]](https://www.microsoft.com/en-us/research/uploads/prod/2023/02/ChatGPT___Robotics.pdf), [[Code]](https://github.com/microsoft/PromptCraft-Robotics) - DIFFUSIONDB [[Page]](https://poloclub.github.io/diffusiondb), [[Paper]](https://arxiv.org/pdf/2210.14896.pdf) - LAION-5B [[Page]](https://laion.ai/laion-5b-a-new-era-of-open-large-scale-multi-modal-datasets/), [[Paper]](https://arxiv.org/pdf/2210.08402.pdf) - LAVIS [[Page]](https://github.com/salesforce/LAVIS), [[Paper]](https://arxiv.org/pdf/2209.09019.pdf) - Imagen Video [[Page]](https://imagen.research.google/video/), [[Paper]](https://imagen.research.google/video/paper.pdf) - Phenaki [[Page]](https://phenaki.video/), [[Paper]](https://openreview.net/pdf?id=vOEXS39nOF) - DREAMFUSION [[Page]](https://dreamfusion3d.github.io/), [[Paper]](https://arxiv.org/pdf/2209.14988.pdf) - MAKE-A-VIDEO [[Page]](https://make-a-video.github.io/), [[Paper]](https://arxiv.org/pdf/2209.14792.pdf) - Stable Difffusion [[Page]](https://ommer-lab.com/research/latent-diffusion-models/), [[Paper]](https://arxiv.org/pdf/2112.10752.pdf) - NUWA-Infinity [[Page]](https://nuwa-infinity.microsoft.com/#/), [[Paper]](https://arxiv.org/pdf/2207.09814.pdf) - Parti [[Page]](https://parti.research.google/), [[Code]](https://github.com/google-research/parti) - Imagen [[Page]](https://imagen.research.google/), [[Paper]](https://arxiv.org/pdf/2205.11487.pdf) - Gato: A Generalist Agent, [[Paper]](https://storage.googleapis.com/deepmind-media/A%20Generalist%20Agent/Generalist%20Agent.pdf) - PaLM: Scaling Language Modeling with Pathways, [[Paper]](https://arxiv.org/pdf/2204.02311.pdf) - DALL·E 2 [[Page]](https://openai.com/dall-e-2/), [[Paper]](https://cdn.openai.com/papers/dall-e-2.pdf) - SCENIC: A JAX Library for Computer Vision Research and Beyond, [[Code]](https://github.com/google-research/scenic) - V-L joint learning study (with good tables): [[METER]](https://arxiv.org/pdf/2111.02387.pdf), [[Kaleido-BERT]](https://arxiv.org/pdf/2103.16110.pdf) - Attention is all you need, [[Paper]](https://arxiv.org/pdf/1706.03762.pdf) - CLIP [[Page]](https://openai.com/blog/clip/), [[Paper]](https://cdn.openai.com/papers/Learning_Transferable_Visual_Models_From_Natural_Language_Supervision.pdf), [[Code]](https://github.com/openai/CLIP), [[arXiv]](https://arxiv.org/pdf/2103.00020.pdf) - DALL·E [[Page]](https://openai.com/blog/dall-e/), [[Code]](https://github.com/openai/DALL-E), [[Paper]](https://arxiv.org/pdf/2102.12092.pdf) - [huggingface/transformers](https://github.com/huggingface/transformers) - [Kyubyong/transformer](https://github.com/Kyubyong/transformer), TF - [jadore801120/attention-is-all-you-need-pytorch](https://github.com/jadore801120/attention-is-all-you-need-pytorch), Torch - [krasserm/fairseq-image-captioning](https://github.com/krasserm/fairseq-image-captioning) - [PyTorch Transformers Tutorials](https://github.com/abhimishra91/transformers-tutorials) - [ictnlp/awesome-transformer](https://github.com/ictnlp/awesome-transformer) - [basicv8vc/awesome-transformer](https://github.com/basicv8vc/awesome-transformer) - [dk-liang/Awesome-Visual-Transformer](https://github.com/dk-liang/Awesome-Visual-Transformer) - [yuewang-cuhk/awesome-vision-language-pretraining-papers](https://github.com/yuewang-cuhk/awesome-vision-language-pretraining-papers) ## Survey - (arXiv 2023.2) TRANSFORMER-BASED **SENSOR FUSION** FOR **AUTONOMOUS DRIVING**: A SURVEY, [[Paper]](https://arxiv.org/pdf/2302.11481.pdf), [[Page]](https://github.com/ApoorvRoboticist/Transformers-Sensor-Fusion) - (arXiv 2023.2) Deep Learning for **Video-Text
Excerpt of 460,658 characters
Read on GitHub827
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:d2ec33ce880f51af, topic:deep-learning, readme:pretraining
matched fp:d2ec33ce880f51af, topic:transformer
matched fp:d2ec33ce880f51af, topic:computer-vision, readme:computer vision