Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A list for Text-to-Video, Image-to-Video works
| Date | Stars |
|---|---|
| 2026-07-31 | 255 |
| 2026-08-05 | 255 |
| 2026-08-06 | 255 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# Awesome-Text-to-Video-Generation [](https://github.com/sindresorhus/awesome)
A curated (continually updated) list of Text-to-Video studies. It's based on our survey paper: [From Sora What We Can See: A Survey of Text-to-Video Generation](https://arxiv.org/pdf/2405.10674).
In this survey, We have conducted a comprehensive exploration of existing works in the Text-to-Video field using OpenAI’s Sora as a clue, and we have also summarized 24 datasets and 9 evaluation metrics in this field. Specifically, we discussed the problems existing in this research area and Sora itself, combined with the advantages of Sora and the characteristics of related fields to provide future research directions. If our work can inspire you, feel free to cite our paper and star our repo.
**This project is curated and maintained by [Rui Sun](https://github.com/ray-ruisun) and [Yumin Zhang](https://github.com/zymvszym).**
```
@article{sun2024sora,
title={From Sora What We Can See: A Survey of Text-to-Video Generation},
author={Sun, Rui and Zhang, Yumin and Shah, Tejal and Sun, Jiahao and Zhang, Shuoying and Li, Wenqi and Duan, Haoran and Wei, Bo and Ranjan, Rajiv},
journal={arXiv preprint arXiv:2405.10674},
year={2024}
}
```


> Topics of this repo cover: <br>
> `Text-to-Seq-Image`, `Text-to-Video`
## Table of Content
* [Text-to-Seq-Image](#text_to_seq_image)
* [Text-to-Video](#text_to_video)
* [Dataset & Metrics](#dataset_and_metrics)
## <a name="text_to_seq_image"></a> Text-to-Seq-Image
- **LivePhoto: Real Image Animation with Text-guided Motion Control** <br>
Team: HKU, Alibaba Group, Ant Group. <br>
*Xi Chen, Zhiheng Liu, Mengting Chen, et al., Hengshuang Zhao* <br>
arXiv, 2023.12 [[Paper](https://arxiv.org/abs/2312.02928)], [[PDF](https://arxiv.org/pdf/2312.02928.pdf)], [[Code](https://github.com/XavierCHEN34/LivePhoto)], [[Demo (Video)](https://www.youtube.com/watch?v=M2vzrTYAsQI)], [[Home Page](https://xavierchen34.github.io/LivePhoto-Page/)] <br>
- **Scalable Diffusion Models with Transformers** `Sequential Images` <br>
Team: UC Berkeley, NYU. <br>
*William Peebles, Saining Xie* <br>
**ICCV'23(Oral)**, arXiv, 2022.12 [[Paper](https://arxiv.org/abs/2212.09748)], [[PDF](https://arxiv.org/pdf/2212.09748.pdf)], [[Code](https://github.com/facebookresearch/DiT)], [[Pretrained Model](https://github.com/facebookresearch/DiT)], [[Home Page](https://www.wpeebles.com/DiT.html)] <br>
## <a name="text_to_video"></a> Text-to-Video
- **Helios: Real Real-Time Long Video Generation Model**<br>
Team: Peking University, ByteDance. <br>
*Shenghai Yuan, Yuanyang Yin, et al., Li Yuan* <br>
arXiv, 2025.03 [[Paper](https://arxiv.org/abs/2603.04379)], [[PDF](https://arxiv.org/pdf/2603.04379)], [[Code](https://github.com/PKU-YuanGroup/Helios)], [[Home Page](https://pku-yuangroup.github.io/Helios-Page)]
- **Identity-Preserving Text-to-Video Generation by Frequency Decomposition**<br>
Team: Peking University, Peng Cheng Laboratory. <br>
*Shenghai Yuan, Jinfa Huang, Xianyi He, et al., Li Yuan* <br>
arXiv, 2024.11 [[Paper](https://arxiv.org/abs/2411.17440)], [[PDF](https://arxiv.org/pdf/2411.17440)], [[Code](https://github.com/PKU-YuanGroup/ConsisID)], [[Home Page](https://pku-yuangroup.github.io/ConsisID/)]
- **Zero-Shot Video Editing through Adaptive Sliding Score Distillation** `Video Editing` <br>
Team: Nanjing University. <br>
*Lianghan Zhu, Yanqi Bao, Jing Huo, et al., Yang Gao* <br>
arXiv, 2024.06 [[Paper](https://arxiv.org/abs/2406.04888)], [[PDF](https://arxiv.org/pdf/2406.04888)], [[Home Page](https://nips24videoedit.github.io/zeroshot_videoedit/)] <br>
- **CoNo: Consistency Noise Injection for Tuning-free Long Video Diffusion** <br>
Team: University of Science and Technology of China. <br>
*Xingrui Wang, Xin Li, Zhibo Chen* <br>
arXiv, 2024.06 [[Paper](htExcerpt of 72,407 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:4d75ee89b6c589cf, name:video generation, name:text-to-video, desc:text-to-video