Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
| Date | Stars |
|---|---|
| 2026-07-24 | 1356 |
| 2026-07-25 | 1356 |
| 2026-07-28 | 1358 |
| 2026-07-30 | 1358 |
| 2026-08-06 | 1358 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<div align="center"> <h2><font color="red"> 🕺🕺🕺 Follow-Your-Pose 💃💃💃 </font></center> <br> <center>Pose-Guided Text-to-Video Generation using Pose-Free Videos (AAAI 2024)</h2> [Yue Ma*](https://mayuelala.github.io/), [Yingqing He*](https://github.com/YingqingHe), [Xiaodong Cun](http://vinthony.github.io/), [Xintao Wang](https://xinntao.github.io/), [Siran Chen](https://github.com/Sranc3), [Ying Shan](https://scholar.google.com/citations?hl=zh-CN&user=4oXBp9UAAAAJ), [Xiu Li](https://scholar.google.com/citations?user=Xrh1OIUAAAAJ&hl=zh-CN), and [Qifeng Chen](https://cqf.io) <a href='https://arxiv.org/abs/2304.01186'><img src='https://img.shields.io/badge/ArXiv-2304.01186-red'></a> <a href='https://follow-your-pose.github.io/'><img src='https://img.shields.io/badge/Project-Page-Green'></a> [](https://colab.research.google.com/github/mayuelala/FollowYourPose/blob/main/quick_demo.ipynb) [](https://huggingface.co/spaces/YueMafighting/FollowYourPose) [](https://openxlab.org.cn/apps/detail/houshaowei/FollowYourPose)  [](https://github.com/mayuelala/FollowYourPose) </div> <!--  --> <table class="center"> <td><img src="gif_results/new_result_0830/a_man_in_the_park.gif"></td> <td><img src="gif_results/new_result_0830/a_Iron_man_in_the_street.gif"></td> <tr> <td width=25% style="text-align:center;">"The man is sitting on chair, on the park"</td> <td width=25% style="text-align:center;">"The Iron man, on the street "</td> <!-- <td width=25% style="text-align:center;">"Wonder Woman, wearing a cowboy hat, is skiing"</td> <td width=25% style="text-align:center;">"A man, wearing pink clothes, is skiing at sunset"</td> --> </tr> <td><img src="gif_results/new_result_0830/a_strom.gif"></td> <td><img src="gif_results/new_result_0830/a_astronaut_cartoon.gif"></td> <tr> <td width=25% style="text-align:center;">"The stormtrooper, in the gym "</td> <td width=25% style="text-align:center;">"The astronaut, earth background, Cartoon Style "</td> </tr> </table > ## 💃💃💃 Demo Video https://github.com/mayuelala/FollowYourPose/assets/38033523/e021bce6-b9bd-474d-a35a-7ddff4ab8e75 ## 💃💃💃 Abstract <b>TL;DR: We tune the text-to-image model (e.g., stable diffusion) to generate the character videos from pose and text description.</b> <details><summary>CLICK for full abstract</summary> > Generating text-editable and pose-controllable character videos have an imperious demand in creating various digital human. Nevertheless, this task has been restricted by the absence of a comprehensive dataset featuring paired video-pose captions and the generative prior models for videos. In this work, we design a novel two-stage training scheme that can utilize easily obtained datasets (i.e., image pose pair and pose-free video) and the pre-trained text-to-image (T2I) model to obtain the pose-controllable character videos. Specifically, in the first stage, only the keypoint-image pairs are used only for a controllable textto-image generation. We learn a zero-initialized convolutional encoder to encode the pose information. In the second stage, we finetune the motion of the above network via a pose-free video dataset by adding the learnable temporal self-attention and reformed cross-frame self-attention blocks. Powered by our new designs, our method successfully generates continuously pose-controllable character videos while keeps the editing and concept composition ability of the pre-trained T2I model. The code and models will be made publicly available. </details> ## 🕺🕺🕺
Excerpt of 17,965 characters
Read on GitHubmayuema · HKUST · Hong Kong
41
Xiaodong Cun · GVC Lab, Great Bay University · China
7
Joy · HKUST
4
Ikko Eltociear Ashimine · Japan
1
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:20e84fd51354c22f, topic:video-generation, desc:video generation, readme:video generation