Weifeng-Chen/control-a-video
quality grade F, 30 out of 100Official Implementation of "Control-A-Video: Controllable Text-to-Video Generation with Diffusion Models"
- stars
- 404
- stars gained this week
- —this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Video synthesis and editing, talking heads, avatars, animation and 3D asset generation.
Signals: video-generation, text-to-video, video-editing, animation, talking-head, avatar, deepfake, lip-sync
331 results
Official Implementation of "Control-A-Video: Controllable Text-to-Video Generation with Diffusion Models"
Evaluation of Text-to-Video Generation Models: A Dynamics Perspective[NeurIPS 2024].
Hermes skill for realistic AI video prompts for Seedance and text-to-video models.
[CVPR 2026 Highlight] High-Quality Text-to-Video Generation with Alpha Channel
World Simulator Assistant for Physics-Aware Text-to-Video Generation
A list for Text-to-Video, Image-to-Video works
SkyReels V1: The first and most advanced open-source human-centric video foundation model
[CVPR2025 Highlight] Video Generation Foundation Models: https://saiyan-world.github.io/goku/
The official implementation of CVPR'25 Oral paper "Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise"
[CVPR 2023] Executing your Commands via Motion Diffusion in Latent Space, a fast and high-quality motion diffusion model
Official implementation of "Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation"
LVDM: Latent Video Diffusion Models for High-Fidelity Long Video Generation
Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models
[ICCV 2025, Oral] TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
A collection of papers on diffusion models for 3D generation.
ViViD: Video Virtual Try-on using Diffusion Models
(CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models
MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model
Official implementation of "DreamPose: Fashion Image-to-Video Synthesis via Stable Diffusion"
🔥🔥🔥自定义Android相机(仿抖音 TikTok),其中功能包括视频人脸识别贴纸,美颜,分段录制,视频裁剪,视频帧处理,获取视频关键帧,视频旋转,添加滤镜,添加水印,合成Gif到视频,文字转视频,图片转视频,音视频合成,音频变声处理,SoundTouch,Fmod音频处理。 Android camera(imitation Tik Tok), which includes video editor,audio editor,video face recognition stickers, segment recording,video cropping, video frame processing, get the first video frame, key frame, video rotation, add filter Mirror ,add watermark ,add gif to video, add text to video, picture to video, audio and video synthesis, audio change processing, SoundTouch, Fmod audio processing.
Allegro is a powerful text-to-video model that generates high-quality videos up to 6 seconds at 15 FPS and 720p resolution from simple text input.
Industry-level video foundation model for unified Text-to-Video (T2V) and Image-to-Video (I2V) generation.
[IJCV] Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
[ICCV 2023] Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation
24,535 repositories in the index in total.