filippogiruzzi/voice_activity_detection
quality grade C, 53 out of 100Voice Activity Detection based on Deep Learning & TensorFlow
- stars
- 372
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Core deep-learning frameworks and libraries for pretraining and distributed training.
Signals: deep-learning, neural-network, pytorch, tensorflow, jax, distributed-training, training, deepspeed
2,672 results
Voice Activity Detection based on Deep Learning & TensorFlow
UniSpeech - Large Scale Self-Supervised Learning for Speech
人工智能学习资料超全整理,包含机器学习基础ML、深度学习基础DL、计算机视觉CV、自然语言处理NLP、推荐系统、语音识别、图神经网路、算法工程师面试题
Allosaurus is a pretrained universal phone recognizer for more than 2000 languages
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
Official Code for Assem-VC @ICASSP2022
WaveNet-Vocoder implementation with pytorch.
VQ-VAE for Acoustic Unit Discovery and Voice Conversion
Voice Conversion With Just Nearest Neighbors
The Implementation of FastSpeech based on pytorch.
GAN-based Mel-Spectrogram Inversion Network for Text-to-Speech Synthesis
so-vits-svc fork with realtime support, improved interface and more features.
Descriptive Deep Learning
Generic template to bootstrap your PyTorch project.
🏆 A ranked gallery of awesome streamlit apps built by the community
NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling @ INTERSPEECH 2021
NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates @ INTERSPEECH 2022
ACM Multimedia 2023: DocDiff: Document Enhancement via Residual Diffusion Models. Also contains 1597 red seals in Chinese scenes, along with their corresponding binary masks.
RealScaler - image/video AI upscaler app (Real-ESRGAN)
Learning Continuous Image Representation with Local Implicit Image Function, in CVPR 2021 (Oral)
Image Restoration Toolbox (PyTorch). Training and testing codes for DPIR, USRNet, DnCNN, FFDNet, SRMD, DPSR, BSRGAN, SwinIR
Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
[ICML 2023] The official implementation of the paper "TabDDPM: Modelling Tabular Data with Diffusion Models"
24,540 repositories in the index in total.