AVSpeech downloader
☆69Jan 30, 2019Updated 7 years ago
Alternatives and similar repositories for avspeech-downloader
Users that are interested in avspeech-downloader are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Luigi pipeline to download VoxCeleb(2) audio from YouTube and extract speaker segments☆43Mar 29, 2021Updated 5 years ago
- Simple python script for downloading AVSpeech Dataset☆47Mar 16, 2024Updated 2 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Include some core functions and model to handle speech separation☆156Jun 24, 2021Updated 5 years ago
- Looking to listen at cocktail party☆36Mar 24, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆10Mar 24, 2023Updated 3 years ago
- CHiME-5 Baseline Array Synchronisation☆12Sep 24, 2018Updated 7 years ago
- ☆16Apr 27, 2025Updated last year
- Executable code based on Google articles☆164Dec 8, 2022Updated 3 years ago
- A python implementation of a traditional Dynamic Range Compressor☆14Oct 30, 2020Updated 5 years ago
- Text Recognition and Detection based on Pixel-Link paper implemented in pytorch☆28May 30, 2023Updated 3 years ago
- Code for "Vid2speech: Speech Reconstruction from Silent Video" ICASSP '17☆115Feb 15, 2017Updated 9 years ago
- Out of time: automated lip sync in the wild☆897Apr 17, 2026Updated 4 months ago
- converting the pretrained tensorflow SoundNet model to pytorch☆14Jun 15, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Pytorch implemention of SDNet☆23Jun 1, 2021Updated 5 years ago
- mirror of VoxCeleb dataset - a large-scale speaker identification dataset☆78Jul 5, 2019Updated 7 years ago
- PyTorch implementation of WASE described in our ICASSP 2021: "Wase: Learning When to Attend for Speaker Extraction in Cocktail Party Envi…☆27Jan 11, 2022Updated 4 years ago
- ☆14Aug 10, 2015Updated 11 years ago
- Face Landmark-based Speaker-Independent Audio-Visual Speech Enhancement in Multi-Talker Environments☆112Mar 19, 2024Updated 2 years ago
- A unofficial Pytorch implementation of Microsoft's PHASEN☆235Apr 10, 2024Updated 2 years ago
- Speech synthesis using LPC☆25Jun 5, 2021Updated 5 years ago
- VGGVox models for Speaker Identification and Verification trained on the VoxCeleb (1 & 2) datasets☆402Feb 4, 2019Updated 7 years ago
- Code for the paper: Audio-Visual Scene Analysis with Self-Supervised Multisensory Features☆225Jul 17, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆53May 15, 2025Updated last year
- LinuxShell编程笔记☆15Aug 29, 2017Updated 8 years ago
- ☆19Apr 1, 2020Updated 6 years ago
- An STFT/iSTFT for PyTorch.☆373Oct 31, 2023Updated 2 years ago
- Problem Agnostic Speech Encoder☆446Jul 6, 2023Updated 3 years ago
- A minimum unofficial implementation of the "A Convolutional Recurrent Neural Network for Real-Time Speech Enhancement" (CRN) using PyTorc…☆350Sep 5, 2020Updated 5 years ago
- a standalone pitch extractor☆13Oct 19, 2017Updated 8 years ago
- Learn and L3 embedding from audio/video pairs☆89Apr 24, 2022Updated 4 years ago
- SMS-WSJ: Spatialized Multi-Speaker Wall Street Journal database for multi-channel source separation and recognition☆131Jun 7, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICML 2020] "AutoGAN-Distiller: Searching to Compress Generative Adversarial Networks" by Yonggan Fu, Wuyang Chen, Haotao Wang, Haoran Li…☆105Dec 30, 2021Updated 4 years ago
- This is the code&dataset for our paper [Modeling Attention and Memory for Auditory Selection in a Cocktail Party Environment. AAAI 2018]☆57Apr 12, 2018Updated 8 years ago
- Code for CVPR 2021 paper Exploring Heterogeneous Clues for Weakly-Supervised Audio-Visual Video Parsing☆24Dec 29, 2021Updated 4 years ago
- Unofficial implementation of music separation model by Luo et.al.☆13Nov 3, 2019Updated 6 years ago
- Abstraction and Reasoning Corpus☆15Nov 22, 2022Updated 3 years ago
- Loader for Google facial expression comparison dataset☆13Feb 26, 2019Updated 7 years ago
- 🎵 muse: Music Separation☆11Feb 14, 2024Updated 2 years ago