这个工程的目的是从视频中获取语音识别的训练数据,用于训练字幕自动生成
☆53Aug 5, 2018Updated 8 years ago
Alternatives and similar repositories for from_video_get_ASR_traindata
Users that are interested in from_video_get_ASR_traindata are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using Baidu ASR auto-generating subtitles for any video file. 使用百度短语音识别技术为视频或音频生成字幕。☆12Jan 23, 2022Updated 4 years ago
- Python implementation of CTC beam search decoder + agnostic LM scorer☆20Dec 16, 2020Updated 5 years ago
- Minimize kaldi nnet3 chain decoder☆45Jan 10, 2020Updated 6 years ago
- 利用文字信息生成文字动画视频☆17Apr 14, 2022Updated 4 years ago
- Code for end-to-end ASR with neural networks, build with TensorFlow☆110Jan 24, 2019Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CS224S / LINGUIST285 - Spoken Language Processing☆24Feb 13, 2020Updated 6 years ago
- A TensorFlow implementation for Chinese speech recognition based on DeepMind's WaveNet☆15Mar 27, 2018Updated 8 years ago
- 基于卷积神经网络的语音识别声学模型的研究☆180Jul 22, 2019Updated 7 years ago
- Chinese Prosodic Structure Prediction☆10May 18, 2019Updated 7 years ago
- 📖 LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.☆22Jul 12, 2019Updated 7 years ago
- ☆14Jun 19, 2019Updated 7 years ago
- it's a train acoustics model code lib☆27May 20, 2020Updated 6 years ago
- MultiSpeaker Tacotron2 using LifeLong Learning.☆13Sep 27, 2019Updated 7 years ago
- implementing beamforming algorithm in C++☆11Jan 9, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Tools for ASR Corpus Generation from Online Video☆140Feb 10, 2019Updated 7 years ago
- implementation of "EFFICIENT KEYWORD SPOTTING USING DILATED CONVOLUTIONS AND GATING"☆38Dec 8, 2019Updated 6 years ago
- python wrap for hts engine☆14Jan 30, 2018Updated 8 years ago
- 将normalize过的中文文本,做逆向normalize。具体功能即实现 chinese_text_normalization的逆向版本。☆13Apr 7, 2021Updated 5 years ago
- In this repository, I try to combine k2 with speechbrain to decode well and fastly.☆16Jun 17, 2022Updated 4 years ago
- Complex Neural Beamformer☆34Oct 15, 2020Updated 5 years ago
- Phone-level evaluation of L2 speakers (GOP algorithm)☆27Mar 1, 2017Updated 9 years ago
- a kws demo on android☆40May 28, 2024Updated 2 years ago
- Tensorflow version of DFSMN☆49Jul 17, 2018Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- An Automatic Speech Recognition Frame ,一个中文语音识别的完整框架, 提供了多个模型☆252Jan 6, 2021Updated 5 years ago
- Chinese Text Normalization and Dataset☆91May 14, 2022Updated 4 years ago
- Papers about the trend of Entity Linking in recent years.☆11Sep 5, 2022Updated 4 years ago
- ☆12May 12, 2016Updated 10 years ago
- The implementation of LSTM with projection layer by PyTorch☆17Sep 1, 2019Updated 7 years ago
- 基于cnn的视频动作分类教程。☆11Aug 17, 2020Updated 6 years ago
- tensorflow speech synthesis c++ inference for voicenet☆16Mar 29, 2019Updated 7 years ago
- ☆277Jan 15, 2021Updated 5 years ago
- Experiment with JNI access to some Kaldi functions.☆12Dec 31, 2018Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Aug 13, 2019Updated 7 years ago
- implement end-to-end asr algorithm with tensorflow☆40Aug 23, 2018Updated 8 years ago
- speex aec kalman filter☆17Mar 17, 2024Updated 2 years ago
- Next word prediction based on N-gram language model☆11Jan 11, 2015Updated 11 years ago
- https://github.com/ARM-software/ML-KWS-for-MCU☆16Jul 8, 2018Updated 8 years ago
- Camera Monitor | 深度学习 视频监控 钉钉预警 人物识别 动作识别☆12Dec 11, 2019Updated 6 years ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 11 months ago