Split audio using the .srt file, clean up annotations, then merge and package into a format suitable for bert-vits2 in a standard manner. 使用.srt文件分割音频并清洗标注,合并封装至适用于bert-vits2的一个较为标准的格式
☆50Jun 17, 2024Updated 2 years ago
Alternatives and similar repositories for Alice_split_toolset
Users that are interested in Alice_split_toolset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SubFix: Efficient Web-Based Audio Subtitle Editing and Multilingual Automatic Annotation Tool.☆207Feb 5, 2024Updated 2 years ago
- Simple data labeling script with funasr inside. 使用阿里fanasr进行VITS训练数据标注☆80Oct 10, 2023Updated 2 years ago
- BertVITS2前端界面☆304Jan 1, 2024Updated 2 years ago
- vits2 backbone with bert☆333Apr 13, 2024Updated 2 years ago
- cpp rotation album,基于cpp eigen实现的3d旋转相册,GAMES101复现内容☆12Jul 25, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Cloned from https://huggingface.co/spaces/aadnk/faster-whisper-webui, and add text post-processing☆27Oct 18, 2024Updated last year
- vits2 backbone with bert☆80Jan 8, 2024Updated 2 years ago
- C++版本的sort算法,可无缝添加在检测器后进行实时多目标跟踪☆12Dec 1, 2022Updated 3 years ago
- 基于中文文本情绪分析自动切换参考音频的 GPT-SoVITS 推理 Demo☆108Mar 8, 2024Updated 2 years ago
- An Prompt Collecton Project which provide a more helpful prompts database☆26Mar 30, 2025Updated last year
- vits2 backbone with multilingual-bert☆8,790Updated this week
- Brand new TTS solution☆11Dec 7, 2024Updated last year
- 基于自回归模型与现有的开源大模型,训练小说大模型☆41Oct 9, 2023Updated 2 years ago
- A voiceprint recognition classifier for audio dataset☆105Jun 21, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Bert-vits2-V2.3 训练和推理☆49Mar 13, 2024Updated 2 years ago
- 通过将APNIC发布的最新国内ip段列表加入到本地路由,实现国内ip段不通过VPN请求互联网☆10Mar 30, 2016Updated 10 years ago
- ☆33Dec 23, 2023Updated 2 years ago
- This project uses gpt-4 to build agents to play one night werewolf.☆10Jul 14, 2023Updated 3 years ago
- Psyche AI Inc release source "CVCUDA_FaceStoreHelper"☆67Jul 14, 2023Updated 3 years ago
- a compact audio-to-phoneme aligner for singing voice☆12Jan 17, 2024Updated 2 years ago
- 日本語N2ワード(新しい標準日本語初級および中級)☆15Dec 29, 2025Updated 7 months ago
- Hamibot钉钉打卡脚本☆12Sep 19, 2022Updated 3 years ago
- 基于GPT-SoVITS的AI读弹幕姬☆25Sep 30, 2025Updated 10 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆17Nov 7, 2023Updated 2 years ago
- ☆10Nov 19, 2023Updated 2 years ago
- Bert-VITS2 onnx推理版本☆44Apr 24, 2024Updated 2 years ago
- Understanding ComfyUI seed☆16May 25, 2024Updated 2 years ago
- Supplementary materials for "Evaluating generalised additive mixed modelling strategies for dynamic speech analysis"☆10Jan 25, 2021Updated 5 years ago
- AITuberのデモリポジトリです☆10Mar 11, 2023Updated 3 years ago
- Interpretability analysis of language model outlier and attempts to distill the model☆13May 8, 2023Updated 3 years ago
- 使用 Python、Streamlit 和 Hugging Face 模型,构建无需 API 令牌的 AI 故事机,应用根据上传的图片创建音频故事。☆11Aug 5, 2023Updated 3 years ago
- A Python client for Deepgram's Voice Agent API☆11Oct 14, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- PFCC 社区博客☆14Updated this week
- ☆17Jun 14, 2024Updated 2 years ago
- ☆10Nov 6, 2019Updated 6 years ago
- A tool for analysing continuous glucose monitoring (CGM) data in epidemiology.☆15Feb 1, 2022Updated 4 years ago
- Use Stable Diffusion intrinsic lora to render texture maps (normal, albedo, shade, depth)☆17Mar 10, 2024Updated 2 years ago
- The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generation…☆13Dec 23, 2023Updated 2 years ago
- A lightweight tool that efficiently isolates target speaker data from your datasets.☆19Nov 23, 2024Updated last year