ONNX-compatible Fast SeamlessM4T—Massively Multilingual & Multimodal Machine Translation
☆43Aug 29, 2023Updated 2 years ago
Alternatives and similar repositories for Fast-SeamlessM4T-ONNX
Users that are interested in Fast-SeamlessM4T-ONNX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ONNX-compatible DocShadow: High-Resolution Document Shadow Removal. Supports TensorRT 🚀☆25Sep 13, 2023Updated 2 years ago
- Belief Revision based Caption Re-ranker with Visual Semantic Information. COLING 2022☆11Apr 13, 2025Updated last year
- A python wrapper for kaldi-online-decoder using Cython☆12Sep 1, 2017Updated 8 years ago
- Chinese and English Bilinguish G2P☆22Jul 16, 2023Updated 3 years ago
- ☆16Aug 23, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆13Apr 9, 2021Updated 5 years ago
- https://wavelandspeech.github.io/☆10Jan 12, 2024Updated 2 years ago
- ☆18Jul 22, 2024Updated 2 years ago
- FunASR安卓端侧离线版本2pass全模式☆15Sep 4, 2023Updated 2 years ago
- tensorflow speech synthesis c++ inference for voicenet☆16Mar 29, 2019Updated 7 years ago
- Full featured web server for TOR Hidden Services with Vanguards, NGINX, PHP-FPM, MariaDB, NYX, Supervisor and dnsmasq. One Container for …☆13Aug 12, 2022Updated 3 years ago
- Keyword Search Recipe for Subword ASR☆30Jul 12, 2019Updated 7 years ago
- An embeddable widget for interacting with openAI api compatable LLM's☆15Sep 18, 2024Updated last year
- ChatTTS is a generative speech model for daily dialogue.☆14Oct 21, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 一个类似llama_index的极简GO版本框架(Qdrant + embedding + openai + gin),本地知识库QA应用后端☆11Jul 12, 2024Updated 2 years ago
- Translated vocal synthesis - Clone a voice and output speech in another language☆26May 3, 2022Updated 4 years ago
- 基于DINet的推理服务,推理视频流和视频☆17Nov 8, 2023Updated 2 years ago
- Tools for convert Text to IPA in python☆19Feb 11, 2023Updated 3 years ago
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- A text to speech web application that speaks word, sentences or even long articles in a music player like interface.☆10Feb 15, 2025Updated last year
- A fork of LatinIME (by Google for Android), targeting marginalised languages that also deserve first-class status on mobile operating sys…☆14Jul 3, 2026Updated 3 weeks ago
- ☆21Jun 25, 2026Updated last month
- ☆25Jan 2, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Unofficial GraphQL Reverse Proxy Server for nHentai written in Rust☆12May 8, 2022Updated 4 years ago
- A pipeline to isolate and transcribe one language in mixed-language speech☆20Oct 25, 2022Updated 3 years ago
- This repository is the implementation of the paper, "Score-balanced Loss for Multi-aspect Pronunciation Assessment" (Interspeech 2023).☆22Apr 29, 2024Updated 2 years ago
- A test validator repo that includes just the regex validator☆15Mar 3, 2026Updated 4 months ago
- AD-TUNING: An Adaptive CHILD-TUNING Approach to Efficient Hyperparameter Optimization of Child Networks for Speech Processing Tasks in th…☆11Feb 23, 2024Updated 2 years ago
- kaldi cnn-tdnnf baseline☆13Aug 31, 2021Updated 4 years ago
- Real-time collaborative kanban board web application.☆16Mar 20, 2024Updated 2 years ago
- ☆21Oct 7, 2020Updated 5 years ago
- This repository contains an asynchronous image processing service built using Golang, Asynq, Redis, Fiber and Docker Compose for easy dep…☆10Dec 13, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official code release for "RTFS-Net: Recurrent time-frequency modelling for efficient audio-visual speech separation", accepted ICLR 2024☆51Oct 14, 2025Updated 9 months ago
- ☆25Jun 14, 2022Updated 4 years ago
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 6 years ago
- 🆔 UUID InBrowser.App is a tool to generate and decode UUIDs. Fully runs in your browser, no data is sent to the server. Fast, secure, an…☆10Nov 13, 2023Updated 2 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- Chinese text normalization. 中文文本规范化。☆60May 3, 2021Updated 5 years ago