TASU: A New Style of Alignment of Speech LLM with only Text Training Data, zero-shot on ASR and Other SU tasks
☆28Jul 20, 2026Updated last month
Alternatives and similar repositories for ps-slm
Users that are interested in ps-slm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, int…☆36Updated this week
- ☆95Jun 25, 2025Updated last year
- [EMNLP 2026' Main] HoliTok:A Continuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding☆49Jun 8, 2026Updated 2 months ago
- X-Talk is an open-source full-duplex cascaded spoken dialogue system framework enabling low-latency, interruptible, and human-like speech…☆244Aug 18, 2026Updated last week
- ☆27Dec 4, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- State-of-the-art continious audio tokenization☆42Mar 9, 2026Updated 5 months ago
- [ACL 2026 Main] Open-Ended Speaking Style Modeling via Fine-Grained and Multi-Granular Contrastive Language-Speech Pre-training☆106Apr 6, 2026Updated 4 months ago
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)