A method that directly addresses the modality gap by aligning speech token with the corresponding text transcription during the tokenization stage.
☆119Sep 3, 2025Updated 10 months ago
Alternatives and similar repositories for TASTE-SpokenLM
Users that are interested in TASTE-SpokenLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The demo page for ALMTokenizer☆59Apr 14, 2025Updated last year
- A lightweight audio codec based on a single quantizer☆72Aug 15, 2025Updated 11 months ago
- Official Code for SyllableLM: Learning Coarse Semantic Units for Speech Language Models