☆11Feb 14, 2025Updated last year
Alternatives and similar repositories for SpeechWellness-1_Baseline
Users that are interested in SpeechWellness-1_Baseline are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Models and codes for INTERSPEECH 2023 paper DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model☆13Mar 30, 2025Updated last year
- Small World of Words - Zhongwen, a Chinese word association norm☆18Apr 30, 2026Updated 3 months ago
- [EMNLP 2024 Oral] PsyGUARD: An Automated System for Suicide Detection and Risk Assessment in Psychological Counseling☆24Apr 21, 2025Updated last year
- Official repository for the paper "Audio xLSTMs: Learning Self-supervised audio representations with xLSTMs"☆21Sep 7, 2025Updated 11 months ago
- MSP-Podcast Challenge Baseline Code for Interspeech 2025☆29Dec 4, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- “莱斯杯”全国第一届“军事智能·机器阅读”挑战赛[初赛top2版]☆11Jan 15, 2019Updated 7 years ago
- Code and data recipes for the paper: Optimal Condition Training for Target Source Separation by Efthymios Tzinis, Gordon Wichern, Paris S…☆14Feb 15, 2023Updated 3 years ago
- ☆19Mar 2, 2024Updated 2 years ago
- The baselines of ARC-Challenge-Interspeech2026☆60Dec 1, 2025Updated 8 months ago
- source code for "Towards Speaker-Unknown Emotion Recognition in Conversation via Progressive Contrastive Deep Supervision"☆11Nov 22, 2024Updated last year
- We propose C2SER, a novel audio-language model designed to enhance the stability and accuracy of speech emotion recognition through conte…☆50Mar 3, 2025Updated last year
- text-only training or language-free training for multimodal tasks (image/audio/video caption, retrieval, text2image)☆13Oct 15, 2024Updated last year
- Official implementation for "Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning"☆12Jun 20, 2025Updated last year
- Official data preparation scripts for the URGENT 2024 Challenge☆90May 21, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- WavReward: Spoken Dialogue Models With Generalist Reward Evaluators☆56May 15, 2025Updated last year
- ☆47Apr 2, 2025Updated last year
- ☆29Jul 25, 2026Updated 3 weeks ago
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆14Oct 31, 2024Updated last year
- ☆11Dec 6, 2024Updated last year
- ☆35Jun 16, 2023Updated 3 years ago
- Non-parallel voice conversion called ICRCycleGAN-VC based on CycleGAN and Inception-resNet module by Afiuny☆15Apr 15, 2026Updated 4 months ago
- uyghur text resource crawled from website☆12Dec 25, 2015Updated 10 years ago
- ☆35Sep 24, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Oct 20, 2022Updated 3 years ago
- ☆18Mar 21, 2024Updated 2 years ago
- EmoLLM: Multimodal Emotional Understanding Meets Large Language Models☆19Jun 24, 2024Updated 2 years ago
- This repository contains the official implementation and pretrained weights for the paper "ReDimNet2: Scaling Speaker Verification via Ti…☆77Jul 30, 2026Updated 2 weeks ago
- ☆23May 25, 2026Updated 2 months ago
- Target speaker automatic speech recognition (TS-ASR)☆15Oct 14, 2023Updated 2 years ago
- ☆20Aug 23, 2024Updated last year
- AD-TUNING: An Adaptive CHILD-TUNING Approach to Efficient Hyperparameter Optimization of Child Networks for Speech Processing Tasks in th…☆11Feb 23, 2024Updated 2 years ago
- ☆17Apr 16, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- kaldi cnn-tdnnf baseline☆13Aug 31, 2021Updated 4 years ago
- First-principles Manim skill for Claude Code — mathematical animation from scratch, paper-explainer patterns, 21 rule files.☆28Jun 27, 2026Updated last month
- Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSI…☆25Oct 8, 2025Updated 10 months ago
- 语音切割,python ,webrtc☆11Sep 28, 2018Updated 7 years ago
- This is the pytorch\DGL implementation of the AMIGO paper.☆10Feb 6, 2024Updated 2 years ago
- Getting confidences from any end-to-end systems☆11May 24, 2023Updated 3 years ago
- Code for InterSpeech 2024 Paper: LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition☆19Jul 16, 2024Updated 2 years ago