Research code for "Towards multi-task learning of speech and speaker recognition" at https://arxiv.org/pdf/2302.12773.pdf
☆12Dec 2, 2024Updated last year
Alternatives and similar repositories for disjoint-mtl
Users that are interested in disjoint-mtl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the EMNLP paper "Improving Detection and Categorization of Task-relevant Utterances through Integration of Discourse Structure a…☆12Nov 23, 2022Updated 3 years ago
- Broadcasted Residual Learning for Efficient Keyword Spotting☆24Jul 9, 2021Updated 5 years ago
- x86汇编语言:从实模式到保护模式_章节源码及检测题答案☆13Aug 13, 2020Updated 5 years ago
- ☆12Jun 14, 2024Updated 2 years ago
- Deformable Speech Transformer (DST)☆35Aug 8, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A dataset of real-world robocall audio recordings☆15Jul 25, 2024Updated 2 years ago
- DO with Terraform and Ansible☆11Jun 5, 2018Updated 8 years ago
- ☆12Oct 19, 2020Updated 5 years ago
- Models and codes for INTERSPEECH 2023 paper DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model☆13Mar 30, 2025Updated last year
- RDLINet: A Novel Lightweight Inception Network for Respiratory Disease Classification Using Lung Sounds (IEEE TIM-2024)☆11Mar 24, 2025Updated last year
- Learning Domain-Invariant Transformation for Speaker Verification.☆11Jun 13, 2023Updated 3 years ago
- ☆14Jun 21, 2022Updated 4 years ago
- THU实验课实验报告模板与数据处理工具整理☆19Dec 15, 2023Updated 2 years ago
- ☆10Dec 22, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆18Feb 11, 2025Updated last year
- Score Normalization for NIST 2019 Speaker Recognition Evaluation☆10Nov 8, 2019Updated 6 years ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- [TCSVT'22] Official Implementation of STI-VQA☆12Oct 18, 2023Updated 2 years ago
- ☆17Oct 24, 2025Updated 9 months ago
- ☆12Oct 17, 2024Updated last year
- A framework for building speech-enabled websites.☆10Jul 10, 2015Updated 11 years ago
- 使用SwiftUI开发的任务管理APP☆11Nov 8, 2023Updated 2 years ago
- Code for building and experimenting on saliency maps for RL agents.☆12Feb 13, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Nov 5, 2025Updated 8 months ago
- This repository contains the code for the paper: "DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utteranc…☆21Oct 13, 2022Updated 3 years ago
- TyDiP Multilingual Politeness dataset and code☆12Oct 15, 2023Updated 2 years ago
- ☆25May 18, 2021Updated 5 years ago
- (NeurIPS 2023 Workshop on DGM4H) Official Implementation of "Adversarial Fine-tuning using Generated Respiratory Sound to Address Class I…☆19Dec 5, 2024Updated last year
- official implementation of paper ExPO: Explainable Phonetic Trait-Oriented Network for Speaker Verification☆15Mar 14, 2025Updated last year
- [CVPR 2024] This is the official implementation of "MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution Disti…☆22Jun 14, 2025Updated last year
- Code for the paper "On the Importance of Feature Decorrelation for Unsupervised Representation Learning for RL" (ICML 2023)☆12Jun 13, 2023Updated 3 years ago
- Data and code for the paper "End-to-End Slot Alignment and Recognition for Cross-Lingual NLU" (Accepted to EMNLP 2020)☆27Jan 13, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code release for paper: The Boombox: Visual Reconstruction from Acoustic Vibrations☆15May 18, 2021Updated 5 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- ☆18Jul 22, 2024Updated 2 years ago
- Moody is a web application allowing the host of online meetings (e.g. via Zoom, Microsoft Teams or Google Meet) to collect real-time feed…☆21Aug 5, 2024Updated last year
- Code for Deep Multimodal Clustering for Unsupervised Audiovisual Learning (CVPR2019)☆15May 27, 2020Updated 6 years ago
- ☆12Apr 26, 2025Updated last year
- [INTERSPEECH 2024] Official pytorch code for the paper "Disentangled Representation Learning for Environment-agnostic Speaker Recognition…☆18Jul 23, 2024Updated 2 years ago