Qualifying Exam Preparing
☆18May 7, 2025Updated last year
Alternatives and similar repositories for QualifyingExamPreparing
Users that are interested in QualifyingExamPreparing are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- The implementation for "Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions"☆51Apr 7, 2025Updated last year
- This repository follows papers and reports on discrete speech representation learning and speech tokenization methods for speech language…☆15Dec 1, 2023Updated 2 years ago
- A casual and simple ChatGPT Python script that can run using terminal (as long as you have an API). Support Azure API.☆20May 3, 2025Updated last year
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models☆124Oct 16, 2025Updated 11 months ago
- ☆25Nov 25, 2025Updated 9 months ago
- This is for EMNLP 2024 Paper: AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction☆16Nov 4, 2024Updated last year
- Project of ACL 2025 "UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models"☆15Mar 25, 2025Updated last year
- [SIGKDD 2024] Rethinking Fair Graph Neural Networks from Re-balancing☆10Jul 15, 2024Updated 2 years ago
- The implementation for "Empowering Whisper as a Joint Multi-Talker and Target-Talker Speech Recognition System".☆34Aug 2, 2025Updated last year
- Latex template for CUHK PhD Thesis☆14Jun 29, 2025Updated last year
- [ICLR 2026] | MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark☆20Feb 12, 2026Updated 7 months ago
- ☆44Jun 3, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Script to perform statistical significance test between ASR hypotheses.☆23Aug 13, 2017Updated 9 years ago
- Fine-Tune Whisper with Transformers and PEFT☆58Nov 4, 2023Updated 2 years ago
- Official implementation of paper "GraphControl: Adding Conditional Control to Universal Graph Pre-trained Models for Graph Domain Transfe…☆18Jan 27, 2024Updated 2 years ago
- ☆33Mar 11, 2022Updated 4 years ago
- The code Implementation of the paper “Universal Prompt Tuning for Graph Neural Networks”.☆26Oct 16, 2023Updated 2 years ago
- Collection of awesome Continual Test-Time Adaptation methods☆24Jun 4, 2024Updated 2 years ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speech☆12May 14, 2025Updated last year
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 6 months ago
- Unofficial PyTorch implementation of "Autoregressive Speech Synthesis without Vector Quantization (MELLE)"☆41Jun 28, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ADAPTING SELF-SUPERVISED MODELS TO MULTI-TALKER SPEECH RECOGNITION USING SPEAKER EMBEDDINGS☆35Mar 16, 2023Updated 3 years ago
- ☆34Jun 12, 2025Updated last year
- Official code for "F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization"☆169Mar 3, 2026Updated 6 months ago
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- Use strategy in stock transaction for high revenue.☆10Dec 24, 2015Updated 10 years ago
- Here the code of EmoAudioNet is a deep neural network for speech classification (published in ICPR 2020)☆14Jul 13, 2020Updated 6 years ago
- Multimodal Speech Recognition for phoneme level prediction using Audio-Visual data from TCDTIMIT dataset implementing RNNs with LSTMs for…☆15Jul 27, 2023Updated 3 years ago
- Europeanized CosyVoice2 for French & German☆18Mar 30, 2026Updated 5 months ago
- ☆53Aug 27, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Apr 25, 2024Updated 2 years ago
- ☆23Sep 14, 2025Updated last year
- ☆49Oct 28, 2024Updated last year
- Real-time Speech-Text Foundation Model Toolkit (wip)☆254Mar 26, 2025Updated last year
- The open source code for LLM-Codec☆147Aug 18, 2024Updated 2 years ago
- FitHuBERT: Going Thinner and Deeper for Knowledge Distillation of Speech Self-Supervised Learning (INTERSPEECH 2022)☆19Nov 15, 2023Updated 2 years ago
- ☆16Aug 10, 2025Updated last year