☆11Mar 24, 2025Updated last year
Alternatives and similar repositories for ICASSP2025-IIICSS
Users that are interested in ICASSP2025-IIICSS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Mar 25, 2025Updated last year
- FCTalker: Fine and Coarse Grained Context Modeling for Expressive Conversational Speech Synthesis (Accepted by ISCSLP'2024)☆26Feb 22, 2024Updated 2 years ago
- 16k Hz Vocoder (HiFiGAN Codes and Pretrained Models)☆18Apr 3, 2023Updated 3 years ago
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)☆61Nov 1, 2024Updated last year
- [TIP2025] The implementation of "Uncertainty Guided Refinement for Fine-grained Salient Object Detection"☆18Apr 20, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Emotion Rendering for Conversational Speech Synthesis with Heterogeneous Graph-Based Context Modeling (Accepted by AAAI'2024)☆59Jun 20, 2024Updated 2 years ago
- Xmake C++23 project template. Using C++ modules, github workflows for CI/CD (Windows and Ubuntu) and gtest for testing. Compiles with bot…☆17Mar 11, 2024Updated 2 years ago
- Rust-style mutex type for C++☆17Jan 12, 2024Updated 2 years ago
- ☆19Dec 22, 2025Updated 7 months ago
- ☆24Oct 23, 2024Updated last year
- ☆11Aug 20, 2025Updated 11 months ago
- Code of Semantic Point Cloud Upsampling (SPU) published on IEEE Transactions on Multimedia.☆23Apr 23, 2022Updated 4 years ago
- Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection☆13Jul 7, 2026Updated 2 weeks ago
- This is a simple test for light field ToolBox.☆27Oct 15, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repo and evaluation implementation of KnowRecall and VisRecall☆10May 22, 2025Updated last year
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- The Poseidon Server Framework☆20Jul 16, 2026Updated last week
- A fast and scalable distributed lock service using programmable switches.☆21Jul 30, 2024Updated last year
- CVPR 2024 accepted paper, An Upload-Efficient Scheme for Transferring Knowledge From a Server-Side Pre-trained Generator to Clients in He…☆68Mar 12, 2025Updated last year
- Java聊天室☆30Jul 30, 2016Updated 9 years ago
- Dubs the video in another language, Powered by Deepgram API and Google Translate.☆16Apr 11, 2022Updated 4 years ago
- ☆71Jun 2, 2023Updated 3 years ago
- [ICML'26] Toward Human-like Audio-Visual Intelligence of Omni-MLLMs☆16Jun 20, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for paper "FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning" Neurips2025.☆15Jan 29, 2026Updated 5 months ago
- ☆17Sep 5, 2025Updated 10 months ago
- A SQL parser written in C++☆32Oct 22, 2021Updated 4 years ago
- [ACM MM 2023] Official PyTorch implementation of "Emo-DNA: Emotion Decoupling and Alignment Learning for Cross-Corpus Speech Emotion Reco…☆12Aug 4, 2023Updated 2 years ago
- create_pg_super_document is a project that generates documentation for all symbols in the PostgreSQL codebase, then utilizes these symbol…☆31May 9, 2026Updated 2 months ago
- PnP-3D: A Plug-and-Play for 3D Point Clouds (TPAMI 2021)☆45Feb 14, 2022Updated 4 years ago
- FunASR实时语音识别版,识别麦克风和电脑内播放的声音,电脑语音打字软件☆19Sep 12, 2025Updated 10 months ago
- [CVPR 2025] Official implementation of paper "Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie…☆23Jun 6, 2025Updated last year
- Prompting Large Language Models with Audio for General-Purpose Speech Summarization☆20May 14, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A binary i/o library for C++, without the agonizing pain☆36Aug 3, 2023Updated 2 years ago
- Visual speech recognition with face inputs: code and models for F&G 2020 paper "Can We Read Speech Beyond the Lips? Rethinking RoI Select…☆19Apr 12, 2021Updated 5 years ago
- The code of paper PCDreamer☆46Oct 15, 2025Updated 9 months ago
- ☆15Apr 9, 2026Updated 3 months ago
- Official Repository for "Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality" (ECCV 2024)☆16Oct 29, 2024Updated last year
- Expression Snippet Transformer for Robust Video-based Facial Expression Recognition☆17Jan 27, 2024Updated 2 years ago
- code of cvpr26 paper Symphony☆17Apr 7, 2026Updated 3 months ago