☆44Jan 5, 2023Updated 3 years ago
Alternatives and similar repositories for acoustic-model
Users that are interested in acoustic-model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Jun 22, 2026Updated last month
- This is the formal code implementation of the CVPR 2024 paper 'Traceable Federated Continual Learning'.☆19May 31, 2024Updated 2 years ago
- ☆25May 9, 2024Updated 2 years ago
- [ACM MM 25] FingER: Content Aware Fine-grained Evaluation with Reasoning for AI-Generated Videos☆17Jul 17, 2025Updated last year
- [CVPR'2022, TPAMI'2024] LAVT: Language-Aware Vision Transformer for Referring Segmentation☆26Jan 21, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 8 months ago
- [AAAI2025] MoRe: Class Patch Attention Needs Regularization for Weakly Supervised Semantic Segmentation☆21Jan 15, 2025Updated last year
- [IEEE TCSVT 2023] The implementation of our paper Semi-Supervised Subspace Clustering via Tensor Low-Rank Representation.☆26Dec 21, 2023Updated 2 years ago
- The official implementation of A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation☆27Aug 17, 2025Updated last year
- Chat about anything on any video!☆39Sep 5, 2023Updated 2 years ago
- Plan, Posture and Go: Towards Open-World Text-to-Motion Generation☆42Nov 19, 2024Updated last year
- Official implementation of the WACV 2024 paper CLIP-DIY☆34Dec 20, 2023Updated 2 years ago
- Official Implementation of "Open-Vocabulary Audio-Visual Semantic Segmentation" [ACM MM 2024 Oral].☆37Nov 2, 2024Updated last year
- Training open models for agentic phone use with real-app and mock-app environments.☆56Jun 26, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ECCV'24] Official PyTorch implementation of In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation☆51Sep 24, 2024Updated last year
- ☆43Aug 5, 2025Updated last year
- cliptrase☆47Sep 1, 2024Updated last year
- [IROS 2025] ManiGaussian++: General Robotic Bimanual Manipulation with Hierarchical Gaussian World Model☆46Jun 26, 2025Updated last year
- The repository contains the official implementation of "DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh Recovery", CVPR 2024☆46Jun 4, 2024Updated 2 years ago
- KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding☆70Apr 5, 2026Updated 4 months ago
- [ML4H'25] m1: Unleash the Potential of Test-Time Scaling for Medical Reasoning in Large Language Models☆51Dec 21, 2025Updated 8 months ago
- [NeurIPS'24] A Simple Image Segmentation Framework via In-Context Examples☆68Oct 29, 2024Updated last year
- Official implement of ICML2024 Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation☆58Aug 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [TIP 2025] Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation☆73Aug 10, 2026Updated last week
- ☆52Aug 22, 2025Updated 11 months ago
- ☆60Aug 12, 2024Updated 2 years ago
- ☆60Sep 14, 2024Updated last year
- [ML4H'25] MedVLThinker: Simple Baselines for Multimodal Medical Reasoning☆60Dec 21, 2025Updated 8 months ago
- PyTorch Implementation of NACLIP in "Pay Attention to Your Neighbours: Training-Free Open-Vocabulary Semantic Segmentation"☆81Aug 4, 2026Updated 2 weeks ago
- [ACL2025 Findings] Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models☆91May 20, 2025Updated last year
- [ICCV 2023] CTVIS: Consistent Training for Online Video Instance Segmentation☆83Oct 15, 2023Updated 2 years ago
- SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis☆71Jul 24, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- What and How Well You Performed? A Multitask Learning Approach to Action Quality Assessment [CVPR 2019]☆77May 5, 2025Updated last year
- [ECCV2024] ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference☆101Mar 26, 2025Updated last year
- [CVPR2024] Separate and Conquer: Decoupling Co-occurrence via Decomposition and Representation for Weakly Supervised Semantic Segmentatio…☆81Oct 10, 2024Updated last year
- [ICLR'26] Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs☆100Jan 26, 2026Updated 6 months ago
- [CVPR2025] Exploring CLIP’s Dense Knowledge for Weakly Supervised Semantic Segmentation☆71Jun 21, 2025Updated last year
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology☆92Jan 26, 2026Updated 6 months ago
- [NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.☆102Jan 3, 2025Updated last year