☆44Jan 5, 2023Updated 3 years ago
Alternatives and similar repositories for acoustic-model
Users that are interested in acoustic-model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Jun 22, 2026Updated 2 months ago
- Official implementation of the paper "ProxyDet: Synthesizing Proxy Novel Classes via Classwise Mixup for Open-Vocabulary Object Detection…☆26Feb 13, 2024Updated 2 years ago
- ☆25May 9, 2024Updated 2 years ago
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 9 months ago
- [AAAI2025] MoRe: Class Patch Attention Needs Regularization for Weakly Supervised Semantic Segmentation☆21Jan 15, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [IEEE TCSVT 2023] The implementation of our paper Semi-Supervised Subspace Clustering via Tensor Low-Rank Representation.☆26Dec 21, 2023Updated 2 years ago
- [ACM MM 26] Peak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic Assessment☆26Aug 5, 2026Updated last month
- Plan, Posture and Go: Towards Open-World Text-to-Motion Generation☆42Nov 19, 2024Updated last year
- ☆36Jul 7, 2025Updated last year
- Training open models for agentic phone use with real-app and mock-app environments.☆57Aug 24, 2026Updated 2 weeks ago
- [CVPR 2024] Narrative Action Evaluation with Prompt-Guided Multimodal Interaction☆43May 16, 2024Updated 2 years ago
- [ECCV'24] Official PyTorch implementation of In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation☆51Sep 24, 2024Updated last year
- The repository contains the official implementation of "DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh Recovery", CVPR 2024☆46Jun 4, 2024Updated 2 years ago
- [ML4H'25] m1: Unleash the Potential of Test-Time Scaling for Medical Reasoning in Large Language Models☆51Dec 21, 2025Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS'24] A Simple Image Segmentation Framework via In-Context Examples☆68Oct 29, 2024Updated last year
- Official implement of ICML2024 Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation☆58Aug 15, 2024Updated 2 years ago
- ☆52Aug 22, 2025Updated last year
- ☆60Sep 14, 2024Updated last year
- ☆56Oct 5, 2022Updated 3 years ago
- [ML4H'25] MedVLThinker: Simple Baselines for Multimodal Medical Reasoning☆61Dec 21, 2025Updated 8 months ago
- Official Repo for PosSAM: Panoptic Open-vocabulary Segment Anything☆71Apr 7, 2024Updated 2 years ago
- [ACL2025 Findings] Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models☆91May 20, 2025Updated last year
- [ICCV 2025 Oral] CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation☆73Aug 16, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV2024] ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference☆101Mar 26, 2025Updated last year
- [ICLR'26] Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs☆101Jan 26, 2026Updated 7 months ago
- [CVPR2025] Exploring CLIP’s Dense Knowledge for Weakly Supervised Semantic Segmentation☆72Jun 21, 2025Updated last year
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology☆93Jan 26, 2026Updated 7 months ago
- [NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.☆102Jan 3, 2025Updated last year
- Next Token Is Enough: Realistic Image Quality and Aesthetic Scoring with Multimodal Large Language Model.☆87Jun 30, 2025Updated last year
- [ICLR 2025] SAMRefiner: Taming Segment Anything Model for Universal Mask Refinement☆108Apr 19, 2025Updated last year
- Dimple, the first Discrete Diffusion Multimodal Large Language Model☆117Aug 23, 2026Updated 2 weeks ago
- Source code for the paper "Long-Tail Learning with Foundation Model: Heavy Fine-Tuning Hurts" (ICML 2024)☆109Oct 22, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NeurlPS 2024] One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos☆150Dec 26, 2024Updated last year
- [CVPR24] Official Implementation of GEM (Grounding Everything Module)☆140Apr 10, 2025Updated last year
- Official implementation of "Open-Vocabulary Multi-Label Classification via Multi-Modal Knowledge Transfer".☆129Nov 7, 2024Updated last year
- [CVPR 2024 oral]This repository contains the official implementation of "FlowIE: Efficient Image Enhancement via Rectified Flow"☆154Jan 13, 2025Updated last year
- [ICLR26] NarrLV: Towards a Comprehensive Narrative-Centric Evaluation for Long Video Generation Models☆113Jul 28, 2025Updated last year
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆146Feb 26, 2026Updated 6 months ago
- [ICCV 2025] Code for Momentum-GS: Momentum Gaussian Self-Distillation for High-Quality Large Scene Reconstruction☆174Dec 15, 2025Updated 8 months ago