☆15Jan 24, 2025Updated last year
Alternatives and similar repositories for gaudi-lavcap
Users that are interested in gaudi-lavcap are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Sep 1, 2024Updated last year
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆15May 13, 2025Updated last year
- [ICCV 2023] Audio-Visual Class-Incremental Learning☆36Sep 29, 2024Updated last year
- Implementation of the paper "CXR-IRGen: An Integrated Vision and Language Model for the Generation of Clinically Accurate Chest X-Ray Ima…☆21Jul 2, 2024Updated 2 years ago
- The official implementation of DMEL the method presented in the paper "DMEL: The differentiable log-Mel spectrogram as a trainable layer …☆24Dec 21, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official repository for "OTMorph: Unsupervised Multi-domain Abdominal Medical Image Registration Using Neural Optimal Transport".☆19Dec 25, 2024Updated last year
- [IJBHI 2024] This is the official implementation of CAMANet: Class Activation Map Guided Attention Network for Radiology Report Generati…☆11May 14, 2025Updated last year
- ☆15Dec 31, 2024Updated last year
- An unofficial code reproduction of Channel Attention Dense U-Net for Multichannel Speech Enhancement☆13Jul 17, 2023Updated 3 years ago
- Implementation of "DeepWriter: A Multi-Stream Deep CNN for Text-independent Writer Identification"☆16Feb 3, 2020Updated 6 years ago
- 🔊 Repository for our NAACL-HLT 2019 paper: AudioCaps☆215Oct 6, 2025Updated 10 months ago
- Repository for reproducing result in journal "Self-supervised learning for Speech Emotion Recognition"☆10Mar 15, 2023Updated 3 years ago
- it makes txt file with chat written by twitch.tv replay.☆13Apr 3, 2023Updated 3 years ago
- (ICLR 2025) Multi-Task Corrupted Prediction for Learning Robust Audio-Visual Speech Representation☆16Apr 29, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR2025] Official code repository for SeTa: "Scale Efficient Training for Large Datasets"☆24Mar 18, 2025Updated last year
- NeurIPS'2023 official implementation code☆70Nov 11, 2023Updated 2 years ago
- Code for "Recognizing Scenes from Novel Viewpoints"☆29Sep 16, 2022Updated 3 years ago
- ☆46Oct 17, 2025Updated 9 months ago
- Tree-Based Diffusion Schrödinger Bridge with Applications to Wasserstein Barycenters☆10Mar 5, 2024Updated 2 years ago
- Pytorch implementation of "Entropic Neural Optimal Transport via Diffusion Processes" (NeurIPS 2023, oral).☆44Mar 11, 2024Updated 2 years ago
- JMLR Cover Letter Template☆10Dec 15, 2021Updated 4 years ago
- ☆10Aug 23, 2022Updated 3 years ago
- [2025 CVPR] Towards Open-Vocabulary Audio-Visual Event Localization☆46Mar 7, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆10Jan 14, 2023Updated 3 years ago
- "Applying Regularized Schrödinger-Bridge-Based Stochastic Process in Generative Modeling"☆11Aug 16, 2022Updated 3 years ago
- ☆13Jan 10, 2017Updated 9 years ago
- An attempt to replicate the results of [1706.08612] VoxCeleb: a large-scale speaker identification dataset☆12Dec 11, 2019Updated 6 years ago
- ☆13Jul 25, 2023Updated 3 years ago
- PrideDiff: Physics-Regularized Generalized Diffusion Model for CT Reconstruction☆12Dec 10, 2024Updated last year
- ☆13Sep 13, 2023Updated 2 years ago
- Code for T-MARS data filtering☆35Aug 23, 2023Updated 2 years ago
- [ICCV2023] DR-Tune: Improving Fine-tuning of Pretrained Visual Models by Distribution Regularization with Semantic Calibration☆12Oct 12, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Multiple Self-Similarity Network Based Plug-and-Play Prior for MRI Reconstruction☆15Mar 6, 2020Updated 6 years ago
- ☆13Jan 5, 2025Updated last year
- For CVPR2022 Submission☆10Sep 4, 2022Updated 3 years ago
- Official Repository for "Learning to Visually Localize Sound Sources from Mixtures without Prior Source Knowledge" (CVPR 2024)☆17Sep 1, 2024Updated last year
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 3 years ago
- k-t CLAIR: Self-Consistency Guided Multi-Prior Learning for Dynamic Parallel MR Image Reconstruction☆11Jan 30, 2025Updated last year
- VoxSRC2022 workshop development kit☆19Jul 21, 2022Updated 4 years ago