☆29Feb 23, 2026Updated 6 months ago
Alternatives and similar repositories for AuroLA
Users that are interested in AuroLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ETH Zürich MSc Thesis: Accelerating Neural Audio Synthesis☆26Apr 10, 2023Updated 3 years ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 11 months ago
- Prompting Large Language Models with Audio for General-Purpose Speech Summarization☆20May 14, 2025Updated last year
- ☆14Apr 4, 2025Updated last year
- Multi-talker ASR based on DiCoW with Serialized Output Training☆21Sep 18, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆48Jul 26, 2026Updated last month
- Nix packages for reproducible MIR research☆16Aug 17, 2026Updated last month
- MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations☆43Aug 29, 2026Updated 3 weeks ago
- Code for WACV24 work for multiview acoustic-visual detection☆13Mar 22, 2024Updated 2 years ago
- OutboundEval, a comprehensive benchmark for evaluating large language models (LLMs) in expert-level intelligent outbound calling scenario…☆17Oct 28, 2025Updated 10 months ago
- Diabetic Retinopathy Two-field image Dataset (DRTiD) & source code of Cross-Field Transformer for Diabetic Retinopathy Grading on Two-fie…☆29May 28, 2024Updated 2 years ago
- ☆18Jun 25, 2026Updated 2 months ago
- A compiler that bridges Faust DSP code with the CLAP plugin standard, producing plugins from .dsp files. It supports both static builds a…☆19Mar 27, 2026Updated 5 months ago
- "Deep Learning Models for Automatic Chord Recognition in Polyphonic Audio” for the EPSRC-funded Bede Supercomputer studentship.☆17Mar 26, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Grouped Feedback Delay Networks for Coupled Room Modeling☆41May 27, 2024Updated 2 years ago
- ☆32Aug 18, 2026Updated last month
- ☆11Dec 28, 2023Updated 2 years ago
- [EMNLP 2024] Official code repository of paper titled "PALM: Few-Shot Prompt Learning for Audio Language Models" accepted in EMNLP 2024 c…☆29Dec 22, 2024Updated last year
- Pytorch implementation for Egoinstructor at CVPR 2024☆28Dec 1, 2024Updated last year
- ☆33Dec 7, 2025Updated 9 months ago
- WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs☆51Jul 12, 2026Updated 2 months ago
- Generate embeddings for audio files (music, speech, sounds) and text using CLAP with llm☆22May 15, 2025Updated last year
- MichiAI: A Low Latency, Full Duplex Speech LLM with zero coherence loss☆120Sep 3, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR2025] Official code for Lost in Translation Found in Context☆24Jan 14, 2026Updated 8 months ago
- Music Demixing Challenge Submission Repo☆16Sep 8, 2023Updated 3 years ago
- ☆29May 22, 2026Updated 3 months ago
- Dark-velvet-noise reverb: Accurate model for late-reverberation with arbitrary temporal energy decay. Offline implementations in Matlab a…☆26Nov 1, 2024Updated last year
- Virtual Model of the Echoplex Ep-3 tape delay, made in Faust. This was a project for the Sound Processing exam (AAU SMC 2019)☆18Oct 19, 2020Updated 5 years ago
- ☆14Nov 22, 2022Updated 3 years ago
- Code for "AudioMarathon: A Comprehensive Benchmark for Long-Context Audio Understanding and Efficiency in Audio LLMs"☆26Oct 9, 2025Updated 11 months ago
- Self-Supervised Contrastive Learning of Music Spectrograms☆31May 10, 2021Updated 5 years ago
- Faust program for generating rain, wind, surf, and other forms of sparse granular sound.☆23Jan 20, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A lightweight end-of-utterance detection model fine-tuned on SmolLM2-135M, optimized for Raspberry Pi and low-power devices.☆65Mar 20, 2026Updated 5 months ago
- The official repository TimeAudio, a comprehensive framework that incorporates fine-grained acoustic cues into LALMs with enhanced module…☆31Nov 18, 2025Updated 10 months ago
- Official inference code for UniSS: Unified Expressive Speech-to-Speech Translation with Your Voice.☆46May 30, 2026Updated 3 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆46Apr 11, 2026Updated 5 months ago
- open-vocabulary sound event detection☆56Dec 17, 2025Updated 9 months ago
- Generate synthesizer sounds from text prompts with a simple evolutionary algorithm.☆28Jan 12, 2026Updated 8 months ago
- This is an ultra-simple, single-file PyTorch implementation of MoonViT, the native-resolution vision encoder from Kimi-VL.☆38Apr 25, 2026Updated 4 months ago