This project use PANNs for audio tagging and sound event detection, and finally get audio embeddings. Then Milvus is used to search the similarity audio items.
☆30Aug 10, 2021Updated 5 years ago
Alternatives and similar repositories for audio_search
Users that are interested in audio_search are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Baseline system for Language-based Audio Retrieval (Task 6B) in DCASE 2023 Challenge☆10Aug 8, 2023Updated 3 years ago
- ☆10Aug 3, 2020Updated 6 years ago
- Keyword extraction using Scake, KeyBERT, Fine-tuning Transformer BERT-like models and ChatGPT.☆12May 22, 2023Updated 3 years ago
- A Rust crate offering similar functionality to the Python transformers package using Candle.☆15Nov 19, 2024Updated last year
- Codebase, data and models for the Headline Grouping paper at NAACL2021☆12Oct 2, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 3D Mesh Generation from 2D Images in Python☆13Feb 12, 2024Updated 2 years ago
- Triton backend for https://github.com/OpenNMT/CTranslate2☆11Aug 20, 2024Updated 2 years ago
- Open source RAG with Llama Index for Japanese LLM in low resource settting☆10May 12, 2025Updated last year
- end-to-end information extraction pipeline built by LayoutLMV2, pretrained model from HuggingFace☆11Aug 15, 2023Updated 3 years ago
- Quora Paraphrasing Dataset Bahasa Indonesia Version☆11Apr 18, 2021Updated 5 years ago
- ☆16Apr 24, 2021Updated 5 years ago
- A demo and tutorial for Council that implements a financial analyst agent.☆11Jun 21, 2024Updated 2 years ago
- Sound Angle Estimation by Fusion of Gaussian Mixture Model and Multiple Signal classification☆12Jan 31, 2019Updated 7 years ago
- Newspaper Segmentation into images and text☆12Jan 11, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An offline CPU-first low-resource chat application to perform RAG on your corpus of data. Powered by OpenChat and CTranslate2.☆15May 14, 2025Updated last year
- This tool can convert picture format(NV12/YUYV/UYVY...) to (png/jpg/bmp)☆10Jul 14, 2018Updated 8 years ago
- Unofficial Tensorflow/Keras implementation of Google AI VoiceFilter☆16Mar 25, 2023Updated 3 years ago
- Deep Learning for HAR: models and tools for Human Activity Recognition from IMU sensor (accelerometer, gyroscope) data☆10Sep 16, 2020Updated 5 years ago
- repo of files pertaining to realtime, offline translations using whisper realtime and argos translate. This repo is marked Creative Commo…☆19May 20, 2025Updated last year
- MUSIC DOA estimation☆14Feb 14, 2019Updated 7 years ago
- PrintCSS Examples created over the time. Mostly HTML some Markdown.☆15Apr 15, 2022Updated 4 years ago
- rewrite python scipy.signal.lfilter in c code☆12Aug 13, 2019Updated 7 years ago
- A repository to store my cuda codes, including some common-used kernels.☆12Sep 19, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- How Media Cloud approaches extracting metadata from online news stories☆18Aug 31, 2026Updated last week
- The code repo for Youtube tutorial series about using Python asyncio with OpenCV to grab frames from video cameras concurrently☆16Oct 3, 2021Updated 4 years ago
- detecting the meotions using by analysing the sound of the person unsing python☆11Oct 7, 2019Updated 6 years ago
- A simple package of face detection☆14Nov 27, 2020Updated 5 years ago
- transformer的 encoder-decoder结构基于tensorflow实现的中文语音识别项目☆34Feb 24, 2021Updated 5 years ago
- ☆12Apr 9, 2021Updated 5 years ago
- Targeted Aspect-based Sentiment Analysis on SentiHood Dataset (PyTorch)☆11Aug 4, 2020Updated 6 years ago
- Segmenting text blocks and baselines from documents using deep learning techniques☆13Jul 27, 2021Updated 5 years ago
- Conv TaSNet follow work of KaiTuo Xu in TF-keras☆14Oct 19, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- SyncTalkFace: Talking Face Generation for Precise Lip-syncing via Audio-Lip Memory☆33Nov 3, 2022Updated 3 years ago
- Bagpipes spaCy is a collection of custom spaCy pipeline components designed to enhance text processing capabilities.☆23Aug 15, 2024Updated 2 years ago
- rockchip_rtsp可以获取rtsp视频流并调用硬件VPU自动解码☆11Mar 4, 2021Updated 5 years ago
- A Deeplearn Model to rec table in photo with ncnn. 一个深度学习模型用于检测图片中的表格 画像内のテーブルを検出するためのディープラーニング モデル☆20Mar 2, 2025Updated last year
- ☆15Nov 28, 2023Updated 2 years ago
- Face detection code for raspberry pi☆13Apr 2, 2021Updated 5 years ago
- Time-domain Audio Separation Network☆24Aug 3, 2018Updated 8 years ago