A Foundation Model for Industrial Signal Comprehensive Representation
☆76Jun 23, 2026Updated last month
Alternatives and similar repositories for FISHER
Users that are interested in FISHER are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ECHO: Frequency-aware Hierarchical Encoding for Variable-length Signal☆19Jul 3, 2026Updated last month
- [IJCAI 2024] EAT: Self-Supervised Pre-Training with Efficient Audio Transformer☆239Nov 30, 2025Updated 8 months ago
- ☆18Updated this week
- [ACL 2026 Main] FineLAP: Taming Heterogeneous Supervision for Fine-grained Language-Audio Pre-training☆36Apr 20, 2026Updated 3 months ago
- Curated list of groundbreaking music generation research.☆21Apr 24, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆26Apr 30, 2026Updated 3 months ago
- [NeurIPS 2025] Benchmark data and code for MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix☆216Feb 25, 2026Updated 5 months ago
- MSGPT: LLM-based Semantic Prompt Framework under Multi-sensor Data for Mechanical Fault Diagnosis☆23Mar 15, 2025Updated last year
- Source code for Consistent ensemble distillation for audio tagging☆76Mar 20, 2026Updated 4 months ago
- Codebase for 'ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining'☆25Jun 20, 2026Updated last month
- [ACL 2026 Main] MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flows☆147Sep 2, 2025Updated 11 months ago
- [ICLR 2025] Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes☆78Oct 8, 2025Updated 10 months ago
- A single-layer, streaming codec model providing SOTA audio quality and discrete tokens designed for superior downstream modelability.☆126Jun 4, 2025Updated last year
- ☆48Apr 27, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML2026] AudioMosaic: Contrastive Masked Audio Representation Learning☆23May 15, 2026Updated 2 months ago
- EAAI: A unified rotating machinery health management framework leveraging large language models for diverse components, conditions, and t…☆18Oct 10, 2025Updated 10 months ago
- A comprehensive gear eccentricity dataset with multiple fault severity levels: Description, characteristics analysis, and fault diagnosis…☆13May 26, 2026Updated 2 months ago
- ☆12Aug 10, 2023Updated 3 years ago
- 🤗 R1-AQA Model: mispeech/r1-aqa☆325Mar 28, 2025Updated last year
- AAAI 2025: BearLLM: A Prior Knowledge-Enhanced Bearing Health Management Framework with Unified Vibration Signal Representation☆128Apr 12, 2025Updated last year
- Code for the paper "DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models"☆17Oct 23, 2025Updated 9 months ago
- Official code for "WavCube: Unifying Speech Representation for Understanding and Generation via Semantic-Acoustic Joint Modeling"☆63Jun 27, 2026Updated last month
- ☆52Jun 17, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- paper for Anomalous sound detection☆46Feb 27, 2026Updated 5 months ago
- FD-MVLLM: Fault Diagnosis Based on Multimodal Vibration Data and Large Language Model for Bearing☆84Jan 21, 2026Updated 6 months ago
- Submission for task 2 "First-Shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring" of the DCASE challenge 2023 (h…☆18May 22, 2023Updated 3 years ago
- ☆29Jul 4, 2025Updated last year
- The official repository of SpeechCraft dataset, a large-scale expressive bilingual speech dataset with natural language descriptions.☆198Feb 28, 2026Updated 5 months ago
- SDUST dataset for fault diagnosis pulished by Shandong University of Science and Technology☆47Oct 15, 2024Updated last year
- [ICASSP 2026] Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis☆41Dec 24, 2025Updated 7 months ago
- ☆131Feb 6, 2025Updated last year
- Pytorch implementation of Deep Generic Representations for Domain-Generalized Anomalous Sound Detection: https://arxiv.org/abs/2409.05035☆29Mar 16, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Repository of UltraVoice☆63Oct 28, 2025Updated 9 months ago
- [EMNLP 2025 Findings] A complete cross-modal RAG system for end-to-end speech-to-speech large models, including ASR-based Retrieval and E…☆31Jul 11, 2025Updated last year
- [INTERSPEECH 2025 Oral]Official code for "Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment"☆67Jun 16, 2025Updated last year
- ☆44Apr 26, 2026Updated 3 months ago
- A curated list of models, benchmarks, tools and guides for audio editing☆35Updated this week
- Official pytorch implementation of AEGAN-AD☆53Jun 23, 2025Updated last year
- 本项目旨在构建一个高精度、跨域鲁棒性强的智能故障诊断系统,应用于旋转机械(如轴承)的故障检测和分类。项目将复杂的时域振动信号转化为时频图像,并利用卷积神经网络 (CNN) 进行特征提取和诊断。 项目的核心创新点在于深度迁移学习 (Deep Transfer Learnin…☆25Dec 2, 2025Updated 8 months ago