☆96Apr 3, 2023Updated 3 years ago
Alternatives and similar repositories for Mod-Squad
Users that are interested in Mod-Squad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2023 Findings] Emergent Modularity in Pre-trained Transformers☆26Jun 7, 2023Updated 3 years ago
- PyTorch implementation of "From Sparse to Soft Mixtures of Experts"☆72Aug 22, 2023Updated 3 years ago
- Project Page for "Multi-Task Dense Prediction via Mixture of Low-Rank Experts"☆91Jun 9, 2025Updated last year
- [NeurIPS 2022] “M³ViT: Mixture-of-Experts Vision Transformer for Efficient Multi-task Learning with Model-Accelerator Co-design”, Hanxue …☆137Nov 30, 2022Updated 3 years ago
- An up-to-date list of works on Multi-Task Learning☆379Mar 2, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [EMNLP 2025 Findings] MEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation☆15Aug 22, 2025Updated last year
- Mixture of Attention Heads☆54Oct 10, 2022Updated 3 years ago
- [ECCV2024] The official implementation of "Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation".☆18Feb 24, 2025Updated last year
- ☆13Sep 26, 2025Updated 11 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 7 months ago
- PyTorch implementation of LIMoE☆52Apr 1, 2024Updated 2 years ago
- The official repository for the experiments included in the paper titled "Patch-level Routing in Mixture-of-Experts is Provably Sample-ef…☆14Feb 12, 2026Updated 7 months ago
- Official Implementation of Frequency-enhanced Data Augmentation for Vision-and-Language Navigation (NeurIPS2023)☆15Jan 8, 2024Updated 2 years ago
- Visual Representation Learning Benchmark for Self-Supervised Models☆35Apr 18, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- hybrid representation learning approach for fully automatic and template-free vessel centerline extraction☆12Jul 17, 2023Updated 3 years ago
- Official PyTorch implementation of ResFormer: Scaling ViTs with Multi-Resolution Training, CVPR2023☆30Jun 22, 2023Updated 3 years ago
- Deep Learning (FS 2020)☆17Oct 10, 2022Updated 3 years ago
- ☆21Oct 31, 2022Updated 3 years ago
- [ICLR 2023] "Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers" by Tianlong Chen*, Zhenyu Zhang*, Ajay Jaiswal…☆56Feb 28, 2023Updated 3 years ago
- ☆29Oct 9, 2024Updated last year
- A Pytorch implementation of Sparsely-Gated Mixture of Experts, for massively increasing the parameter count of language models☆873Sep 13, 2023Updated 3 years ago
- Bibliometric. A Python framework designed for the analysis and evaluation of scholarly publications.☆15Jan 16, 2026Updated 8 months ago
- [ICCV 2023 oral] This is the official repository for our paper: ''Sensitivity-Aware Visual Parameter-Efficient Fine-Tuning''.☆77Sep 24, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [AAAI2025] Dynamic Contrastive Knowledge Distillation for Efficient Image Restoration☆34Feb 11, 2025Updated last year
- The code source of MambaCSR☆27Sep 24, 2024Updated last year
- ☆20Jul 3, 2019Updated 7 years ago
- none official pytorch implement of CVPR2020 GroupFace☆24Feb 7, 2021Updated 5 years ago
- ☆28Mar 20, 2023Updated 3 years ago
- a fast implementation of BM25☆10Sep 15, 2022Updated 4 years ago
- Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs☆25Nov 11, 2025Updated 10 months ago
- ☆290Aug 14, 2025Updated last year
- SVL-Adapter: Self-Supervised Adapter for Vision-Language Pretrained Models☆21Jan 11, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- For MICCAI 2023☆19Jul 21, 2023Updated 3 years ago
- A collection of AWESOME things about mixture-of-experts☆1,288Dec 8, 2024Updated last year
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 7 months ago
- (CVPR 2024) "Unsegment Anything by Simulating Deformation"☆29May 27, 2024Updated 2 years ago
- Implementation of "Towards Understanding Mixture of Experts in Deep Learning", NeurIPS 2022☆10Jan 6, 2023Updated 3 years ago
- PyTorch Re-Implementation of "The Sparsely-Gated Mixture-of-Experts Layer" by Noam Shazeer et al. https://arxiv.org/abs/1701.06538☆1,253Apr 19, 2024Updated 2 years ago
- Official PyTorch Implementation of EMoE: Unlocking Emergent Modularity in Large Language Models [main conference @ NAACL2024]☆40May 28, 2024Updated 2 years ago