[ICML 2026] Prism: Spectral-Aware Block-Sparse Attention
☆27May 22, 2026Updated 3 months ago
Alternatives and similar repositories for prism
Users that are interested in prism are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository for the EMNLP 2025 paper “UnifiedVisual: A Framework for Constructing Unified Vision-Language Datasets”.☆16Sep 19, 2025Updated 11 months ago
- [ICML 2026] Sparser Block-Sparse Attention via Token Permutation☆32May 22, 2026Updated 3 months ago
- [EMNLP Findings'25] Official PyTorch Implementation of Decoupled Proxy Alignment: Mitigating Language Prior Conflict for Multimodal Align…☆16Sep 19, 2025Updated 11 months ago
- A tool for better use of Inspire platform (Beta: Codeberg version is more up-to-date)☆30Apr 2, 2026Updated 4 months ago
- MOSS-VL is the core multimodal model series within the OpenMOSS ecosystem, dedicated to visual understanding.☆484Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- a survey of long-context LLMs from four perspectives, architecture, infrastructure, training, and evaluation☆63Mar 31, 2025Updated last year
- [AAAI 2024] DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning☆15Apr 29, 2024Updated 2 years ago
- ☆25Jan 29, 2026Updated 7 months ago
- [ICLR26] Beyond Real: Imaginary Extension of Rotary Position Embeddings for Long-Context LLMs☆33Dec 9, 2025Updated 8 months ago
- Modified LLaVA framework for MOSS2, and makes MOSS2 a multimodal model.☆13Sep 19, 2024Updated last year
- Inference-time alignment for harmlessness through cross-model guidance (ACL 2024). Code + MM-Harmful Bench.☆38Oct 2, 2024Updated last year
- MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, i…☆254Jun 16, 2026Updated 2 months ago
- [ACL 2025] "World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning." https://arxiv.org/abs/2503.1…☆18Jul 22, 2025Updated last year
- We introduce 'Thinking with Video', a new paradigm leveraging video generation for multimodal reasoning. Our VideoThinkBench shows that S…☆320Aug 23, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- code for Scaling Laws of RoPE-based Extrapolation☆73Oct 16, 2023Updated 2 years ago
- ☆23May 3, 2025Updated last year
- OpenMOSS presents a collection of our research on LLMs, supported by SII, Fudan and Mosi.☆31Updated this week
- A Python library for building simple, modular, multifunctional, and efficient large model training data synthesis/augmentation pipelines.☆35May 29, 2026Updated 3 months ago
- [ICLR 2025] BitStack: Any-Size Compression of Large Language Models in Variable Memory Environments☆39Feb 17, 2025Updated last year
- Official code of "RoboOmni: Proactive Robot Manipulation in Omni-modal Context"☆119Mar 28, 2026Updated 5 months ago
- ☆29Oct 16, 2025Updated 10 months ago
- 一个面向启智平台(Inspire)的 awesome list☆38Mar 29, 2026Updated 5 months ago
- Elastic Attention: Test-time Adaptive Sparsity Ratios for Efficient Transformers☆24May 26, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- VehicleWorld is the first comprehensive multi-device environment for intelligent vehicle interaction that accurately models the complex, …☆25Sep 16, 2025Updated 11 months ago
- ☆145Jun 24, 2026Updated 2 months ago
- ☆12Jul 23, 2024Updated 2 years ago
- A Docker-first, non-preemptive multi-agent coordination runtime☆17Aug 6, 2026Updated 3 weeks ago
- UnifiedToolHub is a comprehensive project supporting LLM-based tool use, designed to unify various tool-use dataset formats and provide t…☆22Jul 23, 2025Updated last year
- trying to reproduce suno v3☆35Jan 29, 2025Updated last year
- ☆12Dec 6, 2024Updated last year
- [ECCV 2024] The first zero-shot setting for spatio-temporal video grounding.☆11Jul 16, 2024Updated 2 years ago
- Revisiting ASR in the Age of Voice Agents [COLM26]☆29Apr 13, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆14Dec 25, 2024Updated last year
- MOSS-Audio is an open-source foundation model for unified audio understanding, enabling speech, sound, music, captioning, QA, and reasoni…☆651Jun 2, 2026Updated 2 months ago
- VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications☆23Oct 17, 2025Updated 10 months ago
- SLMTokBench for paper "SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models"☆37Aug 29, 2023Updated 3 years ago
- [ICLR'2026] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆18Oct 21, 2025Updated 10 months ago
- ☆27Jun 5, 2023Updated 3 years ago
- HTML Agent based on NexAU☆16Nov 20, 2025Updated 9 months ago