OpenMOSS presents a collection of our research on LLMs, supported by SII, Fudan and Mosi.
☆31Sep 21, 2026Updated this week
Alternatives and similar repositories for OpenMOSS
Users that are interested in OpenMOSS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An asynchronous file-based handoff protocol for Claude Code and Codex agents sharing one project workspace☆40Jul 4, 2026Updated 2 months ago
- ☆23Mar 2, 2026Updated 6 months ago
- Low-rank sparse attention decomposition for LLM interpretability; active development continues in Llamascopium☆30Nov 9, 2025Updated 10 months ago
- ☆24Nov 16, 2025Updated 10 months ago
- An open-weight 11B model series for long-form and real-time video understanding☆726Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICML 2026] Prism: Spectral-Aware Block-Sparse Attention☆27May 22, 2026Updated 4 months ago
- FamilyTool benchmark☆14Sep 10, 2025Updated last year
- A multi-device environment for evaluating intelligent vehicle interaction across connected cockpit systems☆25Sep 16, 2025Updated last year
- [ICML 2026] Sparser Block-Sparse Attention via Token Permutation☆33May 22, 2026Updated 4 months ago
- Official implementation of ACL'26 (findings) paper WESR (Word-level Event-Speech Recognition): A comprehensive benchmark and baseline for…☆48Jan 30, 2026Updated 7 months ago
- ☆25Jul 20, 2025Updated last year
- A survey of long-context language models covering architecture, infrastructure, training, and evaluation☆64Mar 31, 2025Updated last year
- ☆29Sep 9, 2026Updated last week
- An imaginary extension of rotary position embeddings for long-context language models☆33Dec 9, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SLMTokBench for paper "SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models"☆37Aug 29, 2023Updated 3 years ago
- A Chinese human-level benchmark for evaluating multimodal large language models☆83Mar 13, 2024Updated 2 years ago
- [AAAI 2024] DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning☆15Apr 29, 2024Updated 2 years ago
- An end-to-end speech-to-speech language model that generates spoken responses without text guidance☆139Feb 13, 2026Updated 7 months ago
- A curated list of models, benchmarks, tools and guides for audio editing☆46Updated this week
- A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning☆1,401Sep 6, 2026Updated 2 weeks ago
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago
- ☆19Nov 6, 2023Updated 2 years ago
- ☆60Jun 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An open-source personal academic homepage template characterized by its user-friendly design and extensive scalability.☆36Oct 6, 2025Updated 11 months ago
- [NeurIPS 2024] Can Language Models Learn to Skip Steps?☆22Jan 25, 2025Updated last year
- code for Scaling Laws of RoPE-based Extrapolation☆73Oct 16, 2023Updated 2 years ago
- Methods and code for extending the context length of diffusion language models☆56Dec 7, 2025Updated 9 months ago
- A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music☆257Jun 16, 2026Updated 3 months ago
- A real-time spoken dialogue system for direct speech-to-speech interaction☆375Jan 27, 2025Updated last year
- ☆37Aug 31, 2025Updated last year
- Official repository for the EMNLP 2025 paper “UnifiedVisual: A Framework for Constructing Unified Vision-Language Datasets”.☆16Sep 19, 2025Updated last year
- Course homepage of Introduction to Computer System, Fall 2025, Fudan University☆31Sep 5, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [EMNLP Findings'25] Official PyTorch Implementation of Decoupled Proxy Alignment: Mitigating Language Prior Conflict for Multimodal Align…☆16Sep 19, 2025Updated last year
- ☆12Jul 23, 2024Updated 2 years ago
- A framework for training, analyzing, and visualizing sparse autoencoders and related interpretability methods☆228Sep 6, 2026Updated 2 weeks ago
- ☆15Apr 11, 2026Updated 5 months ago
- ☆28Mar 31, 2026Updated 5 months ago
- [ACL 2025] "World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning." https://arxiv.org/abs/2503.1…☆18Jul 22, 2025Updated last year
- [ICLR 2026] An official implementation of "STAR-Bench: Probing Deep Spatio-Temporal Reasoning as Audio 4D Intelligence"☆44Apr 19, 2026Updated 5 months ago