"AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation"
☆43May 26, 2026Updated 3 months ago
Alternatives and similar repositories for AR-Omni
Users that are interested in AR-Omni are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- "A Survey on Agent-as-a-Judge"☆145May 11, 2026Updated 3 months ago
- ☆21Aug 9, 2024Updated 2 years ago
- Official repo for SAO-Instruct: Free-form Audio Editing using Natural Language Instructions presented at NeurIPS 2025☆19Oct 28, 2025Updated 10 months ago
- Adaptive Multimodal Reasoning via Reinforcement Learning☆24Jan 11, 2026Updated 7 months ago
- [CVPR-26] Official repository of "CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization"☆19Mar 9, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- AI Tool - A product of The Village in Apex Aurum☆66Feb 2, 2026Updated 7 months ago
- ☆30Aug 25, 2024Updated 2 years ago
- ☆30May 22, 2026Updated 3 months ago
- 用Paddle复现论文ChineseBERT: Chinese Pretraining Enhanced by Glyph and Pinyin Information(ACL2021)☆10Nov 15, 2021Updated 4 years ago
- StreamUni is a framework that efficiently enables unified Large Speech-Language Models to accomplish streaming speech translation in a co…☆26Jul 14, 2025Updated last year
- Codes for our paper "Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation" (EMNLP 2023 Findings)☆47Dec 9, 2023Updated 2 years ago
- BLSP: Bootstrapping Langauge-Speech Pre-training via Behavior Alignment of Continuation Writing☆59Mar 11, 2024Updated 2 years ago
- Official repository for the paper "Audio ControlNet for Fine-Grained Audio Generation and Editing".☆80Feb 7, 2026Updated 6 months ago
- World Model Self-Distillation project website☆19Jun 15, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞☆16Jan 29, 2026Updated 7 months ago
- ☆18Apr 14, 2026Updated 4 months ago
- Real-Time learning "Living entity"☆61Jan 26, 2026Updated 7 months ago
- A real-time Electron-based desktop GUI for DeepSeek-OCR☆30Dec 24, 2025Updated 8 months ago
- WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario (COLING 2025)☆13Jan 5, 2025Updated last year
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago
- Code for the paper "Controllable Video Captioning with an Exemplar Sentence"☆12Apr 14, 2021Updated 5 years ago
- mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local☆29Mar 9, 2026Updated 5 months ago
- ☆52Apr 20, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- TASU: A New Style of Alignment of Speech LLM with only Text Training Data, zero-shot on ASR and Other SU tasks☆28Jul 20, 2026Updated last month
- nlp processing ( pos-tag, parsing , ner , coref resolution) using NLTK Stanford nlp☆10Jul 28, 2017Updated 9 years ago
- ☆14Jul 5, 2024Updated 2 years ago
- Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations☆22Dec 24, 2025Updated 8 months ago
- ☆10Aug 17, 2021Updated 5 years ago
- OccCasNet: Occlusion-aware Cascade Cost Volume for Light Field Depth Estimation☆11May 28, 2023Updated 3 years ago
- [ICML 2026] Prism: Spectral-Aware Block-Sparse Attention☆27May 22, 2026Updated 3 months ago
- The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems☆29Mar 3, 2026Updated 6 months ago
- Code for the Avey-B paper (https://arxiv.org/abs/2602.15814)☆32Feb 21, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Agent-Driven Software Development Lifecycle (AD-SDLC) system built with Claude Agent SDK☆31Updated this week
- Code and data for "Medical Dialogue Generation via Dual Flow Modeling" (ACL 2023 Findings)☆14Nov 22, 2023Updated 2 years ago
- [NeurIPS 2024] Code, Dataset, Samples for the VATT paper “ Tell What You Hear From What You See - Video to Audio Generation Through Text”☆38Jul 24, 2025Updated last year
- [TVCG 2024] Official implementation of "JIMR: Joint Semantic and Geometry Learning for Point Scene Instance Mesh Reconstruction”☆15Jan 7, 2026Updated 7 months ago
- Codes for our paper "Enhancing Continual Relation Extraction via Classifier Decomposition" (Findings of ACL2023)☆10Nov 29, 2023Updated 2 years ago
- BLSP-Emo: Towards Empathetic Large Speech-Language Models☆61Jun 7, 2024Updated 2 years ago
- ☆16Apr 22, 2019Updated 7 years ago