☆28Jan 22, 2026Updated 7 months ago
Alternatives and similar repositories for FutureOmni
Users that are interested in FutureOmni are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Institutional OpenClaw Solution. Share One Claw with Others.☆25Mar 30, 2026Updated 5 months ago
- ☆25Jan 29, 2026Updated 7 months ago
- Official Repository for paper "HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding" [ACL 2026]☆101May 8, 2026Updated 3 months ago
- A python tool help to interact with chatgpt.☆10Dec 11, 2022Updated 3 years ago
- WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs☆51Jul 12, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- FamilyTool benchmark☆14Sep 10, 2025Updated 11 months ago
- MOSS-VL is the core multimodal model series within the OpenMOSS ecosystem, dedicated to visual understanding.☆486Updated this week
- ☆145Jun 24, 2026Updated 2 months ago
- We introduce 'Thinking with Video', a new paradigm leveraging video generation for multimodal reasoning. Our VideoThinkBench shows that S…☆320Aug 23, 2026Updated last week
- Exchange-of-Thought: Enhancing Large Language Model Capabilities through Cross-Model Communication☆21Mar 21, 2024Updated 2 years ago
- Official implementation of TDC.☆15Jul 22, 2025Updated last year
- [CVPR 2026] OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models☆105Apr 20, 2026Updated 4 months ago
- A curated list of awesome resources about reward construction for AI agents. This repository covers cutting-edge research, and practical …☆61Sep 1, 2025Updated 11 months ago
- ☆29Oct 16, 2025Updated 10 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [AAAI26] LongLLaDA: Unlocking Long Context Capabilities in Diffusion LLMs☆56Dec 7, 2025Updated 8 months ago
- [EMNLP 2026 Main] Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding☆28Aug 21, 2026Updated last week
- We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervis…☆432Jul 22, 2026Updated last month
- [NeurIPS 25] InfiniPot-V: Memory-Constrained KV Cache Compression for Streaming Video Understanding☆23Jan 25, 2026Updated 7 months ago
- Official code of "RoboOmni: Proactive Robot Manipulation in Omni-modal Context"☆119Mar 28, 2026Updated 5 months ago
- The official implementation of the paper "MotifRetro: Exploring the Combinability-Consistency Trade-offs in retrosynthesis via Dynamic Mo…☆11Jun 25, 2023Updated 3 years ago
- ☆165Mar 30, 2026Updated 5 months ago
- [ACL2025 main] Official implementation of "LED-Merging: Mitigating Safety-Utility Conflicts in Model Merging with Location-Election-Disjo…☆20Mar 16, 2026Updated 5 months ago
- ☆32Mar 17, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACL 2026 Findings] "Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning"☆63May 26, 2026Updated 3 months ago
- [CVPR 2026] Wavelet-based Frame Selection by Detecting Semantic Boundary for Long Video Understanding☆33Apr 12, 2026Updated 4 months ago
- [ECAI 2023] QCCDM: A Q-Augmented Causal Cognitive Diagnosis Model for Student Learning☆12Aug 4, 2023Updated 3 years ago
- Public codebase for ECONET: EMNLP'21☆12Mar 11, 2022Updated 4 years ago
- ☆40Jan 4, 2026Updated 7 months ago
- Code for paper: “What Data Benefits My Classifier?” Enhancing Model Performance and Interpretability through Influence-Based Data Selecti…☆23May 17, 2024Updated 2 years ago
- ☆20Jan 29, 2026Updated 7 months ago
- Official Repository for NeurIPS'25 Paper "Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task"☆23May 18, 2026Updated 3 months ago
- ☆29Apr 8, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MOVA: Towards Scalable and Synchronized Video–Audio Generation☆1,105Updated this week
- ☆14Feb 26, 2024Updated 2 years ago
- [CVPR 2025] OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts☆20Apr 2, 2025Updated last year
- Some simple tutorials about python☆12Oct 11, 2020Updated 5 years ago
- On Path to Multimodal Generalist: General-Level and General-Bench☆22Jul 11, 2025Updated last year
- Diffusion-based generative drug-like molecular editing with chemical natural language☆18Dec 22, 2024Updated last year
- ☆32Feb 27, 2025Updated last year