The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems
☆29Mar 3, 2026Updated 5 months ago
Alternatives and similar repositories for heterogeneous-latent-mas
Users that are interested in heterogeneous-latent-mas are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository open-sources our GEC system submitted by THU KELab (sz) in the CCL2023-CLTC Track 1: Multidimensional Chinese Learner Tex…☆15Nov 25, 2023Updated 2 years ago
- Simple Telegram bot Framework☆10Apr 8, 2017Updated 9 years ago
- GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant☆19Apr 30, 2026Updated 4 months ago
- 模拟东北大学教务处网站登录 并获取全部学生信息 目前可能随着教务处网站的更新变得不可用☆11Mar 2, 2019Updated 7 years ago
- ImageNet3D: Towards General-Purpose Object-Level 3D Understanding☆22Dec 6, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICLR 2026] Official code for BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step Generation☆23Apr 13, 2026Updated 4 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Jun 17, 2026Updated 2 months ago
- [ICML 2025 Spotlight] Official implementation of the paper: Re-ranking Reasoning Context with Tree Search Makes Large Vision-Language Mod…☆19Sep 1, 2025Updated 11 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆46Oct 3, 2025Updated 10 months ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆88May 12, 2026Updated 3 months ago
- Code release for "Category-Specific Prompts for Animal Action Recognition with Pretrained Vision-Language Models"☆14Feb 21, 2024Updated 2 years ago
- [ICCV 2021] Multimodal Knowledge Expansion☆10Aug 28, 2021Updated 5 years ago
- RENT (Reinforcement Learning via Entropy Minimization) is an unsupervised method for training reasoning LLMs.☆42Oct 31, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for "Retaining Key Information under High Compression Rates: Query-Guided Compressor for LLMs" (ACL 2024)☆19Jun 12, 2024Updated 2 years ago
- THEORY OF SPACE: a benchmark for evaluating whether foundation models can actively explore under partial observability efficiently to bui…☆86Feb 27, 2026Updated 6 months ago
- This code is provided for reproducibility of results in the paper: Multiview Aerial Visual Recognition (MAVREC): Can Multi-view Improve A…☆24Feb 6, 2025Updated last year
- A paper list of Awesome Latent Space.☆961Jul 13, 2026Updated last month
- [ICML 2026] Prism: Spectral-Aware Block-Sparse Attention☆27May 22, 2026Updated 3 months ago
- ☆12Jun 12, 2024Updated 2 years ago
- Video-Text Representation Learning via Differentiable Weak Temporal Alignment (PyTorch implementation for the CVPR 2022 paper)☆11Oct 12, 2022Updated 3 years ago
- [ICCV 2025] A Benchmark for Multi-Step Reasoning in Long Narrative Videos☆28Jun 4, 2026Updated 2 months ago
- Official implementation of "Figure It Out: Improve the Frontier of Reasoning with Active Visual Thinking"☆17Jan 13, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 26 ORAL] When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models☆30Jun 30, 2026Updated 2 months ago
- A benchmark for testing memorization abilities of LMs☆24Oct 15, 2024Updated last year
- 面向大模型的民族文化数据集☆14May 26, 2025Updated last year
- [ Arxiv 2023 ] This repository contains the code for "MUPPET: Multi-Modal Few-Shot Temporal Action Detection"☆16Aug 30, 2023Updated 3 years ago
- [ACM MM 2022] MM_Pyramid: Multimodal Pyramid Attentional Network for Audio-Visual Event Localization and Video Parsing☆15Aug 26, 2022Updated 4 years ago
- High Performance Grouped GEMM in PyTorch☆30May 10, 2022Updated 4 years ago
- ☆22Sep 16, 2025Updated 11 months ago
- Governance substrate for your AI coding agents — adversarial review, drift-detected rules, immutable audit, closed-loop telemetry☆18Updated this week
- [ICLR 2025] Data-Augmented Phrase-Level Alignment for Mitigating Object Hallucination☆21Jan 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2025] The official implementation of paper "Safe-Sora: Safe Text-to-Video Generation via Graphical Watermarking"☆20Oct 10, 2025Updated 10 months ago
- ☆14Nov 13, 2023Updated 2 years ago
- ☆19Jun 6, 2025Updated last year
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videos☆24May 7, 2026Updated 3 months ago
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆52Jul 7, 2026Updated last month
- An arbitrage bot is a smart contract connected to an external automation script that controls its operation.☆2,691Updated this week