[NeurIPS 2025] HoPE: Hybrid of Position Embedding for Long Context Vision-Language Models
☆30Feb 19, 2026Updated 7 months ago
Alternatives and similar repositories for HoPE
Users that are interested in HoPE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding☆61Apr 4, 2026Updated 5 months ago
- [ICML 2025 Oral] An official implementation of VideoRoPE & VideoRoPE++☆224Apr 15, 2026Updated 5 months ago
- Official implementation of Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More☆25Feb 25, 2025Updated last year
- ☆14Jun 13, 2025Updated last year
- Methods and code for extending the context length of diffusion language models☆56Dec 7, 2025Updated 9 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Implementation for "The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer"☆86Oct 29, 2025Updated 10 months ago
- ☆11Jun 11, 2025Updated last year
- Official implementation of TDC.☆15Jul 22, 2025Updated last year
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆19Mar 15, 2024Updated 2 years ago
- The official implementation of ICLR 2025 paper "Polynomial Composition Activations: Unleashing the Dynamics of Large Language Models".☆18Apr 25, 2025Updated last year
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 6 months ago
- [NeurIPS 2024] | An Efficient Recipe for Long Context Extension via Middle-Focused Positional Encoding☆23Oct 10, 2024Updated last year
- ☆30May 22, 2026Updated 4 months ago
- This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception o…☆29Jul 9, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Visualisation of VISOR Segmentations with Annotations and Relations☆22Aug 15, 2022Updated 4 years ago
- [ICLR 2026] Rectifying LLM Thought From Lens of Optimization☆14Dec 5, 2025Updated 9 months ago
- The official repository of "R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Integration"☆143Sep 4, 2025Updated last year
- ☆32Jun 24, 2024Updated 2 years ago
- Implementation of "VQ-HPS: Human Pose and Shape Estimation in a Vector-Quantized Latent Space" - ECCV 2024☆14Mar 24, 2025Updated last year
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- ☆13Apr 23, 2025Updated last year
- Suri: Multi-constraint instruction following for long-form text generation [EMNLP’24]☆27Oct 3, 2025Updated 11 months ago
- ☆21Mar 6, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Robot-agent harness with a CLI and local Web console for LIBERO short evaluation☆95Updated this week
- ☆29Feb 27, 2025Updated last year
- An Experiment on Dynamic NTK Scaling RoPE☆64Nov 26, 2023Updated 2 years ago
- Official repo for Sym2Real: Symbolic Dynamics with Residual Learning for Data-Efficient Adaptive Control☆15Sep 22, 2025Updated last year
- [COLM 2026] Resa: Transparent Reasoning Models via SAEs☆49Sep 23, 2025Updated last year
- [ICML 2025] Fourier Position Embedding: Enhancing Attention’s Periodic Extension for Length Generalization☆120Jun 2, 2025Updated last year
- The repository for papaer "Distance between Relevant Information Pieces Causes Bias in Long-Context LLMs"☆14Dec 16, 2024Updated last year
- UEval: A Benchmark for Unified Multimodal Generation☆26Apr 20, 2026Updated 5 months ago
- Code and data for paper "Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation".☆25Oct 22, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2024] Hybrid Proposal Refiner: Revisiting DETR Series from the Faster R-CNN Perspective☆20Aug 18, 2024Updated 2 years ago
- Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning☆84May 17, 2025Updated last year
- Official PyTorch implementation of "Let RGB Be the Language of Vision".☆49Jul 16, 2026Updated 2 months ago
- [ICML 2025] LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models☆16Nov 4, 2025Updated 10 months ago
- Official code of "Edit Transfer: Learning Image Editing via Vision In-Context Relations"☆89Jun 6, 2025Updated last year
- Learning 1D Causal Visual Representation with De-focus Attention Networks☆35Jun 7, 2024Updated 2 years ago
- ☆18Aug 5, 2025Updated last year