The official repo for [TPAMI'23] "Vision Transformer with Quadrangle Attention"
☆239Sep 25, 2025Updated 10 months ago
Alternatives and similar repositories for QFormer
Users that are interested in QFormer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repo for the technical report "Scalable Mask Annotation for Video Text Spotting"☆16May 3, 2023Updated 3 years ago
- The official repo for [ECCV'22] "VSA: Learning Varied-Size Window Attention in Vision Transformers"☆159Sep 25, 2025Updated 10 months ago
- The official repo for [ACM CSUR'24] "Empowering Agrifood System with Artificial Intelligence: A Survey of the Progress, Challenges and Op…☆12Dec 6, 2024Updated last year
- ☆18Jul 24, 2025Updated last year
- Code of our Neurips2020 paper "Auto Learning Attention", coming soon☆22Apr 14, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official repo for [NeurIPS'21] "ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive Bias" and [IJCV'22] "ViTAEv2: Vis…☆279Apr 15, 2026Updated 3 months ago
- [ICML 2026] Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding☆16Mar 13, 2026Updated 4 months ago
- (CVPR2023/TPAMI2024) Integrally Pre-Trained Transformer Pyramid Networks -- A Hierarchical Vision Transformer for Masked Image Modeling☆216Jul 28, 2024Updated 2 years ago
- Official repo for [CVPR 2026] "GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization"☆39May 14, 2026Updated 2 months ago
- Unofficial implementation for [ECCV'22] "Exploring Plain Vision Transformer Backbones for Object Detection"☆586Apr 24, 2022Updated 4 years ago
- Official repository for RealRain-1k☆33Jul 6, 2025Updated last year
- Repository of Vision Transformer with Deformable Attention (CVPR2022) and DAT++: Spatially Dynamic Vision Transformerwith Deformable Atte…☆940Apr 17, 2024Updated 2 years ago
- [JSTARS'26] S3RNet: Sparse Spatial--Spectral Representation with Hybrid Knowledge Distillation for Efficient Multispectral and Hyperspect…☆14Jun 16, 2026Updated last month
- CIFAR10 ResNets implemented in JAX+Flax☆12Apr 6, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The official PyTorch code for "Traffic Scene Parsing through the TSP6K Dataset".☆34Jul 6, 2025Updated last year
- The official pytorch implementation of ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive Bias☆104Apr 12, 2022Updated 4 years ago
- A comprehensive list [SAMRS@NeurIPS'23, RVSA@TGRS'22, RSP@TGRS'22] of our research works related to remote sensing, including papers, cod…☆486Jun 6, 2024Updated 2 years ago
- [CVPR 2023] Referring Image Matting☆208Apr 17, 2023Updated 3 years ago
- 😎 Awesome lists of papers and codes about Large Vision-Language Models☆13Apr 1, 2024Updated 2 years ago
- Official repo for "REX-RAG: Reasoning Exploration with Policy Correction in Retrieval-Augmented Generation"☆35Sep 28, 2025Updated 10 months ago
- [ECCV 2024] Official repository of Agent Attention☆669Nov 17, 2024Updated last year
- [ICCV 2023] Source code of "Fcaformer: Forward Cross Attention in Hybrid Vision Transformer"☆25Aug 23, 2023Updated 2 years ago
- Salient Objects in Clutter, arXiv, 2021 (ECCV2018 extenstion).☆11Jun 17, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official repo for [IEEE TGRS'26] "SPEX: A Vision-Language Model for Land Cover Extraction on Spectral Remote Sensing Images"☆28Mar 16, 2026Updated 4 months ago
- A Light-weight and Multi-scale Network for Medical Image Segmentation☆30Jun 8, 2025Updated last year
- [NeurIPS'24] GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching☆34May 29, 2025Updated last year
- [ECCV 2022]JPerceiver: Joint Perception Network for Depth, Pose and Layout Estimation in Driving Scenes☆79Nov 4, 2022Updated 3 years ago
- ☆136Jan 19, 2023Updated 3 years ago
- Code of the Grounded MUIE model, REAMO☆11Dec 3, 2024Updated last year
- [EMNLP22] Improving Sharpness-Aware Minimization with Fisher Mask for Better Generalization on Language Models☆22Mar 27, 2023Updated 3 years ago
- SuperpixelGridMasks is an approach for sensor-based data augmentation towards image classification tasks and so on.☆14Jan 18, 2023Updated 3 years ago
- [ICCV 2021] Official implementation of "Scalable Vision Transformers with Hierarchical Pooling"☆32Dec 30, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2024] Code release for TransNeXt model☆574Jun 13, 2024Updated 2 years ago
- ☆13Apr 19, 2024Updated 2 years ago
- VMamba: Visual State Space Models,code is based on mamba☆3,216Mar 7, 2025Updated last year
- [IJCAI 2023] CLE-ViT: Contrastive Learning Encoded Transformer for Ultra-Fine-Grained Visual Categorization.☆11Nov 3, 2023Updated 2 years ago
- ☆44Jun 18, 2026Updated last month
- ☆25Dec 19, 2024Updated last year
- Verify CPU circuits in Logisim or Verilog against MARS simulation☆10Dec 31, 2020Updated 5 years ago