[ICLR 2024] Official PyTorch implementation of FasterViT: Fast Vision Transformers with Hierarchical Attention
☆925Jul 22, 2025Updated last year
Alternatives and similar repositories for FasterViT
Users that are interested in FasterViT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2023] Official PyTorch implementation of Global Context Vision Transformers☆449Dec 22, 2023Updated 2 years ago
- Efficient vision foundation models for high-resolution generation and perception.☆3,369Sep 5, 2025Updated last year
- This repository contains the official implementation of the research paper, "FastViT: A Fast Hybrid Vision Transformer using Structural R…☆2,039Sep 11, 2026Updated 3 weeks ago
- Hiera: A fast, powerful, and simple hierarchical vision transformer.☆1,077Mar 2, 2024Updated 2 years ago
- RepViT: Revisiting Mobile CNN From ViT Perspective [CVPR 2024] and RepViT-SAM: Towards Real-Time Segmenting Anything☆1,127Jun 14, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Neighborhood Attention Transformer, arxiv 2022 / CVPR 2023. Dilated Neighborhood Attention Transformer, arxiv 2022☆1,186May 15, 2024Updated 2 years ago
- [CVPR 2025] Official PyTorch Implementation of MambaVision: A Hybrid Mamba-Transformer Vision Backbone☆2,236Mar 11, 2026Updated 6 months ago
- EfficientFormerV2 [ICCV 2023] & EfficientFormer [NeurIPs 2022]☆1,116Aug 13, 2023Updated 3 years ago
- Code release for ConvNeXt V2 model☆2,067Aug 14, 2024Updated 2 years ago
- Official repository for "AM-RADIO: Reduce All Domains Into One"☆1,970Updated this week
- This repository contains the official implementation of the research paper, "An Improved One millisecond Mobile Backbone" CVPR 2023.☆831Sep 11, 2026Updated 3 weeks ago
- PyTorch code and models for the DINOv2 self-supervised learning method.☆13,405Updated this week
- This is a collection of our NAS and Vision Transformer work.☆1,845Jul 25, 2024Updated 2 years ago
- [CVPR 2023] OneFormer: One Transformer to Rule Universal Image Segmentation☆1,739Oct 3, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- PoolFormer: MetaFormer Is Actually What You Need for Vision (CVPR 2022 Oral)☆1,365Jun 1, 2024Updated 2 years ago
- [NeurIPS 2022] Official code for "Focal Modulation Networks"☆749Nov 7, 2023Updated 2 years ago
- MetaFormer Baselines for Vision (TPAMI 2024)☆504Jun 1, 2024Updated 2 years ago
- [ICML 2024] Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model☆3,908Feb 13, 2025Updated last year
- Fast Segment Anything☆8,421Jul 30, 2024Updated 2 years ago
- CVNets: A library for training computer vision networks☆1,994Sep 11, 2026Updated 3 weeks ago
- [ECCV 2022] Official repository for "MaxViT: Multi-Axis Vision Transformer". SOTA foundation models for classification, detection, segmen…☆501Jun 2, 2023Updated 3 years ago
- This is the official code for MobileSAM project that makes SAM lightweight for mobile applications and beyond!☆5,886May 5, 2026Updated 5 months ago
- Official PyTorch implementation of Fully Attentional Networks☆483Mar 31, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official code for "FeatUp: A Model-Agnostic Frameworkfor Features at Any Resolution" ICLR 2024☆1,656Jun 28, 2024Updated 2 years ago
- Code release for ConvNeXt model☆6,409Jan 8, 2023Updated 3 years ago
- [ICLR 2023] Official implementation of the paper "DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection"☆2,848Jul 31, 2024Updated 2 years ago
- ☆587Jul 23, 2023Updated 3 years ago
- [CVPR 2023 Highlight] InternImage: Exploring Large-Scale Vision Foundation Models with Deformable Convolutions☆2,860Mar 25, 2025Updated last year
- [ICCV - 2023] Official repository of paper SwiftFormer: Efficient Additive Attention for Transformer-based Real-time Mobile Vision Applic…☆317Jul 18, 2025Updated last year
- A method to increase the speed and lower the memory footprint of existing vision transformers.☆1,204Jun 17, 2024Updated 2 years ago
- The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights --…☆37,208Updated this week
- EVA Series: Visual Representation Fantasies from BAAI☆2,694Aug 1, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2023] Official implementation of the paper "Lite DETR : An Interleaved Multi-Scale Encoder for Efficient DETR"☆210Jun 1, 2023Updated 3 years ago
- [CVPR 2024] Deformable Convolution v4☆747May 17, 2024Updated 2 years ago
- [CVPR 2023] Official Implementation of X-Decoder for generalized decoding for pixel, image and language☆1,344Oct 5, 2023Updated 3 years ago
- [ICLR 2023 Spotlight] Vision Transformer Adapter for Dense Predictions☆1,505Jun 3, 2025Updated last year
- An ultimately comprehensive paper list of Vision Transformer/Attention, including papers, codes, and related websites☆5,043Jul 30, 2024Updated 2 years ago
- [CVPR 2024] Official RT-DETR (RTDETR paddle pytorch), Real-Time DEtection TRansformer, DETRs Beat YOLOs on Real-time Object Detection. 🔥…☆5,585Updated this week
- Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"☆8,677May 31, 2024Updated 2 years ago