☆18Aug 23, 2022Updated 4 years ago
Alternatives and similar repositories for Official-ConvMAE-Det
Users that are interested in Official-ConvMAE-Det are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The multi-view version of MonoDETR on nuScenes dataset☆21Nov 4, 2022Updated 3 years ago
- [Codes of paper]: Region-based Non-local operation for Video Classification☆18Nov 28, 2021Updated 4 years ago
- Training LLaMA language model with MMEngine! It supports LoRA fine-tuning!☆40Apr 2, 2023Updated 3 years ago
- ☆16Jul 6, 2023Updated 3 years ago
- General Vision Benchmark, GV-B, a project from OpenGVLab☆186Feb 23, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2024] Data and benchmark code for the EgoExoLearn dataset☆88Aug 26, 2025Updated last year
- [CVPR 2023] Official repository for paper "Stare at What You See: Masked Image Modeling without Reconstruction"☆72Jul 2, 2025Updated last year
- VisualGPTScore for visio-linguistic reasoning☆27Oct 7, 2023Updated 2 years ago
- ConvMAE: Masked Convolution Meets Masked Autoencoders☆531Mar 14, 2023Updated 3 years ago
- ☆18Feb 13, 2026Updated 7 months ago
- An object detection codebase based on MegEngine.☆28Dec 14, 2022Updated 3 years ago
- Official repository of paper: "FeatAug-DETR: Enriching One-to-Many Matching for DETRs with Feature Augmentation"☆26Mar 2, 2023Updated 3 years ago
- Maximize the Resolution Potential of Pre-trained Rectified Flow Transformers☆66Oct 16, 2024Updated last year
- code of [CVPR22] CodedVTR: Codebook-based Sparse Voxel Transformer with Geometric Guidance☆18Jul 10, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [CVPR 2023] RILS: Masked Visual Reconstruction in Language Semantic Space (https://arxiv.org/abs/2301.06958)☆44Sep 5, 2023Updated 3 years ago
- ☆169Oct 14, 2021Updated 4 years ago
- [ICLR 2024 Spotlight] Bounding Box Stability against Feature Dropout Reflects Detector Generalization across Environments☆20Aug 19, 2025Updated last year
- Composition of Multimodal Language Models From Scratch☆15Aug 16, 2024Updated 2 years ago
- PyTorch implementation of Refine and Represent: Region-to-Object Representation Learning.☆21Jun 19, 2025Updated last year
- ☆61Jun 17, 2022Updated 4 years ago
- Improving Classifiers via Internal Augmentation☆16Apr 8, 2021Updated 5 years ago
- ☆70Jun 9, 2026Updated 3 months ago
- ☆25Jun 24, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Unofficial Paddle implementation of "Swin Transformer V2: Scaling Up Capacity and Resolution"☆33Nov 28, 2021Updated 4 years ago
- A Python Open CV project which highlights vacant parking spaces and also calculates no. of cars that can be accommodated in the given par…☆10Nov 17, 2019Updated 6 years ago
- 生僻字OCR识别优化训练☆17Feb 16, 2023Updated 3 years ago
- Unofficial implement of "Pix2seq: A Language Modeling Framework for Object Detection" on mmdetection☆34Apr 18, 2022Updated 4 years ago
- An Examination of the Compositionality of Large Generative Vision-Language Models☆19Apr 9, 2024Updated 2 years ago
- [NIPS2023]Implementation of Foundation Model is Efficient Multimodal Multitask Model Selector☆37Mar 7, 2024Updated 2 years ago
- ☆53May 3, 2023Updated 3 years ago
- 可以随机生成制定数量的车牌号,因为用到停车场的虚假数据生成,所以地区集中在一个地方。支 持各类车辆的生成,只需在注释的地方修改即可。☆10May 30, 2021Updated 5 years ago
- Official code for PixMamba☆40Feb 5, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 基于Paddle的手势关键点检测☆13Aug 22, 2021Updated 5 years ago
- [ECCV 2022] AMixer: Adaptive Weight Mixing for Self-attention Free Vision Transformers☆29Nov 14, 2022Updated 3 years ago
- This repo contains the code and configuration files for reproducing object detection results of FocalNets with DINO☆68Mar 10, 2023Updated 3 years ago
- [CVPR 2023]Implementation of Siamese Image Modeling for Self-Supervised Vision Representation Learning☆41Jun 6, 2024Updated 2 years ago
- Proteus (ICLR2025)☆61Mar 26, 2025Updated last year
- ☆28Nov 24, 2020Updated 5 years ago
- This is official implementation of KCR.☆22Aug 17, 2023Updated 3 years ago