[CVPR 2024] Asymmetric Masked Distillation for Pre-Training Small Foundation Models
☆18Jan 11, 2026Updated 7 months ago
Alternatives and similar repositories for AMD
Users that are interested in AMD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2023] Masked Video Distillation: Rethinking Masked Feature Modeling for Self-supervised Video Representation Learning (https://arxiv…☆135May 21, 2023Updated 3 years ago
- [IJCV] Progressive Visual Prompt Learning with Contrastive Feature Re-formation☆15Aug 10, 2024Updated 2 years ago
- [T-PAMI 2023] Temporal Perceiver: A General Architecture for Arbitrary Boundary Detection☆39Aug 29, 2023Updated 2 years ago
- [ICML2026] FreeRet: MLLMs as Training-Free Retrievers☆24May 25, 2026Updated 2 months ago
- [ICCV 2025] p-MoD: Building Mixture-of-Depths MLLMs via Progressive Ratio Decay☆43Jun 26, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Fine-grained Benchmark for Video Captioning and Retrieval☆30Jul 16, 2025Updated last year
- [ICCV 2023] MGMAE: Motion Guided Masking for Video Masked Autoencoding☆26Oct 16, 2023Updated 2 years ago
- ☆20Dec 24, 2025Updated 7 months ago
- Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval☆16Nov 29, 2025Updated 8 months ago
- ☆10Dec 3, 2024Updated last year
- [AAAI 2023 Oral] CoMAE: Single Model Hybrid Pre-training on Small-Scale RGB-D Datasets☆37Aug 20, 2024Updated last year
- [NeurIPS 2022] PointTAD: Multi-Label Temporal Action Detection with Learnable Query Points☆48Nov 24, 2023Updated 2 years ago
- Code for Semantic-Aware Dynamic Generation Networks for Few-Shot Human-Object Interaction Recognition☆10May 26, 2021Updated 5 years ago
- [ECCV 2022] Joint-Modal Label Denoising for Weakly-Supervised Audio-Visual Video Parsing☆27Jul 15, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2021] CGA-Net: Category Guided Aggregation for Point Cloud Semantic Segmentation☆24Jan 30, 2022Updated 4 years ago
- Generate a denotation graph from a set of image captions☆16Sep 4, 2018Updated 7 years ago
- [CVPR 2024] SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos☆18May 21, 2024Updated 2 years ago
- [CVPR 2023] This repository includes the official implementation our paper "Masked Autoencoders Enable Efficient Knowledge Distillers"☆109Jul 24, 2023Updated 3 years ago
- ☆11Jan 8, 2022Updated 4 years ago
- A visualization tool for temporal action localization (detection/segmentation).☆13Mar 30, 2023Updated 3 years ago
- ☆12Sep 11, 2021Updated 4 years ago
- a pytorch implement of Supervised Contrastive Learning with memory bank(queue)☆15May 25, 2022Updated 4 years ago
- [ACM MM 2023] PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose Estimation☆12Aug 28, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Aug 29, 2019Updated 6 years ago
- DIVA: A Dirichlet Process Mixtures Based Incremental Deep Clustering Algorithm via Variational Auto-Encoder☆13Sep 18, 2023Updated 2 years ago
- The official codebase of FineAction dataset. We will update the data and code of our FineAction.☆24Apr 10, 2025Updated last year
- ☆58Dec 2, 2025Updated 8 months ago
- ☆16Jul 6, 2023Updated 3 years ago
- Implementation of GPU-friendly differentiable DLT transform proposed in "Lightweight Multi-View 3D Pose Estimation through Camera-Disenta…☆13Sep 15, 2020Updated 5 years ago
- Custom layers for pytorch☆15Mar 16, 2024Updated 2 years ago
- [AAAI2023] Revisiting the Spatial and Temporal Modeling for Few-shot Action Recognition (SloshNet)☆14Jan 10, 2024Updated 2 years ago
- ☆16Aug 5, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AIRS-2025赛道二:「星际矿脉」火星矿物高光谱分类挑战赛☆13May 7, 2025Updated last year
- [TPAMI 2024] Dynamic MDETR: A Dynamic Multimodal Transformer Decoder for Visual Grounding☆29Sep 11, 2024Updated last year
- ☆16Dec 15, 2025Updated 8 months ago
- This repository is related to 'Intriguing Properties of Hyperbolic Embeddings in Vision-Language Models', published at TMLR (2024), https…☆21Jul 5, 2024Updated 2 years ago
- [NeurIPS 2023] MixFormerV2: Efficient Fully Transformer Tracking☆227Apr 20, 2024Updated 2 years ago
- Code and model for "Multi-dataset Training of Transformers for Robust Action Recognition", NeurIPS 2022 Spotlight☆20Aug 1, 2023Updated 3 years ago
- A Graph Attention Spatio-temporal Convolutional Networks for 3D Human Pose Estimation in Video (GAST-Net)☆14Dec 2, 2020Updated 5 years ago