Implementation of ViViT: A Video Vision Transformer - Zipping Coding Challenge
☆32Jun 10, 2021Updated 5 years ago
Alternatives and similar repositories for vivit_pytorch
Users that are interested in vivit_pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆69Apr 26, 2021Updated 5 years ago
- Implementation of ViViT: A Video Vision Transformer☆560Jun 21, 2021Updated 5 years ago
- Open-source evaluation toolkit of large vision-language models (LVLMs), support ~100 VLMs, 30+ benchmarks☆15Feb 17, 2025Updated last year
- Load and visualize different datasets in video question answering☆10May 11, 2021Updated 5 years ago
- Use tensorflow object_detection to finetuning a Mask R-CNN model☆22Oct 1, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Excitation Backprop for RNNs☆15Jul 25, 2018Updated 8 years ago
- HEtero-Assists Distillation for Heterogeneous Object Detectors☆10Jul 3, 2023Updated 3 years ago
- ☆15Apr 26, 2024Updated 2 years ago
- Unoffitial implementation of the MICLe in pytorch.☆12Dec 3, 2022Updated 3 years ago
- Music large model based on InternLM2-chat.☆22Dec 21, 2024Updated last year
- 基于pycorrector以及chatglm3-6b的文本纠错☆12Mar 10, 2024Updated 2 years ago
- 你是否是 每天都在等待你最喜欢的YouTuber发布新视频?你想把他们的精彩内容分享到中国最大的视频分享平台Bilibili上吗? 那么就来试试YouTube2Bili 📹🚀🚀 - 一个能够从任何你想要的YouTuber那里下载视频并上传到Bilibili的Pytho…☆10Apr 14, 2023Updated 3 years ago
- TCM: Temporal Correlation Module☆17Apr 24, 2021Updated 5 years ago
- ☆12Apr 6, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for TIP 2024 paper: Sparse Coding Inspired LSTM and Self-Attention Integration for Medical Image Segmentation☆13Oct 28, 2024Updated last year
- An implementation of the paper "End-to-End Human-Gaze-Target Detection with Transformers"☆21Jul 23, 2026Updated 2 months ago
- The code for AAAI 2025 “Large Language Models Are Read/Write Policy-Makers for Simultaneous Generation”☆15Jan 3, 2025Updated last year
- Pytorch code for ECCVW 2022 paper "Consistency-based Self-supervised Learning for Temporal Anomaly Localization"☆14Jul 9, 2024Updated 2 years ago
- ICML2019 Accepted Paper. Overcoming Multi-Model Forgetting☆14Jun 5, 2019Updated 7 years ago
- ☆12Sep 1, 2024Updated 2 years ago
- Unsupervised Model-based Dense Face Alignment☆10Aug 27, 2020Updated 6 years ago
- ☆22Nov 4, 2024Updated last year
- Revisiting Light Field Rendering with Deep Anti-Aliasing Neural Network☆14Nov 27, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code used in the paper "Learning to Learn from Web Data through Deep Semantic Embeddings" ECCV 2018 MULA Workshop☆11Aug 1, 2018Updated 8 years ago
- Abdominal Organ Segmentation using Multi Decoder Network (MDNet) [Accepted at ICASSP 2025]☆13Apr 15, 2025Updated last year
- ☆16Mar 12, 2025Updated last year
- Fast instruction tuning with Llama2☆10Apr 8, 2024Updated 2 years ago
- The speaker-labeled information of LRW dataset, which is the outcome of the paper "Speaker-adaptive Lip Reading with User-dependent Paddi…☆10Oct 12, 2023Updated 2 years ago
- AFNet(NeurIPS 2022)☆20Nov 24, 2022Updated 3 years ago
- An unofficial (PyTorch) implementation for the paper Deep Lip Reading: A comparison of models and an online application.☆10May 13, 2020Updated 6 years ago
- ☆15Apr 9, 2023Updated 3 years ago
- SA2-Net: Scale-aware Attention Network for Microscopic Image Segmentation (BMVC'23 -- Oral)☆18Dec 14, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official code for "FaceCom: Towards High-fidelity 3D Facial Shape Completion via Optimization and Inpainting Guidance", CVPR 2024.☆23Sep 11, 2024Updated 2 years ago
- Benchmarking Attention Mechanism in Vision Transformers.☆20Oct 10, 2022Updated 3 years ago
- Paper List for Dialogue and Interactive Systems☆16Jun 5, 2020Updated 6 years ago
- custom pytorch implementation of MoCo v3☆47Apr 7, 2021Updated 5 years ago
- CAPG-GAN for face frontalization task☆11May 28, 2020Updated 6 years ago
- 尝试将PP-LCNet转为PyTorch版本,并测试相关指标☆13Dec 3, 2022Updated 3 years ago
- Implementation of brand new video augmentation strategy for video action recognition with 3D CNN☆27Jun 26, 2021Updated 5 years ago