Implementations of Transformers for Video
☆24Mar 26, 2021Updated 5 years ago
Alternatives and similar repositories for video-transformers
Users that are interested in video-transformers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code release to our paper on an agent-based model of the Ramsey-Cass-Koopmans macroeconomic model. In this model, the households imitate …☆13Jun 3, 2021Updated 5 years ago
- BLIND NATURAL IMAGE QUALITY PREDICTION USING CONVOLUTIONAL NEURAL NETWORKS AND WEIGHTED SPATIAL POOLING(ICIP2020)☆12Sep 30, 2020Updated 5 years ago
- Code for generating a single image pretraining dataset☆13Aug 3, 2021Updated 5 years ago
- a pytorch implementation for MoCo V3☆32Apr 14, 2021Updated 5 years ago
- Resources related to the model cards for ML☆11Mar 16, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Binary storage format for deep learning on videos.☆12Mar 10, 2022Updated 4 years ago
- ☆11Mar 25, 2024Updated 2 years ago
- 🎮 Use a Raspberry Pi to control a LoPy over UART☆12Mar 9, 2017Updated 9 years ago
- ☆20Feb 27, 2023Updated 3 years ago
- Implementation of Bitune: Bidirectional Instruction-Tuning☆27Jun 19, 2025Updated last year
- This is an image dataset in raw format.☆13Apr 14, 2021Updated 5 years ago
- Time Does Tell: Self-Supervised Time-Tuning of Dense Image Representations ICCV23☆31Dec 30, 2024Updated last year
- Deformable 3D ConvNets for Action Recognition☆10Jan 21, 2018Updated 8 years ago
- This project applies Monte Carlo Tree Search (MCTS) to a simple grid world.☆10May 30, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch demo code for "Spatial-Temporal Pyramid Based Convolutional Neural Network for Action Recognition"☆15Oct 17, 2018Updated 7 years ago
- The implementation of 'Watch, Listen, Attend and Spell’ (WLAS) network that learns to transcribe videos of mouth motion to character on p…☆11Mar 23, 2018Updated 8 years ago
- ☆25Dec 26, 2019Updated 6 years ago
- [CVPR2024] Learning from Synthetic Human Group Activities☆14Feb 24, 2025Updated last year
- Multimodal Speech Recognition for phoneme level prediction using Audio-Visual data from TCDTIMIT dataset implementing RNNs with LSTMs for…☆15Jul 27, 2023Updated 3 years ago
- ☆15Apr 27, 2017Updated 9 years ago
- The notebook explains the various steps to obtain the results of publication: "Is Space-Time Attention All You Need for Video Understandi…☆42Mar 19, 2021Updated 5 years ago
- Implementation for <Orthogonal Over-Parameterized Training> in CVPR'21.☆22Jul 16, 2021Updated 5 years ago
- Dual cross modality attention audio-visual speech recognition model based on vgg transformer with hybrid CTC/attention architecture using…☆14Jul 2, 2020Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆73Jun 3, 2022Updated 4 years ago
- Temporal Compact Bilinear Pooling (TCBP)☆11May 27, 2020Updated 6 years ago
- HARPER is a HRI dataset for 3D Human Pose Estimation and Forecasting from the Robot’s Perspective.☆13Sep 2, 2025Updated last year
- Official implementation of the ECCV2024 paper: Generalizable Facial Expression Recognition☆22Sep 20, 2024Updated last year
- a Meal App developed on flutter with firebase database integeration.☆10May 19, 2020Updated 6 years ago
- ☆160Jul 21, 2021Updated 5 years ago
- This repo covers the implementation for Labelling unlabelled videos from scratch with multi-modal self-supervision, which learns clusters…☆118Apr 26, 2021Updated 5 years ago
- Development kit for Pandora☆14Aug 4, 2020Updated 6 years ago
- Video Representation Learning by Recognizing Temporal Transformations. In ECCV, 2020.☆49Mar 18, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The official pytorch implementation of our paper "Is Space-Time Attention All You Need for Video Understanding?"☆1,863Apr 9, 2024Updated 2 years ago
- Implementation of TimeSformer from Facebook AI, a pure attention-based solution for video classification☆729Aug 25, 2021Updated 5 years ago
- [MM'20] Aesthetic-Aware Image Style Transfer☆15Sep 16, 2021Updated 4 years ago
- Learning to Count without Annotations☆24May 24, 2024Updated 2 years ago
- [NeurIPS 2024] CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition☆16Nov 12, 2025Updated 9 months ago
- Self-Calibrated Efficient Transformer for Lightweight Super-Resolution (offical)☆38Jun 7, 2022Updated 4 years ago
- UcosII 全源码解读☆14Dec 31, 2019Updated 6 years ago