Reproduction of the first step in the text-to-video model Phenaki. Code and model weights for the Transformer-based autoencoder for videos called CViViT.
☆29Aug 4, 2023Updated 2 years ago
Alternatives and similar repositories for phenaki-cvivit
Users that are interested in phenaki-cvivit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Dec 14, 2024Updated last year
- SFT+RL boosts multimodal reasoning☆47Jun 27, 2025Updated last year
- Region Proposal generation on images using clustering in Pointcloud - Currently only for Pedestrians☆11Jul 13, 2020Updated 6 years ago
- Implementation of MagViT2 Tokenizer in Pytorch☆668Jan 12, 2025Updated last year
- This repository contains tools for visualization of keypoint matches over two images (ORB, SIFT, LIFT, SuperPoint, D2-Net).☆13Jul 23, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation Models, ICCV 2023☆13Jul 13, 2024Updated 2 years ago
- This repo consist of some experimental results on bdd100k datasets using different object detection algorithms(Faster-RCNN, FCOS, ATSS)☆11Jun 27, 2020Updated 6 years ago
- ☆14Aug 9, 2024Updated last year
- Lidar line downsampling for KITTI dataset, transfer lidar the number of lidar lines from 64 to 32, 16, 8, etc.☆13Jun 3, 2020Updated 6 years ago
- Pytorch implementation of deep fill v2 (original by Jiayu et al.)☆10Jun 26, 2019Updated 7 years ago
- Unofficial implement of "Pix2seq: A Language Modeling Framework for Object Detection" on mmdetection☆34Apr 18, 2022Updated 4 years ago
- ☆132Feb 22, 2025Updated last year
- code for the paper Imitation Learning from Observation with Automatic Discount Scheduling☆13Mar 27, 2024Updated 2 years ago
- Official implementation of MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis☆86Jul 16, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Main code of Dolphins dataset☆16Dec 29, 2022Updated 3 years ago
- This is a toolbox repository to help evaluate various methods that perform image matching from a pair of images.☆12Jul 5, 2023Updated 3 years ago
- ☆12Oct 12, 2020Updated 5 years ago
- SEED-Voken: A Series of Powerful Visual Tokenizers☆1,017Nov 25, 2025Updated 7 months ago
- RAST 1.0: Restorable Arbitrary Style Transfer via Multi-restoration☆13Jun 18, 2024Updated 2 years ago
- This is an OCR program designed for travel document. It can now support 23 types of documents with pre-defined template. You can add what…☆10Nov 22, 2022Updated 3 years ago
- Style Transfer by Deep Learning, overview and TensorFlow implementations (UNDER CONSTRUCTION)☆14Jul 25, 2017Updated 8 years ago
- ☆89Jan 4, 2024Updated 2 years ago
- Google MobileNets Implementation using Tensorflow☆18Jun 6, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A partial implementation of Generative Infinite Vocabulary Transformer (GIVT) from Google Deepmind, in PyTorch.☆21Mar 28, 2024Updated 2 years ago
- ☆38Dec 25, 2025Updated 6 months ago
- A PyTorch re-implementation of Weakly Supervised Facial Action Unit Recognition through Adversarial Training☆10Apr 23, 2019Updated 7 years ago
- Annotated Tutorial for PerAct☆19Sep 11, 2023Updated 2 years ago
- ☆41Sep 21, 2023Updated 2 years ago
- Unofficial Pytorch Implementation of "A Simple Framework for Contrastive Learning of Visual Representations"☆10Mar 11, 2020Updated 6 years ago
- EgoToM is an egocentric theory-of-mind benchmark built on Ego4D videos, containing multi-choice questions that evaluate multimodal large …☆16Apr 1, 2025Updated last year
- This is the official implementation of paper "Evaluate and Improve the Quality of Neural Style Transfer" (CVIU 2021))☆11Feb 14, 2022Updated 4 years ago
- Framework to achieve context distillation in LLMs☆15Nov 24, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Multi-temporal Scene dataset for Scene Change Detection.☆15Apr 14, 2021Updated 5 years ago
- [NeurIPS 2024] Data exporter for SS3DM: Benchmarking Street-View Surface Reconstruction with a Synthetic 3D Mesh Dataset☆16Nov 8, 2024Updated last year
- Official JAX implementation of MAGVIT: Masked Generative Video Transformer☆1,002Jan 17, 2024Updated 2 years ago
- Toolkit for VIPER benchmark☆16Aug 11, 2020Updated 5 years ago
- ☆10Jan 20, 2021Updated 5 years ago
- A list of robotics related papers accepted by ICLR'25☆25Aug 28, 2025Updated 10 months ago
- ☆28Feb 7, 2024Updated 2 years ago