Code for NeurIPS 2022 paper "Learning Viewpoint-Agnostic Visual Representations by Recovering Tokens in 3D Space"
☆20Apr 20, 2023Updated 3 years ago
Alternatives and similar repositories for 3DTRL
Users that are interested in 3DTRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Video + CLIP Baseline for Ego4D Long Term Action Anticipation Challenge (CVPR 2022)☆15Jul 4, 2022Updated 4 years ago
- Code for the paper Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers☆22Aug 2, 2024Updated 2 years ago
- This is a python library. Install with "python3 -m pip install rp" then run with "python3 -m rp" or just "rp". Requires python≥3.5☆13Jul 13, 2026Updated last month
- This is the offical repository of LLAVIDAL☆25Oct 4, 2025Updated 11 months ago
- An unofficial pytorch dataloader for Open X-Embodiment Datasets https://github.com/google-deepmind/open_x_embodiment☆25Jan 9, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ECCV 2024] Official Implementation of CoPT: Unsupervised Domain Adaptive Segmentation using Domain-Agnostic Text Embeddings☆10Feb 24, 2025Updated last year
- Code for LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos☆33Oct 27, 2025Updated 10 months ago
- ☆18Dec 17, 2022Updated 3 years ago
- Perceptual Grouping in Contrastive Vision-Language Models (ICCV'23)☆37Jan 1, 2024Updated 2 years ago
- WACV 2024: "PathLDM: Text conditioned Latent Diffusion Model for Histopathology"☆51Jul 7, 2024Updated 2 years ago
- [ICRA'24] Crossway Diffusion: Improving Diffusion-based Visuomotor Policy via Self-supervised Learning☆73Aug 4, 2024Updated 2 years ago
- ☆31Oct 27, 2022Updated 3 years ago
- Implementation of paper 'Helping Hands: An Object-Aware Ego-Centric Video Recognition Model'☆33Nov 7, 2023Updated 2 years ago
- [WACV 2024] Code for "Limited Data, Unlimited Potential: A Study on ViTs Augmented by Masked Autoencoders"☆25Aug 16, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆18Jan 4, 2024Updated 2 years ago
- [ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation☆179Jun 19, 2025Updated last year
- Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise☆42Oct 7, 2025Updated 10 months ago
- This is a repo of extension of VPN for Recognition of Activities of Daily Living☆16May 17, 2021Updated 5 years ago
- ☆40Jun 2, 2026Updated 3 months ago
- Menagerie of video models trained on various video datasets☆10Oct 13, 2024Updated last year
- [ICLR'25] LLaRA: Supercharging Robot Learning Data for Vision-Language Policy☆228Mar 29, 2025Updated last year
- This code is provided for reproducibility of results in the paper: Multiview Aerial Visual Recognition (MAVREC): Can Multi-view Improve A…☆24Feb 6, 2025Updated last year
- ☆19May 3, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- dist☆10Dec 14, 2018Updated 7 years ago
- ☆30Sep 4, 2025Updated last year
- [ICLR'25] Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?☆14Apr 11, 2025Updated last year
- ☆14Nov 28, 2022Updated 3 years ago
- [ACCV 2020] HDD-Net: Hybrid Detector Descriptor with Mutual Interactive Learning☆17Dec 9, 2020Updated 5 years ago
- [CVPR 2024] Code and models for pi-ViT, a video transformer for understanding activities of daily living☆31Nov 12, 2025Updated 9 months ago
- [ECCV2022] [T-PAMI] StARformer: Transformer with State-Action-Reward Representations.☆97May 21, 2023Updated 3 years ago
- [MICCAI 2024 🔥] HLSS, the first study to explore hierarchical information inherent in histopathology images and their language descripti…☆28Aug 5, 2024Updated 2 years ago
- ☆13Sep 23, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆23Aug 17, 2025Updated last year
- [ACM MM 2024] Frequency Guidance Matters: Skeletal Action Recognition by Frequency-Aware Mixed Transformer☆20Apr 28, 2026Updated 4 months ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 3 months ago
- Linear Conditional Random Field implementation in Pytorch☆11Mar 15, 2021Updated 5 years ago
- ☆20Mar 10, 2025Updated last year
- UniCon: A Simple Approach to Unifying Diffusion-based Conditional Generation (ICLR 2025)☆38Jun 21, 2025Updated last year
- A python package for DICOM to NifTi and NifTi to DICOM-SEG and GSPS conversion☆12Sep 25, 2023Updated 2 years ago