Vector-Quantized Vision Foundation Models for Object-Centric Learning, ACM MM 2025.
☆16May 30, 2026Updated last month
Alternatives and similar repositories for VQ-VFM-OCL
Users that are interested in VQ-VFM-OCL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SAVi -- Unofficial But 3x Faster Training @ Better Performance. Implementation of ICLR 2022 Paper "Conditional Object-Centric Learning fr…☆17Oct 29, 2023Updated 2 years ago
- Official implementation of: "PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning" by…☆22Apr 1, 2026Updated 3 months ago
- (CVPR 2025) A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning☆25Mar 11, 2025Updated last year
- [NeurIPS 2023] Self-supervised Object-Centric Learning for Videos☆32Nov 28, 2024Updated last year
- Official implementation of the CVPR'24 paper [Adaptive Slot Attention: Object Discovery with Dynamic Slot Number]☆76Jan 25, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of: "Object-Centric Video Prediction via Decoupling of Object Dynamics and Interactions" by Villar-Corrales et al…☆25Oct 16, 2023Updated 2 years ago
- ☆14May 31, 2023Updated 3 years ago
- [CVPR 2024 Highlight] SPOT: Self-Training with Patch-Order Permutation for Object-Centric Learning with Autoregressive Transformers☆77Jun 11, 2024Updated 2 years ago
- Official PyTorch Implementation of Masked Temporal Interpolation Diffusion for Procedure Planning in Instructional Videos☆11Jul 10, 2026Updated last week
- Repository for our paper "Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities"☆42Feb 12, 2025Updated last year
- ☆33Jul 23, 2025Updated 11 months ago
- Contact point & surface normal estimation from force-torque measurements using adaptive control.☆16Sep 7, 2014Updated 11 years ago
- Code for paper "PoseEmbroider:Towards a 3D, Visual, Semantic-aware Human Pose Representation" (ECCV 2024)☆18Nov 18, 2024Updated last year
- Official PyTorch implementation of: "Cannot See the Forest for the Trees: Aggregating Multiple Viewpoints to Better Classify Objects in V…☆14Aug 29, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆21Mar 12, 2025Updated last year
- [ECCV 2024] Official Implementation of "Appearance-Based Refinement for Object-Centric Motion Segmentation" Junyu Xie, Weidi Xie, Andrew …☆13Oct 23, 2024Updated last year
- ☆23Jun 17, 2025Updated last year
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 4 months ago
- Spectral embedding using Laplacian Eigenmaps☆15Oct 15, 2018Updated 7 years ago
- [BMVC 2021]OMAD: Object Model with Articulated Deformations for Pose Estimation and Retrieval☆12Dec 17, 2021Updated 4 years ago
- Code for "Unsupervised Space-Time Network for Temporally-Consistent Segmentation of Multiple Motions." (CVPR 2023)☆11Jun 15, 2023Updated 3 years ago
- ☆13Oct 22, 2020Updated 5 years ago
- This is anonymous repository for submitting our work to a conference☆14Dec 17, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS 2022] Unsupervised Multi-View Object Segmentation Using Radiance Field Propagation☆14Nov 9, 2022Updated 3 years ago
- ☆39Mar 31, 2023Updated 3 years ago
- ☆19Jul 12, 2021Updated 5 years ago
- Indexity is a web-based tool designed for medical video annotation in surgical data science projects.☆11Jun 27, 2023Updated 3 years ago
- This is the official PyTorch implementation of the CVPR 2023 paper: "GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot A…☆10Mar 17, 2024Updated 2 years ago
- MR.Q is a general-purpose model-free reinforcement learning algorithm.☆153Apr 7, 2026Updated 3 months ago
- ☆15Jan 11, 2022Updated 4 years ago
- Official implementation of the ICML 2025 paper "SOLD: Slot Object-Centric Latent Dynamics Models for Relational Manipulation Learning fro…☆21Sep 30, 2025Updated 9 months ago
- Dur360BEV: (ICRA 2025) A Real-world 360-degree Single Camera Dataset and Benchmark for Bird-Eye View Mapping in Autonomous Driving☆23Feb 2, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICRA 2026] History-Aware Visuomotor Policy Learning via Point Tracking☆27Jan 10, 2026Updated 6 months ago
- CrossMatch: Enhance Semi-Supervised Medical Image Segmentation with Perturbation Strategies and Knowledge Distillation☆52Dec 10, 2025Updated 7 months ago
- CVPR2022:Learning from Untrimmed Videos: Self-Supervised Video Representation Learning with Hierarchical Consistency☆18Aug 10, 2022Updated 3 years ago
- [ICRA 2025] CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation☆35Jan 14, 2025Updated last year
- (3DV Poster) Keypoint Cascade Voting for Point Cloud Based 6DoF Pose Estimation☆20Feb 14, 2024Updated 2 years ago
- Official Pytorch codebase for Open-Vocabulary Instance Segmentation without Manual Mask Annotations [CVPR 2023]☆52Oct 26, 2025Updated 8 months ago
- ☆19Jun 15, 2026Updated last month