[ICLR 2024] Seer: Language Instructed Video Prediction with Latent Diffusion Models
☆35May 23, 2024Updated 2 years ago
Alternatives and similar repositories for SeerVideoLDM
Users that are interested in SeerVideoLDM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Jun 13, 2024Updated 2 years ago
- ReMoDetect: Reward Models Recognize Aligned LLM's Generations (NeurIPS 2024)☆17Nov 15, 2024Updated last year
- [NeurIPS'21] RoMA: Robust Model Adaptation for Offline Model-based Optimization☆15Oct 28, 2021Updated 4 years ago
- From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data☆18Jun 2, 2026Updated 4 months ago
- Codebase for HiP☆90Dec 15, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Improving Motion in Image-to-Video Models via Adaptive Low-Pass Guidance (CVPR 2026 Highlight)☆60Feb 23, 2026Updated 7 months ago
- [arXiv 2024] I4VGen: Image as Free Stepping Stone for Text-to-Video Generation☆24Oct 6, 2024Updated 2 years ago
- ☆17Jan 21, 2026Updated 8 months ago
- [IROS 2025] Codebase for DG16M: A Large-Scale Dataset for Dual-Arm Grasping with Force-Optimized Grasps☆17Feb 13, 2026Updated 7 months ago
- The official PyTorch implementation of "The 18th European Conference on Computer Vision" (ECCV 2024) paper Length-Aware Motion Synthesis …☆19Dec 15, 2024Updated last year
- Bidirectional Mapping between Action Physical-Semantic Space☆33Sep 7, 2025Updated last year
- [CVPR 2026] Implementation of HAMMER: Harnessing MLLMs via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding☆26Aug 2, 2026Updated 2 months ago
- Source code for "Improving the Diffusability of Autoencoders" [ICML 2025]☆22Jan 6, 2026Updated 9 months ago
- ☆16Apr 20, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Codebase for PRISE: Learning Temporal Action Abstractions as a Sequence Compression Problem☆24Jul 11, 2024Updated 2 years ago
- ☆31Apr 7, 2024Updated 2 years ago
- ☆12Jul 30, 2025Updated last year
- ChangeIt dataset with more than 2600 hours of video with state-changing actions published at CVPR 2022☆11Mar 23, 2022Updated 4 years ago
- ☆83May 23, 2025Updated last year
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".☆13Jan 25, 2025Updated last year
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆12Jul 26, 2024Updated 2 years ago
- UniVid: The Open-Source Unified Video Model☆32Oct 13, 2025Updated 11 months ago
- Placeholder☆10Jul 17, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Tutorials for using pybullet_planning☆32Jun 8, 2020Updated 6 years ago
- Official repository of Learning to Act from Actionless Videos through Dense Correspondences.☆263Apr 25, 2024Updated 2 years ago
- ☆19Aug 15, 2026Updated last month
- ☆22Apr 8, 2024Updated 2 years ago
- [ICRA 2026] Official codebase for DAGDiff: Guiding Dual-Arm Grasp Diffusion to Stable and Collision-Free Grasps☆28Feb 1, 2026Updated 8 months ago
- ☆14Oct 17, 2024Updated last year
- [ICLR'25] MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequences☆324Aug 10, 2024Updated 2 years ago
- 为墨水屏设备打造的番茄钟,开箱即用,具备记录并可视化展示历史学习时长、计划倒计时、设置时间长度、本地保存/读取等功能。☆15Apr 18, 2021Updated 5 years ago
- Multi-agent harness + complete run record of the 1st-place entry at Ralphthon@ICML2026 — three AI agents wrote a workshop paper in 3 hour…☆23Jul 14, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Contrastive self-supervised learning using Rényi divergence☆14Oct 21, 2022Updated 3 years ago
- Imitation Learning from Observation Through Generative Modelling☆27Feb 12, 2025Updated last year
- Code for the paper "What Makes Better Augmentation Strategies? Augment Difficult but Not too Different" (ICLR 22)☆12Aug 28, 2023Updated 3 years ago
- [CVPR'23] Video Probabilistic Diffusion Models in Projected Latent Space☆323May 14, 2024Updated 2 years ago
- Code for subgoal synthesis via image editing☆161Oct 23, 2023Updated 2 years ago
- [ICLR'24] Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition☆55May 14, 2024Updated 2 years ago
- Dreamitate: Real-World Visuomotor Policy Learning via Video Generation (CoRL 2024)☆59Jun 7, 2025Updated last year