☆18Mar 25, 2026Updated 4 months ago
Alternatives and similar repositories for auto-hf-papers
Users that are interested in auto-hf-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Own eidetic-memory like Sheldon☆32Apr 8, 2026Updated 3 months ago
- A Simple Way to Eliminate Reward Hacking in GRPO Diffusion Alignment☆21Apr 14, 2026Updated 3 months ago
- Geo-Align: Video Generation Alignment via Metric Geometry Reward☆34May 25, 2026Updated 2 months ago
- 将 B 站视频转化为结构化的 Markdown 阅读笔记 —— 看视频太慢,不如读笔记。☆61Jul 22, 2026Updated 2 weeks ago
- Code of BRIDGE: Building Reinforcement-Learning Depth-to-Image Data Generation Engine for Monocular Depth Estimation☆117Sep 30, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Blender addon for Pi3 3D reconstruction☆15Jul 25, 2025Updated last year
- An easy way for debug python for Slurm HPC users.☆28Mar 23, 2025Updated last year
- Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video☆226May 30, 2026Updated 2 months ago
- A Unified Visual Generator with Interleaved OmniModal Context☆233Mar 5, 2026Updated 5 months ago
- DeepVerse: 4D Autoregressive Video Generation as a World Model☆230Aug 11, 2025Updated 11 months ago
- [ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation☆179Jun 19, 2025Updated last year
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆489Apr 16, 2026Updated 3 months ago
- [ICLR 2025] Where Am I and What Will I See : An Auto-Regressive Model for Spatial Localization and View Prediction☆45Aug 9, 2025Updated 11 months ago
- [CVPR 2025] Tra-MoE: Learning Trajectory Prediction Model from Multiple Domains for Adaptive Policy Conditioning☆56Apr 1, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆133Nov 15, 2025Updated 8 months ago
- [NeurIPS'24] NeuRodin: A Two-stage Framework for High-Fidelity Neural Surface Reconstruction☆128Sep 26, 2024Updated last year
- Open-source implementations on real robots☆35Nov 25, 2024Updated last year
- Code of WinT3R: Window-Based Streaming Rrconstruction With Camera Token Pool☆230Mar 4, 2026Updated 5 months ago
- iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation☆188Dec 1, 2025Updated 8 months ago
- Infinite-Forcing: Towards Infinite-Long Video Generation☆154Nov 13, 2025Updated 8 months ago
- [NeurIPS 2024 D&B] Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning☆92Oct 14, 2024Updated last year
- The repository for a thorough empirical evaluation of pre-trained vision model performance across different downstream policy learning me…☆24Aug 19, 2023Updated 2 years ago
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆606Oct 26, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- PyTorch re-implementation for MeanFlow☆126Jul 17, 2025Updated last year
- Evaluation for 3D reconstruction, includes monocular depth, video depth, relative camera pose & multi-view point map estimation.☆23Aug 26, 2025Updated 11 months ago
- 🎭 Know yourself as a developer. One command → AI analyzes your coding history → beautiful personality portrait + persona skill. Works wi…☆25Apr 8, 2026Updated 3 months ago
- ☆38Mar 25, 2025Updated last year
- ☆14Jan 22, 2025Updated last year
- Code for "BoxDreamer: Dreaming Box Corners for Generalizable Object Pose Estimation", ICCV 2025.☆109Oct 6, 2025Updated 10 months ago
- Depth Any Video with Scalable Synthetic Data (ICLR 2025)☆518Dec 4, 2024Updated last year
- ☆42Feb 17, 2026Updated 5 months ago
- ☆10Oct 27, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Exploring Representation-Aligned Latent Space for Better Generation☆19Mar 17, 2026Updated 4 months ago
- From Automated Idea Factory to Realization☆1,368Jul 18, 2026Updated 2 weeks ago
- ☆15Jun 2, 2025Updated last year
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,102Jul 3, 2026Updated last month
- 🎥 [Awesome] Egocentric / First-Person Video Datasets 📚 Papers, Benchmarks & Resources for Ego Vision☆195Jul 6, 2026Updated 3 weeks ago
- The official github repo for MixEval-X, the first any-to-any, real-world benchmark.☆17Feb 15, 2025Updated last year
- [WACV 2023] XNeRF: Explicit Neural Radiance Field for Multi-Scene 360° Insufficient RGB-D Views☆54Oct 26, 2022Updated 3 years ago