Real-time YOLOv5 + Intel RealSense D435 pipeline for depth-aware object detection and 3D coordinate extraction, enabling precise robotic arm grasping.
☆17Nov 29, 2025Updated 8 months ago
Alternatives and similar repositories for Real-Time-Object-Detection-with-Depth
Users that are interested in Real-Time-Object-Detection-with-Depth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Intelligent video-learning platform that analyzes viewing behavior to generate personalized questions, track errors, and enable social, f…☆20Dec 15, 2025Updated 7 months ago
- VisionZip-enhanced LLaDA-V for DLM inference, compressing visual tokens for faster, plug-and-play vision-aware reasoning with minimal qua…☆15Nov 28, 2025Updated 8 months ago
- [ECCV2026] FreeSwim: Revisiting Sliding-Window Attention Mechanisms for Training-Free Ultra-High-Resolution Video Generation☆59Dec 2, 2025Updated 7 months ago
- [ECCV2026] ViBe: Ultra-High-Resolution Video Synthesis Born from Pure Images☆31May 21, 2026Updated 2 months ago
- A Lightweight, Configuration-Driven, Flexible Fine-Tuning Framework for 🤗 Diffusers☆17Apr 15, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICML 2026] Official repository for the paper "Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention"☆42May 24, 2026Updated 2 months ago
- Minute-long video generation at 24FPS.☆69Mar 28, 2026Updated 4 months ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆82May 31, 2025Updated last year
- A comprehensive benchmark specifically designed to evaluate the interactive response capabilities of world models in 4D settings.☆106Mar 24, 2026Updated 4 months ago
- A Minimalist, Batteries-included Repository for Advancing World Model Science.☆694Jun 15, 2026Updated last month
- [ICML 2026] d3LLM: Ultra-Fast Diffusion LLM 🚀☆148May 1, 2026Updated 2 months ago
- 4-steps distilled version of Wan2.2-TI2V-5B☆163Mar 15, 2026Updated 4 months ago
- [ICML 2026] Pytorch implementation of Self-Refining Video Sampling☆185May 1, 2026Updated 2 months ago
- [NeurIPS 2025] Official PyTorch implementation of paper "CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up".☆219Sep 27, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Interactive World Model papers organized by core research challenges.☆273Jul 16, 2026Updated last week
- VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model☆2,262Mar 19, 2026Updated 4 months ago
- ☆251Nov 19, 2025Updated 8 months ago
- [ICLR 2026] When it comes to optimizers, it's always better to be safe than sorry☆418Sep 26, 2025Updated 10 months ago
- SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse–Linear Attention☆324Feb 24, 2026Updated 5 months ago
- [ICLR 2026] Official implementation of JavisDiT and JavisDiT++ series.☆376Mar 29, 2026Updated 4 months ago
- Dynamic 3D Foundation Model using Causal Transformer. [ICLR 2026]☆392May 8, 2026Updated 2 months ago
- [ICCV 2025] Official implementation of the paper: REPA-E: Unlocking VAE for End-to-End Tuning of Latent Diffusion Transformers☆511Dec 6, 2025Updated 7 months ago
- rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale☆774Jun 25, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆774Apr 16, 2026Updated 3 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆882Updated this week
- [ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.☆1,019Feb 25, 2026Updated 5 months ago
- Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"☆1,064May 30, 2026Updated last month
- MOVA: Towards Scalable and Synchronized Video–Audio Generation☆1,087Jun 18, 2026Updated last month
- Master the Toolkit of AI and Machine Learning. Mathematics for Machine Learning and Data Science is a beginner-friendly Specialization wh…☆866Jul 10, 2023Updated 3 years ago
- [ICLR 26 Oral] Stable Video Infinity: Infinite-Length Video Generation with Error Recycling☆2,538Jun 3, 2026Updated last month
- PyTorch implementation of JiT https://arxiv.org/abs/2511.13720☆2,473Dec 8, 2025Updated 7 months ago
- Official codebase for "Self Forcing: Bridging Training and Inference in Autoregressive Video Diffusion" (NeurIPS 2025 Spotlight)☆3,465Sep 12, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- TurboDiffusion: 100–200× Acceleration for Video Diffusion Models☆3,588Jul 16, 2026Updated last week
- [ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-t…☆3,518Jan 17, 2026Updated 6 months ago
- 分享一些好用的 Dify DSL 工作流程,自用、学习两相宜。 Sharing some Dify workflows.☆10,720Mar 25, 2026Updated 4 months ago
- Advancing Open-source World Models☆4,298Jul 9, 2026Updated 2 weeks ago
- 🚀 Efficient implementations for emerging model architectures☆5,463Updated this week
- Align Anything: Training All-modality Model with Feedback☆4,664Nov 27, 2025Updated 8 months ago
- SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer☆8,591Updated this week