[CVPR 2025] ITA-MDT official implementation
☆67Dec 21, 2025Updated 9 months ago
Alternatives and similar repositories for ita-mdt_code
Users that are interested in ita-mdt_code are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML'25 Spotlight] FlowDrag: 3D-aware Drag-based Image Editing with Mesh-guided Deformation Vector Flow Fields☆47Dec 28, 2025Updated 9 months ago
- (ICCV2025) Occlusion-robust Stylization for Drawing-based 3D Animation☆50Dec 26, 2025Updated 9 months ago
- HEAR: Hearing Enhanced Audio Response for Video-grounded Dialogue, EMNLP 2023 (long, findings) [STARLAB] Audio Enhancement for video-dial…☆57Dec 23, 2023Updated 2 years ago
- Implementation of Uncertainty-Aware Rank-One MIMO Q Network Framework for Accelerated Offline Reinforcement Learning☆32Apr 7, 2026Updated 5 months ago
- ☆33Nov 26, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 2024] FlexiEdit: Frequency-Aware Latent Refinement for Enhanced Non-Rigid Editing☆75Aug 13, 2025Updated last year
- Policy Learning from Large Vision-Language Model Feedback Without Reward Modeling (IROS 2025)☆39Dec 26, 2025Updated 9 months ago
- [IEEE Access 2022] AI for detecting BPPV disorders specified by beatings, torsional movements of the eyes☆37Nov 25, 2022Updated 3 years ago
- [ICLR'25] MDSGen: Fast and Efficient Masked Diffusion Temporal-Aware Transformers for Open-Domain Sound Generation☆39Dec 25, 2025Updated 9 months ago
- ☆39Dec 21, 2024Updated last year
- SCANet: Scene Complexity Aware Network for Weakly-Supervised Video Moment Retrieval (ICCV'2023), [STARLAB] This repositery is a system to…☆57Apr 14, 2025Updated last year
- [ECCV'24] Official code for "BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation"☆43Nov 19, 2024Updated last year
- PyTorch implementation of **Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses**☆31Dec 9, 2025Updated 9 months ago
- Predictive Coding for Decision Transformer (IROS 2024)☆42Jun 19, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR'25] Official code for "Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models"☆35Dec 26, 2025Updated 9 months ago
- Test-time Procrustes Calibration for Diffusion-based Human Image Animation, NeurIPS 2024☆52Aug 23, 2025Updated last year
- Mitigating Adversarial Perturbations for Deep Reinforcement Learning via Vector Quantization (IROS 2024)☆44Jun 19, 2025Updated last year
- This repository is the official implementation of the paper: Physics Informed Distillation for Diffusion Models, accepted by Transactions…☆54Nov 27, 2025Updated 10 months ago
- Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models (ICML 2025)☆55Dec 26, 2025Updated 9 months ago
- Winning SubNetwork (WSN)☆59Jan 17, 2024Updated 2 years ago
- Causal Localization Network for Radar Human Localization with micro-Doppler signature☆62Sep 26, 2024Updated 2 years ago
- [ECCV'22] SQuiDNet: Selective Query-guided Debiasing Network for Video Corpus Moment Retrieval☆74Nov 23, 2022Updated 3 years ago
- [ICML'25] Official code for "ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization"☆18Mar 15, 2026Updated 6 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Winning SubNetwork (WSN), Fourier Subneural Operator (FSO), Video-Incremental Learning (VIL), Sequential Neural Implicit Representation (…☆49Nov 19, 2024Updated last year
- ICML 2024, Official Implementation of "Cross-view Masked Diffusion Transformers for Person Image Synthesis."☆53Aug 10, 2026Updated last month
- 비디오 기반 인공지능 대화시스템☆14Dec 23, 2023Updated 2 years ago
- ☆18Nov 19, 2024Updated last year
- Official code for "SimPSI: A Simple Strategy to Preserve Spectral Information in Time Series Data Augmentation", AAAI 2024.☆37Jan 22, 2025Updated last year
- Text-based Video Retrieval☆15Dec 4, 2024Updated last year
- Retrieval_OOD_for_Multimodal_AI☆11Dec 4, 2024Updated last year
- Multimodal_AI_Video_Dialogue☆16Dec 3, 2024Updated last year
- Fast and Efficient MMD-based Fair PCA via Optimization over Stiefel Manifold (AAAI 2022)☆11Sep 27, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2022 Oral] SoftGroup for Instance Segmentation on 3D Point Clouds☆461Jan 22, 2024Updated 2 years ago
- ☆28Mar 13, 2025Updated last year
- [ACM Multimedia 2024] Shape-Guided Clothing Warping for Virtual Try-On☆33May 14, 2025Updated last year
- This is official repository for Dual Temperature Helps Contrastive Learning without Many Negative Samples (CVPR2022)☆27Dec 1, 2022Updated 3 years ago
- [ICCV 2025] Code Implementation of "Any2AnyTryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing Tasks"☆167Sep 27, 2025Updated last year
- [IEEE Access] ProNeRF: Learning Efficient Projection-Aware Ray Sampling for Fine-Grained Implicit Neural Radiance Fields☆13Apr 26, 2026Updated 5 months ago
- [ICCV2025] OmniVTON: Training-Free Universal Virtual Try-On☆79Oct 25, 2025Updated 11 months ago