☆17Apr 20, 2025Updated last year
Alternatives and similar repositories for PerturboLLaVA
Users that are interested in PerturboLLaVA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Jun 13, 2026Updated last month
- TVRBench: Target Viewpoint Reproduction Benchmark for Active Spatial Intelligence☆25Jun 2, 2026Updated last month
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 4 months ago
- [ICLR 2024] Official PyTorch/Diffusers implementation of "Object-aware Inversion and Reassembly for Image Editing"☆87Aug 23, 2024Updated last year
- [ICML2026] ACTIVE-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO☆83Apr 30, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICCV'25] Unified Open-World Segmentation with Multi-Modal Prompts☆16Jun 16, 2026Updated last month
- ☆47May 6, 2026Updated 2 months ago
- [ACL'26] EvoToken-DLM (Beyond Hard Masks: Progressive Token Evolution for Diffusion Language)☆48Apr 7, 2026Updated 3 months ago
- MUltiple SUV Thresholding (MUST)-segmenter is a semi-automated PET image segmentation tool that enables delineation of multiple lesions a…☆12Mar 18, 2026Updated 4 months ago
- ☆22Jun 30, 2023Updated 3 years ago
- Training-Free OOD Medical Tumor Segmentation via Anatomical Reasoning and Statistical Rejection☆24Jun 29, 2026Updated 3 weeks ago
- ☆39Mar 5, 2026Updated 4 months ago
- ☆36Oct 21, 2022Updated 3 years ago
- [CVPR2026] Exploring Spatial Intelligence from a Generative Perspective☆30Jun 3, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML 2024] Floating Anchor Diffusion Model for Multi-motif Scaffolding☆34Aug 23, 2024Updated last year
- ☆35Apr 10, 2026Updated 3 months ago
- vscode extension for showing images in tile view☆11Mar 6, 2023Updated 3 years ago
- One-shot and Few-shot 3D Editing without Per-Scene Optimization☆175Aug 21, 2025Updated 11 months ago
- [AAAI, 2026] Official implementation of "CoCoLIT: ControlNet-Conditioned Latent Image Translation for MRI to Amyloid PET Synthesis".☆18Mar 23, 2026Updated 4 months ago
- Official implementation for "K-Forcing: Joint Next-K-Token Decoding via Push-Forward Language Modeling"☆17Jun 14, 2026Updated last month
- Are Binary Annotations Sufficient? Video Moment Retrieval via Hierarchical Uncertainty-based Active Learning☆15Dec 12, 2023Updated 2 years ago
- SurfaceSplat: Connecting Surface Reconstruction and Gaussian Splatting☆60Jul 21, 2025Updated last year
- [ICCV2023] 🧊FrozenRecon: Pose-free 3D Scene Reconstruction with Frozen Depth Models☆131Aug 23, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- PyTorch code for the CVPR'23 paper: "ConStruct-VL: Data-Free Continual Structured VL Concepts Learning"☆13Feb 5, 2024Updated 2 years ago
- ☆51Oct 6, 2024Updated last year
- DiverGen (CVPR 2024) & BSGAL (ICML 2024)☆53Jul 6, 2025Updated last year
- Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning☆45Mar 2, 2026Updated 4 months ago
- [NeurIPS'24] Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation (Diffews)☆51Apr 14, 2025Updated last year
- [AAAI'26] PET2Rep: Towards Vision-Language Model-Drived Automated Radiology Report Generation for Positron Emission Tomography☆24Dec 26, 2025Updated 7 months ago
- [NeurIPS 2025] Official Repo of Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration☆126Dec 3, 2025Updated 7 months ago
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)☆15May 21, 2025Updated last year
- [ICLR 2025] Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding☆44Mar 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [WACV 2025] Official Pytorch code for "Background-aware Moment Detection for Video Moment Retrieval"☆16Feb 24, 2025Updated last year
- ☆54Aug 3, 2023Updated 2 years ago
- The code for the NeurIPS19 paper and blog on "Uniform convergence may be unable to explain generalization in deep learning".☆10Oct 26, 2019Updated 6 years ago
- ☆25Jan 10, 2024Updated 2 years ago
- Official implementation of High-Fidelity Zero-Shot Texture Anomaly Localization Using Feature Correspondence Analysis.☆12Dec 18, 2023Updated 2 years ago
- [ICLR'25] MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequences☆323Aug 10, 2024Updated last year
- ☆13Aug 14, 2022Updated 3 years ago