[AAAI 2026 Oral] HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
☆29Dec 17, 2025Updated 9 months ago
Alternatives and similar repositories for HiMo-CLIP
Users that are interested in HiMo-CLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Optimizing for the Shortest Path in Denoising Diffusion Model (CVPR2025)☆20Dec 17, 2025Updated 9 months ago
- Collection of Acceleration Methods for Generative AI☆29Dec 9, 2025Updated 9 months ago
- ☆32Nov 17, 2025Updated 10 months ago
- [NeurIPS 2025 Spotlight] LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation☆122Jun 22, 2026Updated 2 months ago
- This is the Github repository for a fault tolerant optimal ZNN controller with state constraints and precribed performance constraints de…☆20Nov 11, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆46Jan 14, 2026Updated 8 months ago
- [AAAI 2025] Enhance Vision-Language Alignment with Noise☆27Dec 19, 2024Updated last year
- ☆45Jan 12, 2026Updated 8 months ago
- [MM2024] LDA-AQU: Adaptive Query-guided Upsampling via Local Deformable Attention☆13Dec 24, 2024Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 4 months ago
- ☆23Apr 27, 2026Updated 4 months ago
- The official code repo of 1.x-Distill, is a stagewise distillation framework for diversity, high-quality and efficient few-step generatio…☆21Apr 10, 2026Updated 5 months ago
- A variant of Varibad that is robust to difficult tasks☆11Aug 30, 2023Updated 3 years ago
- CoTj (Chain-of-Trajectories) upgrades diffusion models from fixed System-1 denoising schedules to System-2 style, condition-adaptive traj…☆23Mar 24, 2026Updated 5 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- RayGen: Multi-Modal Dataset Reinforcement for MobileCLIP and MobileCLIP2☆40Sep 11, 2026Updated last week
- ICCV 2023: The Euclidean Space is Evil: Hyperbolic Attribute Editing for Few-shot Image Generation☆15Sep 29, 2023Updated 2 years ago
- The official implementation for "Learning Expandable and Adaptable Representations for Continual Learning" (NeurIPS2025)☆15Jan 18, 2026Updated 8 months ago
- The official Pytorch implementation of paper Where is My Spot? Few-shot Image Generation via Latent Subspace Optimization, CVPR 2023.☆11Jan 6, 2024Updated 2 years ago
- Multi-modal Multi-platform Person Re-Identification: Benchmark and Method☆25Aug 21, 2025Updated last year
- yolov5第四版☆15Oct 13, 2021Updated 4 years ago
- Given an RGBD image and a text prompt, ForceSight produces visual-force goals for a robot, enabling mobile manipulation in unseen environ…☆25Nov 6, 2023Updated 2 years ago
- Official repository for the paper "PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset"☆35Jul 10, 2026Updated 2 months ago
- code for FineLIP☆43Nov 25, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- teacache for lumina-image-2 in comfyui☆20Feb 2, 2026Updated 7 months ago
- Official Codebase for "Generative Multimodal Model Features Are Discriminative Vision-Language Classifiers"☆29Jun 7, 2025Updated last year
- ☆15Apr 5, 2023Updated 3 years ago
- Code release for Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning | IROS 2024☆55Dec 30, 2024Updated last year
- [AAAI 2026 Oral] The official code of "UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning"☆76Dec 8, 2025Updated 9 months ago
- AI agents weaving intelligence, execution, and automation into DeFi☆10Feb 28, 2025Updated last year
- [ICLR 2026] Official repo for S3OD: Towards Generalizable Salient Object Detection with Synthetic Data☆50Jun 3, 2026Updated 3 months ago
- LW-CTrans☆18Jun 4, 2024Updated 2 years ago
- Few-Shot Image Generation by Conditional Relaxing Diffusion Inversion☆16Mar 14, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 📖Curated list about reasoning abilitiy of MLLM, including OpenAI o1, OpenAI o3-mini, and Slow-Thinking.☆14Feb 7, 2025Updated last year
- 国家税务总局全国增值税发票查验平台(https://inv-veri.chinatax.gov.cn/) 测试查询☆12Jan 3, 2023Updated 3 years ago
- ScreenExplorer: Training a Vision-Language Model for Diverse Exploration in Open GUI World☆26Jun 17, 2025Updated last year
- Code for paper "A Large Language Model-Driven Reward Design Framework via Dynamic Feedback for Reinforcement Learning"☆16Jul 15, 2025Updated last year
- modules required for running Roboy at fairs☆12Sep 24, 2018Updated 7 years ago
- MICCAI2022 GOALS Challenge & Paper accepted by TMI2023 (Retinal Layer Segmentation in OCT images with Boundary Regression and Feature Pol…☆18Oct 16, 2023Updated 2 years ago
- Official FlashDecoder Github☆19Apr 4, 2026Updated 5 months ago