[AAAI 2026 Oral] HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
☆29Dec 17, 2025Updated 7 months ago
Alternatives and similar repositories for HiMo-CLIP
Users that are interested in HiMo-CLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Collection of Acceleration Methods for Generative AI☆29Dec 9, 2025Updated 7 months ago
- ☆26Nov 17, 2025Updated 8 months ago
- [ICLR 2026] Any-step Generation via N-th Order Recursive Consistent Velocity Field Estimation☆40Feb 4, 2026Updated 5 months ago
- Official PyTorch Implementation for the "What if...?: Thinking Counterfactual Keywords Helps to Mitigate Hallucination in Large Multi-mod…☆20Sep 26, 2024Updated last year
- [NeurIPS 2025 Spotlight] LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation☆122Jun 22, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- MSCA: Multi-Scale Channel Attention Module☆16Nov 24, 2021Updated 4 years ago
- This is the Github repository for a fault tolerant optimal ZNN controller with state constraints and precribed performance constraints de…☆21Nov 11, 2024Updated last year
- [AAAI 2025] Enhance Vision-Language Alignment with Noise☆26Dec 19, 2024Updated last year
- Official Codebase for "Generative Multimodal Model Features Are Discriminative Vision-Language Classifiers"☆26Jun 7, 2025Updated last year
- ☆11Apr 4, 2021Updated 5 years ago
- [MM2024] LDA-AQU: Adaptive Query-guided Upsampling via Local Deformable Attention☆13Dec 24, 2024Updated last year
- [ICLR 2025] Official repo for paper: "GRACE: Generative Representation Learning via Contrastive Policy Optimization"☆39Feb 3, 2026Updated 5 months ago
- ICCV 2023: The Euclidean Space is Evil: Hyperbolic Attribute Editing for Few-shot Image Generation☆15Sep 29, 2023Updated 2 years ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code for "Harnessing Textual Semantic Priors for Knowledge Transfer and Refinement in CLIP-Driven Continual Learning" (AAAI-2026 poster)☆16Mar 13, 2026Updated 4 months ago
- Working note for WSI analysis☆10Apr 3, 2023Updated 3 years ago
- RayGen: Multi-Modal Dataset Reinforcement for MobileCLIP and MobileCLIP2☆40Mar 12, 2026Updated 4 months ago
- The official implementation for "Learning Expandable and Adaptable Representations for Continual Learning" (NeurIPS2025)☆15Jan 18, 2026Updated 6 months ago
- The official Pytorch implementation of paper Where is My Spot? Few-shot Image Generation via Latent Subspace Optimization, CVPR 2023.☆11Jan 6, 2024Updated 2 years ago
- Log-Polar Space Convolution for Convolutional Neural Networks☆13Dec 12, 2022Updated 3 years ago
- Multi-modal Multi-platform Person Re-Identification: Benchmark and Method☆25Aug 21, 2025Updated 11 months ago
- [NeurIPS 2025] Scaling Language-centric Omnimodal Representation Learning☆47Apr 13, 2026Updated 3 months ago
- ☆15May 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Given an RGBD image and a text prompt, ForceSight produces visual-force goals for a robot, enabling mobile manipulation in unseen environ…☆25Nov 6, 2023Updated 2 years ago
- ☆14Oct 22, 2024Updated last year
- (ICCV 2023) Generative Multiplane Neural Radiance for 3D Aware Image Generation.☆18Sep 28, 2023Updated 2 years ago
- AI agents weaving intelligence, execution, and automation into DeFi☆10Feb 28, 2025Updated last year
- Coursera MOOC on Robotics: Estimation and Learning☆16Feb 17, 2017Updated 9 years ago
- ☆15Apr 5, 2023Updated 3 years ago
- [AAAI 2026 Oral] The official code of "UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning"☆74Dec 8, 2025Updated 7 months ago
- ☆21Aug 27, 2025Updated 10 months ago
- 国家税务总局全国增值税发票查验平台(https://inv-veri.chinatax.gov.cn/) 测试查询☆12Jan 3, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Cross-Self KV Cache Pruning for Efficient Vision-Language Inference☆10Dec 15, 2024Updated last year
- ScreenExplorer: Training a Vision-Language Model for Diverse Exploration in Open GUI World☆26Jun 17, 2025Updated last year
- Implementation of the accepted paper "2D Laser SLAM with Closed Shape Features: Fourier Series Parameterization and Submap Joining"☆16Jul 18, 2021Updated 5 years ago
- Official FlashDecoder Github☆15Apr 4, 2026Updated 3 months ago
- Code for paper "A Large Language Model-Driven Reward Design Framework via Dynamic Feedback for Reinforcement Learning"☆16Jul 15, 2025Updated last year
- modules required for running Roboy at fairs☆12Sep 24, 2018Updated 7 years ago
- MICCAI2022 GOALS Challenge & Paper accepted by TMI2023 (Retinal Layer Segmentation in OCT images with Boundary Regression and Feature Pol…☆18Oct 16, 2023Updated 2 years ago