☆211Jul 18, 2025Updated last year
Alternatives and similar repositories for large-scale-image-deduplication
Users that are interested in large-scale-image-deduplication are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆38Jul 8, 2025Updated last year
- A repo of resource for the GPU Mode talk on OpenEnv.☆16Jan 14, 2026Updated 8 months ago
- MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer (EMNLP 2025)☆12Apr 18, 2025Updated last year
- Minimal Differentiable Image Reward Functions☆120Mar 30, 2026Updated 5 months ago
- [CVPR 2025] Exploring the Deep Fusion of Large Language Models and Diffusion Transformers for Text-to-Image Synthesis☆140May 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Apr 16, 2025Updated last year
- A modular, first-principles walk-through of Vision-Language-Action models☆33Nov 30, 2025Updated 9 months ago
- ☆59Feb 27, 2025Updated last year
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆5,027Oct 27, 2025Updated 10 months ago
- A missing piece of the Python multitask (both threads and processes) API: An extension that supports stateful worker pools & size-aware i…☆29Mar 8, 2026Updated 6 months ago
- Fork of Flame repo for training of some new stuff in development☆20Aug 27, 2026Updated 3 weeks ago
- Scaling Vision Pre-Training to 4K Resolution☆225Jan 4, 2026Updated 8 months ago
- Image Gaussian Splatting☆27Jul 21, 2025Updated last year
- [NeurIPS 2025] Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations☆202Sep 18, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for "CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning"☆34Mar 26, 2025Updated last year
- Compare Savant and PyTorch performance☆13Feb 9, 2024Updated 2 years ago
- FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens☆17Sep 8, 2025Updated last year
- ☆106Jul 16, 2026Updated 2 months ago
- Model code for inferencing T5☆67Mar 10, 2025Updated last year
- Code for "Scaling Language-Free Visual Representation Learning" paper (Web-SSL).☆216Mar 20, 2026Updated 6 months ago
- A suite of SMPL functionality built in Rust☆28Dec 18, 2025Updated 9 months ago
- ☆32Nov 4, 2024Updated last year
- DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception☆161Dec 6, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!☆2,367Apr 13, 2026Updated 5 months ago
- Framework for processing and filtering datasets☆31Aug 1, 2024Updated 2 years ago
- Seed1.5-VL, a vision-language foundation model designed to advance general-purpose multimodal understanding and reasoning, achieving stat…☆1,589Jun 14, 2025Updated last year
- [ICLR 2025 Spotlight] OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text☆429May 5, 2025Updated last year
- Codebase for FinePDFs☆191Jan 9, 2026Updated 8 months ago
- Making Flux go brrr on GPUs.☆172Jan 5, 2026Updated 8 months ago
- Code for ECCV 2022 paper “Learning with Recoverable Forgetting”☆21Jul 27, 2022Updated 4 years ago
- PyLate efficient inference engine☆91Jan 7, 2026Updated 8 months ago
- ContextBLIP : Doubly Contextual Alignment for Contrastive Image Retrieval from Linguistically Complex Descriptions [ACL 2024]☆11May 17, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- LLaVA-UHD v3: Progressive Visual Compression for Efficient Native-Resolution Encoding in MLLMs☆425Jul 6, 2026Updated 2 months ago
- Visual demo of DSPy's prompt optimization on Gradio☆17Apr 14, 2025Updated last year
- [ICML 2026] 🎨 Occluded 3D Scene Reconstruction from a Single Image.☆97Jun 9, 2026Updated 3 months ago
- Codebase for EMNLP 2025 Findings paper "Text or Pixels? Evaluating Efficiency and Understanding of LLMs with Visual Text Inputs"☆19Nov 14, 2025Updated 10 months ago
- ☆22Mar 25, 2025Updated last year
- ☆16Aug 1, 2024Updated 2 years ago
- 🦄 Serving Platform for Spatial AI and Robotics.☆23Jun 19, 2025Updated last year