☆73Jun 28, 2023Updated 3 years ago
Alternatives and similar repositories for zh-clip
Users that are interested in zh-clip are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Chinese CLIP models with SOTA performance.☆62Aug 28, 2023Updated 3 years ago
- Chinese version of CLIP which achieves Chinese cross-modal retrieval and representation generation.☆6,006Mar 31, 2026Updated 5 months ago
- Official codes for ConMIM (ICLR 2023)☆58Feb 8, 2023Updated 3 years ago
- ALIGN trained on COYO-dataset☆29Apr 30, 2024Updated 2 years ago
- MeloTTS demo on Axera☆14Jul 1, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2023] Learning Fine-Grained Features for Pixel-wise Video Correspondences☆18Mar 3, 2024Updated 2 years ago
- SDXL API provides a seamless interface for image generation and retrieval using Stable Diffusion XL integrated with Cloudflare AI Workers…☆14Feb 29, 2024Updated 2 years ago
- Repository of paper: Position-Enhanced Visual Instruction Tuning for Multimodal Large Language Models☆37Sep 19, 2023Updated 2 years ago
- CoTj (Chain-of-Trajectories) upgrades diffusion models from fixed System-1 denoising schedules to System-2 style, condition-adaptive traj…☆23Mar 24, 2026Updated 5 months ago
- [CVPR 2024] CapsFusion: Rethinking Image-Text Data at Scale☆216Feb 27, 2024Updated 2 years ago
- ☆91Jul 4, 2024Updated 2 years ago
- ☆170Nov 9, 2023Updated 2 years ago
- ☆21Feb 29, 2024Updated 2 years ago
- ☆13Sep 6, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- TaiSu(太素)--a large-scale Chinese multimodal dataset(亿级大规模中文视觉语言预训练数据集)☆192Nov 17, 2023Updated 2 years ago
- ManifoldAlignmentStyleTransfer☆46Feb 24, 2022Updated 4 years ago
- 基于wav2lip进行虚拟数字人训练,唇形驱动,包括数据处理流程等,模型包括96x96,192x192,192x288,288x288。☆22May 7, 2024Updated 2 years ago
- ☆21Apr 10, 2018Updated 8 years ago
- 探索智能零售领域的图像识别方案,从而让机器更精准地识别商品,通过更快捷地购物带来全新的用户体验。☆12Jun 15, 2021Updated 5 years ago
- ☆16Jul 29, 2025Updated last year
- ☆28Oct 2, 2024Updated last year
- ☆20Nov 21, 2019Updated 6 years ago
- [ICLR 2024 Spotlight] DreamLLM: Synergistic Multimodal Comprehension and Creation☆462Dec 2, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 使用OpenCV部署L2CS-Net人脸朝向估计,包含C++和Python两个版本的程序,只依赖opencv库就可以运行☆21Aug 12, 2023Updated 3 years ago
- VLE: Vision-Language Encoder (VLE: 视觉-语言多模态预训练模型)☆196Mar 13, 2023Updated 3 years ago
- Train InternViT-6B in MMSegmentation and MMDetection with DeepSpeed☆108Oct 25, 2024Updated last year
- ☆65Feb 5, 2024Updated 2 years ago
- Bridging Vision and Language Model☆287Mar 27, 2023Updated 3 years ago
- 模型 llava-Qwen2-7B-Instruct-Chinese-CLIP 增强中文文字识别能力和表情包内涵识别能力,接近gpt4o、claude-3.5-sonnet的识别水平!☆28Jul 23, 2024Updated 2 years ago
- Unofficial implementation for Sigmoid Loss for Language Image Pre-Training☆11Sep 26, 2023Updated 2 years ago
- Qwen1.5-SFT(阿里, Ali), Qwen_Qwen1.5-2B-Chat/Qwen_Qwen1.5-7B-Chat微调(transformers)/LORA(peft)/推理☆73May 17, 2024Updated 2 years ago
- [CVPR 2023] Official implementation of "SAP-DETR: Bridging the Gap between Salient Points and Queries-Based Transformer Detector for Fast…☆30May 28, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Easily compute clip embeddings and build a clip retrieval system with them☆2,795Mar 28, 2026Updated 5 months ago
- RepGhost: A Hardware-Efficient Ghost Module via Re-parameterization☆182Aug 15, 2023Updated 3 years ago
- ☆13May 11, 2021Updated 5 years ago
- ☆134Dec 22, 2023Updated 2 years ago
- 首届“全国人工智能大赛”(行人重识别 Person ReID 赛项)☆29Jan 3, 2020Updated 6 years ago
- Tight Mutual Information Estimation With Contrastive Fenchel-Legendre Optimization☆11Nov 29, 2022Updated 3 years ago
- This repository contains the CUDA implementation of the paper "Work-efficient Parallel Non-Maximum Suppression Kernels".☆15Aug 21, 2020Updated 6 years ago