Official PyTorch codebase for the Modeling Caption Diversity in ContrastiveVision-Language Pretraining paper.
☆19Mar 28, 2025Updated last year
Alternatives and similar repositories for Llip
Users that are interested in Llip are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2024] Z-GMOT: Zero-shot Generic Multiple Object Tracking☆12May 19, 2026Updated 2 months ago
- A curated list of Survey Papers on Deep Learning.☆13Sep 5, 2023Updated 2 years ago
- ☆15Jun 9, 2025Updated last year
- [CVPR2025] Official implementation of the paper "Multi-Layer Visual Feature Fusion in Multimodal LLMs: Methods, Analysis, and Best Practi…☆48Oct 29, 2025Updated 9 months ago
- ☆11Sep 30, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆11Mar 13, 2024Updated 2 years ago
- Code of "Robustifying Token Attention for Vision Transformers"☆20Dec 31, 2023Updated 2 years ago
- Retrieval_OOD_for_Multimodal_AI☆11Dec 4, 2024Updated last year
- ☆18Nov 19, 2024Updated last year
- Code accompanying paper "SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation"☆29May 8, 2026Updated 3 months ago
- [EMNLP 2024 Main] Official implementation of the paper "To Preserve or To Compress: An In-Depth Study of Connector Selection in Multimoda…☆16Dec 13, 2024Updated last year
- Text-based Video Retrieval☆15Dec 4, 2024Updated last year
- ☆14Jan 5, 2022Updated 4 years ago
- ☆11May 1, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- tlspyo - secure transfer of python objects over network☆18Jan 23, 2024Updated 2 years ago
- ☆29Apr 8, 2025Updated last year
- Official implementation of "In-style: Bridging Text and Uncurated Videos with Style Transfer for Cross-modal Retrieval." ICCV 2023☆11Oct 5, 2023Updated 2 years ago
- ☆28Mar 13, 2025Updated last year
- In this project, facial recognition algorithm is implemented with python using PCA and SVD dimensionality reduction tools.☆11Sep 2, 2019Updated 6 years ago
- Hướng dẫn tạo một hệ thống Log Remote dùng chung cho nhiều dự án/server☆15Feb 26, 2020Updated 6 years ago
- My PhD manuscript LaTeX code and the slides for the defense☆11Feb 2, 2022Updated 4 years ago
- Official repository of the paper "JIST: Joint Image and Sequence Training for Sequential Visual Place Recognition"☆24Dec 15, 2023Updated 2 years ago
- This repository including most of cnn visualizations techniques using pytorch☆14Apr 14, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- All the codes of ACMICPC for two categories: Competition(Codeforces, Bestcoder, Regional, etc..) and Algorithms☆13Sep 23, 2017Updated 8 years ago
- ☆28Mar 13, 2025Updated last year
- This repository contains code for deploying a Gradio application using the SAM2 model for video processing. The application allows users …☆47Sep 24, 2024Updated last year
- Awesome-Text2Motion-Generation☆18Oct 26, 2023Updated 2 years ago
- Official PyTorch implementation of our CVPR 2025 paper: "SwiftEdit: Lightning Fast Text-guided Image Editing via One-step Diffusion"☆53Jan 7, 2026Updated 7 months ago
- X-MIC: Cross-Modal Instance Conditioning for Egocentric Action Generalization, CVPR 2024☆11Nov 7, 2024Updated last year
- [ECCV 2024] Official Release of SILC: Improving vision language pretraining with self-distillation☆48Oct 3, 2024Updated last year
- Official Release of NeurIPS 2024 paper "Slot State Space Models"☆11Mar 22, 2025Updated last year
- ☆30Jun 1, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ACM Multimedia 2023 (Oral) - RTQ: Rethinking Video-language Understanding Based on Image-text Model☆15Apr 7, 2026Updated 4 months ago
- GPT-style network for phonemization with durations of text☆68Mar 21, 2024Updated 2 years ago
- use Blender software to visualize mesh sequences☆24Sep 2, 2019Updated 6 years ago
- Implementation of Recurrent Hidden Semi-Markov Model http://www.cc.gatech.edu/~lsong/papers/DaiDaiZhaLietal17.pdf☆13Mar 31, 2019Updated 7 years ago
- ☆12Oct 10, 2023Updated 2 years ago
- code for "Delving into Probabilistic Uncertainty for Unsupervised Domain Adaptive Person Re-Identification" in AAAI2022☆18Apr 8, 2022Updated 4 years ago
- [ICASSP'23] Online speaker clustering☆19Feb 22, 2026Updated 5 months ago