Official repo for "Let ViT Speak: Generative Language-Image Pre-training"
☆135Jun 10, 2026Updated 2 months ago
Alternatives and similar repositories for GenLIP
Users that are interested in GenLIP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ThinkGen: Generalized Thinking for Visual Generation☆61Dec 30, 2025Updated 8 months ago
- ☆26Apr 17, 2024Updated 2 years ago
- [NeurIPS 2025] Official repo of "Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions"☆21Aug 6, 2025Updated last year
- Memory Efficient Matting with Adaptive Token Routing (AAAI 2025)☆80Mar 30, 2026Updated 5 months ago
- ☆49Jun 7, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆31Jul 25, 2026Updated last month
- Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders [Technical Report]☆209Mar 30, 2026Updated 5 months ago
- ☆141Dec 26, 2025Updated 8 months ago
- Official implementation of Tuna-2: Pixel Embeddings Beat Vision Encoders for Unified Understanding and Generation☆752Jul 22, 2026Updated last month
- TIPSv2 (CVPR'26) and TIPS (ICLR'25)☆613Jun 1, 2026Updated 2 months ago
- OpenVision (ICCV 2025), OpenVision 2 (CVPR 2026), and OpenVision 3☆494Aug 22, 2026Updated last week
- The download methods of Vision-language Continual Pretraining Dataset P9D.☆12Jan 3, 2025Updated last year
- [NeurIPS2025] The official implementation of MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO☆140Oct 15, 2025Updated 10 months ago
- ☆19Aug 15, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 🍨 Gelato — From Data Curation to Reinforcement Learning: Building a Strong Grounding Model for Computer-Use Agents☆46Dec 22, 2025Updated 8 months ago
- [ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generation☆504Apr 15, 2026Updated 4 months ago
- Revisiting Multi-Task Visual Representation Learning☆22Jan 21, 2026Updated 7 months ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation