☆29Jul 25, 2026Updated 2 weeks ago
Alternatives and similar repositories for UniVR
Users that are interested in UniVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CAD - Memory Efficient Convolutional Adapter for Segment Anything☆12Oct 4, 2024Updated last year
- ThinkGen: Generalized Thinking for Visual Generation☆61Dec 30, 2025Updated 7 months ago
- A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing☆20Mar 13, 2026Updated 4 months ago
- ☆21Apr 2, 2026Updated 4 months ago
- Official repo for "Let ViT Speak: Generative Language-Image Pre-training"☆133Jun 10, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology☆92Jan 26, 2026Updated 6 months ago
- 小红书笔记 | 评论爬虫、抖音视频 | 评论爬虫、快手视频 | 评论爬虫、B 站视频 | 评论爬虫、微博帖子 | 评论爬虫☆11Mar 28, 2024Updated 2 years ago
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 2 months ago
- [NeurIPS'23] Binary Classification with Confidence Difference☆10May 13, 2024Updated 2 years ago
- TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment☆10Mar 1, 2025Updated last year
- ☆10Jun 17, 2023Updated 3 years ago
- Official Pytorch implementation of 'Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning'? (ICLR2024)☆13Mar 8, 2024Updated 2 years ago
- Röttger et al. (2025): "MSTS: A Multimodal Safety Test Suite for Vision-Language Models"☆20Mar 31, 2025Updated last year
- The repo of the paper: Generalist Vision Foundation Models for Medical Imaging: A Case Study of Segment Anything Model on Zero-Shot Medic…☆11May 26, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- FNIN: A Fourier Neural Operator-based Numerical Integration Network for Surface-form-gradients☆13Jan 22, 2025Updated last year
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆95Jul 13, 2025Updated last year
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- ☆18May 18, 2026Updated 2 months ago
- Pattern Expansion and Consolidation on Evolving Graphs for Continual Traffic Prediction in KDD2023☆13Dec 9, 2025Updated 8 months ago
- The official implementation of "Semi-supervised Segmentation of Histopathology Images with Noise-Aware Topological Consistency".☆14Jul 16, 2024Updated 2 years ago
- code release☆38Jun 22, 2026Updated last month
- D-ORCA: Dialogue-Centric Optimization for Robust Audio-Visual Captioning☆15Updated this week
- [MICCAI 2024] MoRA: LoRA Guided Multi-Modal Disease Diagnosis with Missing Modality☆14Sep 26, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆18Mar 6, 2026Updated 5 months ago
- ICDE'24 "Time-aware Graph Structure Learning for Spatiao-temporal Forecasting"☆15Jan 26, 2026Updated 6 months ago
- Opearting system lab(2023) of BUAA☆14Feb 14, 2026Updated 5 months ago
- 【2024 ECAI】First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending☆14Jun 16, 2025Updated last year
- The Source Code for OmniVideoBench @ICLR 2026☆78Feb 12, 2026Updated 5 months ago
- [ICCV 2025] Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation☆91Sep 29, 2025Updated 10 months ago
- ☆16Jun 23, 2026Updated last month
- [CVPR-26] Official repository of "CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization"☆19Mar 9, 2026Updated 5 months ago
- This is the 2024 OS lab repository.☆11Jun 27, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SpeedVision is an AI-powered tool that detects and calculates vehicle speed from video footage using YOLO-based object detection and fram…☆15Sep 22, 2024Updated last year
- [ICML 2026] The official implementation of paper "Unified Multimodal Autoregressive Modeling with Shared Context—Visual Tokenizer is Key …☆51Jul 13, 2026Updated 3 weeks ago
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- [Roadmap] Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling☆129Jun 9, 2026Updated 2 months ago
- 2022高教社杯数学建模 C题 古代玻璃制品的成分分析与鉴别☆13Nov 26, 2022Updated 3 years ago
- Multimodal Safety Awareness Benchmark for Large Language Models☆16Jun 3, 2025Updated last year
- [ICASSP 2025] Diffusion Features to Bridge Domain Gap for Semantic Segmentation☆18Nov 21, 2024Updated last year