☆31Jul 25, 2026Updated last month
Alternatives and similar repositories for UniVR
Users that are interested in UniVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing☆21Mar 13, 2026Updated 5 months ago
- ☆21Apr 2, 2026Updated 4 months ago
- Official repo for "Let ViT Speak: Generative Language-Image Pre-training"☆135Jun 10, 2026Updated 2 months ago
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology☆93Jan 26, 2026Updated 7 months ago
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS'23] Binary Classification with Confidence Difference☆10May 13, 2024Updated 2 years ago
- Useful tools for pulling data from CalTrans-PeMS.☆11Jan 20, 2023Updated 3 years ago
- Official Pytorch implementation of 'Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning'? (ICLR2024)☆13Mar 8, 2024Updated 2 years ago
- Röttger et al. (2025): "MSTS: A Multimodal Safety Test Suite for Vision-Language Models"☆20Mar 31, 2025Updated last year
- Reliable Source Approximation: Source-Free Domain Adaptation for Vestibular Schwannoma MRI Segmentation☆11Dec 28, 2024Updated last year
- The repo of the paper: Generalist Vision Foundation Models for Medical Imaging: A Case Study of Segment Anything Model on Zero-Shot Medic…☆11May 26, 2023Updated 3 years ago
- SODA: Story Oriented Dense Video Captioning Evaluation Framework☆14May 3, 2024Updated 2 years ago
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆98Jul 13, 2025Updated last year
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆18May 18, 2026Updated 3 months ago
- [CVPR 2023] LOGO: A Long-Form Video Dataset for Group Action Quality Assessment☆48Apr 9, 2024Updated 2 years ago
- The official implementation of "Semi-supervised Segmentation of Histopathology Images with Noise-Aware Topological Consistency".☆14Jul 16, 2024Updated 2 years ago
- code release☆43Jun 22, 2026Updated 2 months ago
- [MICCAI 2024] MoRA: LoRA Guided Multi-Modal Disease Diagnosis with Missing Modality☆14Sep 26, 2025Updated 11 months ago
- ☆18Mar 6, 2026Updated 5 months ago
- ICDE'24 "Time-aware Graph Structure Learning for Spatiao-temporal Forecasting"☆16Jan 26, 2026Updated 7 months ago
- Opearting system lab(2023) of BUAA☆14Feb 14, 2026Updated 6 months ago
- [ICCV 2023] Label-Efficient Online Continual Object Detection in Streaming Video☆23Jan 8, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 【2024 ECAI】First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending☆14Jun 16, 2025Updated last year
- [ICCV 2025] Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation☆91Sep 29, 2025Updated 11 months ago
- ☆16Jun 23, 2026Updated 2 months ago
- [CVPR-26] Official repository of "CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization"☆19Mar 9, 2026Updated 5 months ago
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- This is the 2024 OS lab repository.☆11Jun 27, 2024Updated 2 years ago
- Official Repo For PerceptionDLM Codebase☆77Jun 22, 2026Updated 2 months ago
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- Multimodal Safety Awareness Benchmark for Large Language Models☆15Jun 3, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The good practice in the VQA system such as pos-tag attention, structed triplet learning and triplet attention is very general and can be…☆19Jan 23, 2018Updated 8 years ago
- ☆19Jan 2, 2026Updated 7 months ago
- code for downloading videos from HowTo100M dataset☆18May 13, 2021Updated 5 years ago
- Official repository of **CloudFixer: Test-Time Adaptation for 3D Point Clouds via Diffusion-Guided Geometric Transformation, ECCV'24**☆13Jun 16, 2025Updated last year
- 面向对象学习小项目,学生信息管理系统☆10Oct 6, 2019Updated 6 years ago
- SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability☆17May 8, 2025Updated last year
- imagetokenizer is a python package, helps you encoder visuals and generate visuals token ids from codebook, supports both image and video…☆40Jun 22, 2024Updated 2 years ago