Official repository for the General Robust Image Task (GRIT) Benchmark
☆56Mar 29, 2023Updated 3 years ago
Alternatives and similar repositories for grit_official
Users that are interested in grit_official are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- EPIC-Kitchens-100 Action Recognition baselines: TSN, TRN, TSM☆33Mar 15, 2022Updated 4 years ago
- ☆78Jul 3, 2024Updated 2 years ago
- When do we not need larger vision models?☆420Feb 8, 2025Updated last year
- General-purpose Visual Understanding Evaluation☆20Dec 21, 2023Updated 2 years ago
- [ICCV 2023] Code for "Multi-task View Synthesis with Neural Radiance Fields"☆12Oct 2, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆231Dec 18, 2023Updated 2 years ago
- [WACV 2023] Code for "Beyond RGB: Scene-Property Synthesis with Neural Radiance Fields"☆30Jan 13, 2023Updated 3 years ago
- Original reference implementation of "Analyzing the Internals of Neural Radiance Fields"☆11Apr 10, 2024Updated 2 years ago
- Real-time application of DIVeR☆59Nov 23, 2021Updated 4 years ago
- Code for ECCV 2020 paper - LEMMA: A Multi-view Dataset for LEarning Multi-agent Multi-task Activities☆32Apr 8, 2021Updated 5 years ago
- [ECCV2022] Unstructured Feature Decoupling for Vehicle Re-Identification (UFDN)☆27Oct 13, 2022Updated 3 years ago
- ☆11Mar 4, 2025Updated last year
- ☆28Aug 9, 2024Updated last year
- 🤖 Autonomous Drone DJI Tello with OpenCv and ImageAI☆12Jun 18, 2019Updated 7 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Continual Learning Toolbox for Computer Vision Tasks☆20Oct 3, 2023Updated 2 years ago
- [ICML2024] Repo for the paper `Evaluating and Analyzing Relationship Hallucinations in Large Vision-Language Models'☆24Jan 1, 2025Updated last year
- ☆74Sep 23, 2025Updated 9 months ago
- SlideVQA: A Dataset for Document Visual Question Answering on Multiple Images (AAAI2023)☆106Mar 31, 2025Updated last year
- ☆11Sep 29, 2018Updated 7 years ago
- ☆14Aug 22, 2025Updated 10 months ago
- ☆12Jun 17, 2023Updated 3 years ago
- OpenLock Environment for OpenAI Gym☆19Feb 16, 2021Updated 5 years ago
- Official Implementation for "ESCAPE: Encoding Super-keypoints for Category-Agnostic Pose Estimation", CVPR 2024.☆10Jun 17, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LL3M: Large Language and Multi-Modal Model in Jax☆74Apr 23, 2024Updated 2 years ago
- ☆14Oct 10, 2022Updated 3 years ago
- DataComp: In search of the next generation of multimodal datasets☆787Apr 28, 2025Updated last year
- Official code repository for Findings of EMNLP 2022 paper: PseudoReasoner: Leveraging Pseudo Labels for Commonsense Knowledge Base Popula…☆11Oct 18, 2022Updated 3 years ago
- code release of research paper "Exploring Long-Sequence Masked Autoencoders"☆100Oct 14, 2022Updated 3 years ago
- Blender rendering script for multi-view images of 3D objects (ModelNet, ShapeNet, ...)☆13Oct 22, 2024Updated last year
- Market-1501 dataset with super-resolution quality☆21May 12, 2022Updated 4 years ago
- Towers of Babel: Combining Images, Language, and 3D Geometry for Learning Multimodal Vision. ICCV 2021.☆43Apr 30, 2024Updated 2 years ago
- This is the code repo for Findings of EMNLP2022 paper: MICO: a multi-alternative contrastive learning framework for commonsense knowledg…☆10Nov 29, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- https://jiaweisii.github.io/gorgeous/☆18Feb 24, 2026Updated 4 months ago
- ☆14Nov 19, 2017Updated 8 years ago
- Filtering, Distillation, and Hard Negatives for Vision-Language Pre-Training☆141Dec 16, 2025Updated 7 months ago
- A curve-editor for Stable Diffusion prompt interpolation☆21Oct 3, 2022Updated 3 years ago
- [NeurIPS2022] This is the official implementation of the paper "Expediting Large-Scale Vision Transformer for Dense Prediction without Fi…☆87Oct 29, 2023Updated 2 years ago
- Conceptual 12M is a dataset containing (image-URL, caption) pairs collected for vision-and-language pre-training.☆426Jul 14, 2025Updated last year
- LobotoMl is a set of scripts and tools to assess production deployments of ML services☆10May 16, 2022Updated 4 years ago