Scaffold Prompting to promote LMMs
☆46Dec 16, 2024Updated last year
Alternatives and similar repositories for Scaffold
Users that are interested in Scaffold are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Dettoolchain: A new prompting paradigm to unleash detection ability of MLLM☆45Oct 12, 2024Updated last year
- Official Implementation of UA^{2}-Agent and other baseline algorithms of "Towards Unified Alignment Between Agents, Humans, and Environme…☆20Nov 12, 2024Updated last year
- Repo for paper "CODIS: Benchmarking Context-Dependent Visual Comprehension for Multimodal Large Language Models".☆13Oct 14, 2024Updated last year
- ☆12Dec 20, 2024Updated last year
- The repository of the ACCV 2024 paper "FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Ge…☆12Aug 15, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Language-Conditioned Path Planning (Amber Xie, Youngwoon Lee, Pieter Abbeel, Stephen James)☆27Sep 1, 2023Updated 3 years ago
- ☆19Oct 28, 2025Updated 10 months ago
- ☆20Sep 3, 2025Updated last year
- [CVPR 2024] Official Code for the Paper "Compositional Chain-of-Thought Prompting for Large Multimodal Models"☆144Jun 20, 2024Updated 2 years ago
- The Code for Lever LM: Configuring In-Context Sequence to Lever Large Vision Language Models☆19Oct 4, 2024Updated last year
- 目标检测,关键点检测。A pure version of CenterNet, convenient for secondary development and easy to understand.☆21Dec 9, 2020Updated 5 years ago
- [ICCV 2023] Going Beyond Nouns With Vision & Language Models Using Synthetic Data☆13Sep 30, 2023Updated 2 years ago
- ☆21Aug 9, 2024Updated 2 years ago
- Official Code Repository for EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents (COLM 2024)☆41Jul 13, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- (CVPR 2023) Official implemention of the paper "Weakly Supervised Video Representation Learning with Unaligned Text for Sequential Videos…☆31Apr 2, 2024Updated 2 years ago
- ☆20Aug 29, 2026Updated 3 weeks ago
- ☆35Apr 14, 2023Updated 3 years ago
- ☆15Apr 25, 2025Updated last year
- Repo for 2020 EMNLP paper "Conditional Causal Relationships between Emotions and Causes in Texts"☆14Apr 8, 2021Updated 5 years ago
- The collection of medical VLP papars☆20Jul 24, 2024Updated 2 years ago
- Official repo of Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains.☆43Jun 6, 2025Updated last year
- [NeurIPS 2024] Calibrated Self-Rewarding Vision Language Models☆87Oct 26, 2025Updated 10 months ago
- Graph Cut Algorithm in CUDA☆28Jun 1, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch code for "Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training"☆39Mar 4, 2024Updated 2 years ago
- 【AAAI 2026 🔥】A benchmark that evaluates multimodel knowledge conflicts for large multimodal model☆25May 27, 2025Updated last year
- [CVPR 2025] Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention☆69Jul 16, 2024Updated 2 years ago
- [EMNLP 2024] This is the code for our paper "BMRetriever: Tuning Large Language Models as Better Biomedical Text Retrievers".☆27Sep 19, 2024Updated 2 years ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆23Sep 5, 2026Updated 2 weeks ago
- ☆10Dec 15, 2024Updated last year
- ☆32Jun 30, 2026Updated 2 months ago
- Official implementation of "AnyPlace: Learning Generalized Object Placement for Robot Manipulation"☆104Mar 25, 2025Updated last year
- ☆17Mar 10, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2025] VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning☆69Sep 20, 2025Updated last year
- AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World | CoRL 2025☆107Mar 26, 2026Updated 5 months ago
- Code repo for the paper: Attacking Vision-Language Computer Agents via Pop-ups☆52Dec 23, 2024Updated last year
- ☆14Oct 25, 2024Updated last year
- A toolkit for dialogue system evaluation via crowdsourcing☆18Apr 25, 2023Updated 3 years ago
- [CVPR'24 Highlight] The official code and data for paper "EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Lan…☆67Mar 25, 2025Updated last year
- A password manager☆12Jun 22, 2026Updated 3 months ago