InstructionGPT-4
☆42Dec 29, 2023Updated 2 years ago
Alternatives and similar repositories for InstructionGPT-4
Users that are interested in InstructionGPT-4 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Less is More: High-value Data Selection for Visual Instruction Tuning☆20Jan 18, 2025Updated last year
- Official PyTorch implementation of “MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation”☆18Dec 5, 2024Updated last year
- ☆11Jun 2, 2019Updated 7 years ago
- The code for paper: PeFoMed: Parameter Efficient Fine-tuning on Multi-modal Large Language Models for Medical Visual Question Answering☆64Dec 21, 2025Updated 8 months ago
- This is AlpaGasus2-QLoRA based on LLaMA2 with AlpaGasus mechanism using QLoRA!☆15Nov 22, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2025] First SFT, Second RL, Third UPT: Continual Improving Multi-Modal LLM Reasoning via Unsupervised Post-Training☆90Oct 29, 2025Updated 10 months ago
- ☆10Apr 3, 2021Updated 5 years ago
- AutoHallusion Codebase (EMNLP 2024)☆22Dec 6, 2024Updated last year
- (ArXiv25) Vision Matters: Simple Visual Perturbations Can Boost Multimodal Math Reasoning☆61Sep 30, 2025Updated 11 months ago
- ☆22May 21, 2025Updated last year
- Creative AI for Visual Art and Music slides and demos.☆11May 2, 2023Updated 3 years ago
- ☆15Apr 23, 2026Updated 4 months ago
- ☆13Jul 17, 2024Updated 2 years ago
- ☆23Jan 17, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official code of "PixelWorld: Towards Perceiving Everything as Pixels" [TMLR25]☆15Sep 12, 2025Updated 11 months ago
- (ACL 2025) MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale☆50Jun 4, 2025Updated last year
- ☆28Jul 10, 2025Updated last year
- Official implementation of EMNLP'2022 paper "Non-Parametric Domain Adaptation for End-to-End Speech Translation"☆11Oct 26, 2022Updated 3 years ago
- ☆14Jun 11, 2024Updated 2 years ago
- Code for the paper CVPR‘17 “Zero Shot Learning from Noisy Text Description at Part Precision”☆16Nov 22, 2019Updated 6 years ago
- ☆16Jul 1, 2024Updated 2 years ago
- Artistic Vision-Language Understanding with Adapter-enhanced MiniGPT-4☆30May 31, 2023Updated 3 years ago
- List of papers on Hallucination in LMM☆10Nov 29, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2025] NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation☆112Sep 18, 2025Updated 11 months ago
- ☆37Jan 25, 2024Updated 2 years ago
- Official implementation of SIGIR 2022 Paper "Task-Oriented Dialogue System as Natural Language Generation".☆14Apr 6, 2022Updated 4 years ago
- ☆11Feb 25, 2024Updated 2 years ago
- ☆10Mar 4, 2024Updated 2 years ago
- A collection of visual instruction tuning datasets.☆76Mar 14, 2024Updated 2 years ago
- SFT+RL boosts multimodal reasoning☆50Jun 27, 2025Updated last year
- Cross-Perspective Topic Modeling☆11Oct 27, 2017Updated 8 years ago
- Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents, CVPR 2025☆26Jan 25, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning☆296Mar 13, 2024Updated 2 years ago
- ☆16Aug 11, 2025Updated last year
- Exploring Hierarchical Graph Representation for Large-Scale Zero-Shot Image Classification. ECCV 2022.☆18Jul 12, 2022Updated 4 years ago
- TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones☆1,316Feb 5, 2026Updated 6 months ago
- MoCLE (First MLLM with MoE for instruction customization and generalization!) (https://arxiv.org/abs/2312.12379)☆46Jul 1, 2025Updated last year
- ☆14May 20, 2025Updated last year
- ☆35Nov 18, 2025Updated 9 months ago