[COLING 2025π₯] Evolver: Chain-of-Evolution Prompting to Boost Large Multimodal Models for Hateful Meme Detection
β17Jan 21, 2025Updated last year
Alternatives and similar repositories for Evolver
Users that are interested in Evolver are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VideoNIAH: A Flexible Synthetic Method for Benchmarking Video MLLMsβ57Mar 9, 2025Updated last year
- β10Jun 18, 2023Updated 3 years ago
- Official Code for the WWW'24 Paper: "Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language Models"β26Apr 16, 2025Updated last year
- Official repository for WWW'24 paper "MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation"β12Jul 25, 2024Updated 2 years ago
- Official Repository: A Comprehensive Benchmark for Logical Reasoning in MLLMsβ45Jun 17, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- π ACL 2024: RGCL, Retrieval-Guided Contrastive Learning for Hateful Meme Detection π EMNLP 2025 (Oral): RA-HMD, Robust Adaptation of Laβ¦β40Mar 1, 2026Updated 5 months ago
- Official repository for ACM Multimedia'23 paper "MATK: The Meme Analytical Tool Kit"β14May 29, 2024Updated 2 years ago
- Repo for paper "T2Vid: Translating Long Text into Multi-Image is the Catalyst for Video-LLMs"β48Sep 3, 2025Updated 11 months ago
- GPT-4V(ision) as A Social Media Analysis Engineβ39Dec 20, 2024Updated last year
- MR. Video: MapReduce is the Principle for Long Video Understandingβ31Jun 18, 2026Updated last month
- β15Apr 25, 2025Updated last year
- [ICCV 2025] The official pytorch implement of "LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs".β24Oct 28, 2025Updated 9 months ago
- [ECCV 2026π₯] This is the official implementation of our paper "SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perceptionβ¦β65Apr 2, 2026Updated 4 months ago
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videosβ24May 7, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β13May 17, 2025Updated last year
- Developer project for getting basic API integrations working in under 5 minutesβ11May 22, 2026Updated 2 months ago
- β27Apr 5, 2024Updated 2 years ago
- Efficient Foundation Model Design: A Perspective From Model and System Co-Design [Efficient ML System & Model]β31Feb 23, 2025Updated last year
- Code for ICML21 paper "Learning Self-Modulating Attention in Continuous Time Space with Applications to Sequential Recommendation"β12Feb 8, 2023Updated 3 years ago
- A Massive Multi-Discipline Lecture Understanding Benchmarkβ34Apr 20, 2026Updated 3 months ago
- [EMNLP 2024 Findingsπ₯] Official implementation of ": LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inβ¦β103Nov 9, 2024Updated last year
- [ICLR2025] Ξ³ -MOD: Mixture-of-Depth Adaptation for Multimodal Large Language Modelsβ45Oct 28, 2025Updated 9 months ago
- Video Summarization Transformer: Implementation in PyTorch of the Transformer model for video summarisationβ10Oct 27, 2020Updated 5 years ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictorsβ23May 19, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Unleashing Hour-Scale Video Training for Long Video-Language Understandingβ19Jun 24, 2025Updated last year
- β15Oct 7, 2024Updated last year
- [NeurIPS 2025] AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understandingβ15Updated this week
- An official implementation of SwapAnyone.β77Mar 14, 2025Updated last year
- [ICCV 2025] p-MoD: Building Mixture-of-Depths MLLMs via Progressive Ratio Decayβ43Jun 26, 2025Updated last year
- A transformer model that should be able to solve a simple NER taskβ11Mar 7, 2019Updated 7 years ago
- β41Sep 9, 2025Updated 11 months ago
- β67Aug 7, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β11May 24, 2024Updated 2 years ago
- Unleashing Reasoning in Medical Large Language Modelsβ12Mar 19, 2025Updated last year
- To mitigate position bias in LLMs, especially in long-context scenarios, we scale only one dimension of LLMs, reducing position bias and β¦β12Jun 18, 2024Updated 2 years ago
- SafeSora is a human preference dataset designed to support safety alignment research in the text-to-video generation field, aiming to enhβ¦β36Aug 20, 2024Updated last year
- Adaptive Topology Reconstruction for Robust Graph Representation Learning [Efficient ML Model]β10Feb 11, 2025Updated last year
- π This is a repository for organizing papers, codes, and other resources related to personalized video generation and editing.β64Dec 9, 2025Updated 8 months ago
- Code implementation of paper "MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval (AAAI2025)"β26Feb 2, 2025Updated last year