☆21Jul 9, 2025Updated last year
Alternatives and similar repositories for VisuLogic-Train
Users that are interested in VisuLogic-Train are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆38Aug 18, 2025Updated last year
- ☆18Nov 30, 2025Updated 9 months ago
- LMM for VQA, tcsvt version☆10Jul 19, 2024Updated 2 years ago
- ☆12Dec 20, 2024Updated last year
- PKU-I2IQA: An Image-to-Image Quality Assessment Database for AI Generated Images☆15Dec 4, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Jul 4, 2024Updated 2 years ago
- ☆49Oct 28, 2024Updated last year
- EventHallusion: Diagnosing Event Hallucinations in Video LLMs☆34Aug 5, 2025Updated last year
- ☆11Aug 20, 2025Updated last year
- 首届社交群体智能算法大赛 【赛题1:社交媒体舆论场虚假账号检测】第三名(0.8248)方案☆12May 30, 2024Updated 2 years ago
- Official eval code for ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation☆29Dec 12, 2025Updated 9 months ago
- Official PyTorch implementation of the paper "Equivariant Image Modeling"(https://arxiv.org/abs/2503.18948)☆37Aug 1, 2025Updated last year
- Repository containing code for CoRL 2020 paper on "Learning Object Manipulation Skills via Approximate State Estimation from Real Videos"☆17Dec 15, 2021Updated 4 years ago
- A collection of research papers related to Natural Language Reasoning☆10May 27, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year
- [ICML 2026] Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMs☆35Apr 27, 2026Updated 4 months ago
- Github repository for "Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging" (ICML 2025)☆91Jun 9, 2026Updated 3 months ago
- [TCSVT'22] Official Implementation of STI-VQA☆12Oct 18, 2023Updated 2 years ago
- [AAAI'26] Official implementation of CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augm…☆11Dec 5, 2025Updated 9 months ago
- Dual-Branch Network for Portrait Image Quality Assessment☆19Aug 28, 2026Updated 3 weeks ago
- "Blind Image Quality Assessment for Pathological Microscopic Image under Screen and Immersion Scenarios"☆15Aug 29, 2023Updated 3 years ago
- [EMNLP 2024] SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information☆11Oct 11, 2024Updated last year
- Tis is code for Few-Shot Joint Multimodal Entity-Relation Extraction via Knowledge-Enhanced Cross-modal Prompt Model (ACM MM 2024))☆12Aug 27, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Dec 9, 2024Updated last year
- X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains☆48Feb 4, 2026Updated 7 months ago
- [NAACL 2025] Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning☆11Feb 9, 2025Updated last year
- ☆10Apr 13, 2020Updated 6 years ago
- [AAAI 2024] MESED: A Multi-modal Entity Set Expansion Dataset with Fine-grained Semantic Classes and Hard Negative Entities☆15Apr 26, 2024Updated 2 years ago
- ☆14May 23, 2022Updated 4 years ago
- ☆15Jul 17, 2025Updated last year
- Benchmarking Multi-Image Understanding in Vision and Language Models☆11Jul 29, 2024Updated 2 years ago
- code for G2LTraj☆18Dec 23, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Codebase for VidHal: Benchmarking Hallucinations in Vision LLMs☆14Apr 23, 2026Updated 4 months ago
- [EMNLP2022] Transformer-based Entity Typing in Knowledge Graphs☆15Nov 26, 2024Updated last year
- SQAD: Automatic Smartphone Camera Quality Assessment and Benchmarking☆28Aug 23, 2025Updated last year
- The official code of "VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning" [NeurIPS25]☆192Jun 5, 2025Updated last year
- ☆36Feb 25, 2020Updated 6 years ago
- Official Repo for MageBench: Bridging Large Multimodal Models to Agents☆21Jan 8, 2025Updated last year
- Official code for Attention-driven GUI Grounding, AAAI2025☆16Dec 17, 2024Updated last year