☆64Aug 29, 2026Updated this week
Alternatives and similar repositories for awesome_ai_paper
Users that are interested in awesome_ai_paper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17May 2, 2024Updated 2 years ago
- Official Implement of the paper "Unifying Segment Anything in Microscopy with Multimodal Large Language Model"☆20Apr 27, 2026Updated 4 months ago
- [KDD 2026 ADS Track] Pytorch implementation of the paper "Hi-Guard: Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasonin…☆26Jan 13, 2026Updated 7 months ago
- 基于PaddleNLP的对话意图识别☆10Apr 11, 2023Updated 3 years ago
- [ICLR 2026] Constructive Distortion: Multimodal LLMs with Attention‑Aware Image Warping☆24Feb 9, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [CVPR 2025] PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models☆54Jun 12, 2025Updated last year
- [CVPR 2024] Improving language-visual pretraining efficiency by perform cluster-based masking on images.☆33May 16, 2024Updated 2 years ago
- ☆17Jul 30, 2024Updated 2 years ago
- ☆136Mar 22, 2025Updated last year
- 中文原生多层次文生视频测评基准☆18Jul 8, 2024Updated 2 years ago
- An implementation of vConv layer.☆11Apr 28, 2021Updated 5 years ago
- ☆13Oct 17, 2024Updated last year
- Official implementation of ICLR 2026: Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement☆15May 24, 2026Updated 3 months ago
- [ICML2022] "Identity-Disentangled Adversarial Augmentation for Self-Supervised Learning"☆10Jul 24, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LLM wiki that evolves with you☆15Aug 23, 2026Updated last week
- [NAACL 2024] Vision language model that reduces hallucinations through self-feedback guided revision. Visualizes attentions on image feat…☆49Aug 21, 2024Updated 2 years ago
- (ICML 2024) Improve Context Understanding in Multimodal Large Language Models via Multimodal Composition Learning☆28Sep 27, 2024Updated last year
- BioSeq-BLM: a platform for analyzing DNA, RNA and protein sequences based on biological language models☆14Aug 21, 2022Updated 4 years ago
- Weakly-supervised road-lane markings detection for autonomous driving, mitigating the lack of training data☆14Oct 8, 2024Updated last year
- Code implementation of our ICCV 2025 paper: On Large Multimodal Models as Open-World Image Classifiers☆26Dec 4, 2025Updated 8 months ago
- ☆15Sep 2, 2024Updated last year
- ☆15Dec 28, 2022Updated 3 years ago
- Code for paper "Spider: Any-to-Many Multimodal LLM"☆16Apr 26, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Currently collecting some awesome Manus replays. Feel free to share your use cases.☆18Mar 9, 2025Updated last year
- 🌟 A curated list of papers, methods, and resources on long-horizon credit assignment for agentic RL.☆62Aug 6, 2026Updated 3 weeks ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- Code for "Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models"☆18Feb 16, 2026Updated 6 months ago
- [NeurIPS 2025] Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs☆39Sep 21, 2025Updated 11 months ago
- Official implementation for "Causal Intervention for Subject-Deconfounded Facial Action Unit Recognition" (AAAI 2022 Oral).☆17Mar 11, 2025Updated last year
- [ECCV2024] Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty Correction☆18Jul 12, 2024Updated 2 years ago
- Code for ICCV2023 paper: Homography Guided Temporal Fusion for Road Line and Marking Segmentation☆14Oct 13, 2024Updated last year
- Official Implementation of Moment Alignment Transformer☆17Oct 18, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Survey on Benchmarks of Multimodal Large Language Models☆159Jul 13, 2026Updated last month
- [NeurIPS 2024] Can Language Models Learn to Skip Steps?☆22Jan 25, 2025Updated last year
- Multi-Label Classification with Weighted Classifier Selection and Stacked Ensemble☆15Mar 30, 2020Updated 6 years ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- ☆10Jul 4, 2024Updated 2 years ago
- Datasets & Code for the WACV 2024 paper 'Robust Source-Free Domain Adaptation for Fundus Image Segmentation'☆13Jan 26, 2024Updated 2 years ago
- Watch Every Step! LLM Agent Learning via Iterative Step-level Process Refinement (EMNLP 2024 Main Conference)☆67Oct 18, 2024Updated last year