《多模态大模型:新一代人工智能技术范式》配套教学资源
☆311Jun 17, 2026Updated last month
Alternatives and similar repositories for Book-of-MLM
Users that are interested in Book-of-MLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Embodied-AI-Survey-2025] Paper List and Resource Repository for Embodied AI☆2,129Jun 10, 2026Updated last month
- ☆14Oct 23, 2023Updated 2 years ago
- The official repository of [CVPR2025] DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering☆28Apr 18, 2025Updated last year
- [CVPR 2026] AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition☆40Apr 27, 2026Updated 2 months ago
- 🧑🚀 全世界最好的LLM资料总结(多模态生成、Agent、辅助编程、AI审稿、数据处理、模型训练、模型推理、o1 模型、MCP、小语言模型、视觉语言模型) | Summary of the world's best LLM resources.☆8,737Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 通用简单工具项目☆22Oct 6, 2024Updated last year
- ☆86Jun 16, 2026Updated last month
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆24Dec 19, 2023Updated 2 years ago
- 2025.01:从零到一实现了一个多模态大模型,并命名为Reyes(睿视),R:睿,eyes:眼。Reyes的参数量为8B,视觉编码器使用的是InternViT-300M-448px-V2_5,语言模型侧使用的是Qwen2.5-7B-Instruct,Reyes也通过一个两…☆34Feb 10, 2026Updated 5 months ago
- [ICLR 2024] Official repository for "Vision-by-Language for Training-Free Compositional Image Retrieval"☆89Jul 4, 2024Updated 2 years ago
- 《从零构建AIAgent:大模型驱动的智能体设计与实战》配套代码☆19Mar 2, 2026Updated 4 months ago
- The collections of MOE (Mixture Of Expert) papers, code and tools, etc.☆12Mar 15, 2024Updated 2 years ago
- The code of YOLOv5 inferencing with TensorRT C++ api is packaged into a dynamic link library , then called through Python.☆15Oct 23, 2025Updated 9 months ago
- ☆13Feb 23, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆40Aug 26, 2025Updated 10 months ago
- 本项目旨在分享大模型相关 技术原理以及实战经验(大模型工程化、大模型应用落地)☆24,800Updated this week
- 主要记录大语言大模型(LLMs) 算法(应用)工程师多模态相关知识☆288May 12, 2024Updated 2 years ago
- Code for "Harnessing Textual Semantic Priors for Knowledge Transfer and Refinement in CLIP-Driven Continual Learning" (AAAI-2026 poster)☆16Mar 13, 2026Updated 4 months ago
- ☆21Mar 1, 2022Updated 4 years ago
- 【GRSL 2021】ASEA for Change Detection☆11Oct 28, 2025Updated 8 months ago
- Visual Delta Generator with Large Multi-modal Model for Semi-supervised Composed Image Retrieval - CVPR2024☆21May 30, 2024Updated 2 years ago
- [CVPR 2026 Highlight 🔥] PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training☆43May 6, 2026Updated 2 months ago
- ☆14Sep 27, 2025Updated 9 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Latest Advances on Multimodal Large Language Models☆17,956Jul 2, 2026Updated 3 weeks ago
- Open Source Road Datasets☆19Aug 30, 2024Updated last year
- Multi-task Learning for Multi-modal Emotion Recognition and Sentiment Analysis☆13Mar 17, 2021Updated 5 years ago
- ☆13Aug 3, 2024Updated last year
- 大模型基础: 一文了解大模型基础知识☆7,507Jun 22, 2026Updated last month
- Integrating Task-Specific and Universal Adapters for Pre-Trained Model-based Class-Incremental Learning (ICCV 2025)☆18Sep 23, 2025Updated 10 months ago
- This repository integrates the codes for some feature selection & clustering methods.☆18Jun 9, 2022Updated 4 years ago
- ☆24Apr 16, 2022Updated 4 years ago
- 可以成功Lora微调的Qwen-VL模型☆16Oct 27, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Llama3-Tutorial(XTuner、LMDeploy、OpenCompass)☆506May 10, 2024Updated 2 years ago
- 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians (ACM MM 25)☆79Jul 21, 2025Updated last year
- [IEEE T-PAMI 2023] Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering☆78Jul 6, 2023Updated 3 years ago
- A book for Learning the Foundations of LLMs☆16,503Dec 12, 2025Updated 7 months ago
- Large Language Model in Action☆342Jan 28, 2025Updated last year
- 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程☆31,410Jul 15, 2026Updated last week
- ☆434Apr 29, 2025Updated last year