A systematic AI Agent development tutorial covering LLM agents, RAG, tool use, memory systems, multi-agent systems, LangChain, LangGraph, MCP, and agentic RL.|从零开始学 AI Agent 开发 | 系统、全面、实战导向的 Agent 开发教程 | 每日自动追踪 arXiv 最新论文 | Learn AI Agent Development from Scratch
☆550Sep 27, 2026Updated this week
Alternatives and similar repositories for agent_learning
Users that are interested in agent_learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆143Jul 2, 2026Updated 2 months ago
- 超简单使用监督微调SFT和强化学习RL去训练领域Agent☆35Oct 20, 2025Updated 11 months ago
- 调用大模型已经是如今做 ai 项目习以为常的工作的,但是大模型的输出很多时候是不可控的,我们又需要使用大模型去做各种下游任务,实现可控可解析的输出。我们探索了一种和 python 开发可以紧密合作的开发方法。☆30Jun 18, 2024Updated 2 years ago
- 100 天搞定 Agent 开发☆747Apr 26, 2026Updated 5 months ago
- 🚀 2026 最系统的 AI Agent 速成指南|智能体实战教程 · 完整学习路径 + 实战项目 + 面试题库 · 对标大模型应用开发工程师岗位 · 覆盖LangChain / LangGraph / Coze / Dify / MCP / skills / LLM…☆4,933Sep 10, 2026Updated 2 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆11Feb 26, 2024Updated 2 years ago
- Source code of “Reinforcement Learning with Token-level Feedback for Controllable Text Generation (NAACL 2024)☆17Dec 8, 2024Updated last year
- ☆27Sep 25, 2025Updated last year
- Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning☆25Jun 25, 2025Updated last year
- Build your own Claude Code from scratch. 🔍 Claude Code 开源了 50 万行代码,读不动?用 ~5000 行 TypeScript / Python 从零复现核心架构,11 章分步教程带你理解 coding agent…☆2,715Jul 9, 2026Updated 2 months ago
- 2022年华为软件精英赛初赛☆11Apr 2, 2022Updated 4 years ago
- Topic Embedding, Text Generation and Modeling using diffusion☆15Jun 10, 2026Updated 3 months ago
- AI Agent 面试知识库 100 题 | 涵盖 Agent 架构、RAG、工具使用、多 Agent、记忆、规划推理、提示工程、评估、安全对齐、生产部署、框架选型☆359Aug 15, 2026Updated last month
- CFT-RAG: An Entity Tree Based Retrieval Augmented Generation Algorithm With Cuckoo Filter☆26May 28, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for experiments in the paper: "Compositional Reinforcement Learning from Logical Specifications" (https://arxiv.org/abs/2106.13906).☆16Oct 26, 2021Updated 4 years ago
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision☆18Apr 1, 2025Updated last year
- Code for NeurIPS 2021 paper "Curriculum Offline Imitation Learning"☆18Oct 21, 2022Updated 3 years ago
- (NBCE)Naive Bayes-based Context Extension on ChatGLM-6b☆15Jun 7, 2023Updated 3 years ago
- A Chinese-focused PyTorch framework for exploring Attention Residuals in Qwen3-style causal LMs, with baseline, Block AttnRes, Full AttnR…☆21May 3, 2026Updated 4 months ago
- 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程☆80,985Updated this week
- PMD Homepage☆12Updated this week
- Multi-Object Tracking with Ultralytics YOLO11☆13Oct 5, 2024Updated last year
- 大模型应用开发 -上下文工程与运行空间实践指南,上下文工程是设计原则,Agent Harness 是构建目标,本项目的目标,是为开发者和研究者提供一份大模型应用开发的骨架思路☆821Jul 15, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- realtime-emd(Empirical Mode Decomposition) from code.google.com/archive/p/realtime-emd/☆14Aug 9, 2016Updated 10 years ago
- EMVCC:Enhanced Multi-View Contrastive Clustering for Hyperspectral Images☆15Mar 5, 2026Updated 6 months ago
- Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1☆77,675Aug 26, 2026Updated last month
- 从0到1多模态大模型 · 理论与实战学习记录 From0to1-MLLM-StudyLog 是一个个人从零自学多模态大模型(MLLM)的系统记录仓库,覆盖约 24 周的学习与实践过程。 仓库按 Week1–Week24 组织,每周包含: 精简的理论理解与知识梳理 关键论…☆186Sep 10, 2026Updated 2 weeks ago
- Meta-Reinforcement Learning with Policy Residual Representation☆11Aug 15, 2019Updated 7 years ago
- AI Agent 面试全攻略:从零到Offer,包含200+面试题、企业级项目(Python/Java/Go)、简历模板、STAR面试稿、哆啦A梦漫画图解☆2,834Apr 1, 2026Updated 5 months ago
- MSTI☆16Mar 6, 2024Updated 2 years ago
- https://adongwanai.github.io/AgentGuide | AI Agent开发指南 | LangGraph实战 | 高级RAG | 转行大模型 | 大模型面试 | 算法工程师 | 面试题库 | 强化学习|数据合成☆10,147Sep 15, 2026Updated last week
- Pure Java library for creating LLVM IR☆12Jun 9, 2013Updated 13 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The SQL-RL-GEN is an algorithm based on a Reinforcement Learning approach with a reward function generated by a LLM to guide the agent's …☆25Sep 18, 2025Updated last year
- ☆25Jul 24, 2023Updated 3 years ago
- Agentic Reinforcement Learning 101. A pragmatic course for AI/ML Engineers based on "The Landscape of Agentic Reinforcement Learning for …☆35Sep 16, 2025Updated last year
- 人脸识别、人脸细粒度表情识别、异常行为检测和识别☆12Apr 11, 2022Updated 4 years ago
- ☆34Mar 20, 2026Updated 6 months ago
- A classic 5-stage rv32i(incomplete) toy implementation based on powerful SpinalHDL☆10Jul 5, 2021Updated 5 years ago
- This is a repo consisting of papers about LLMs' perception of their knowledge boundaries; Uncertainty Quantification; Honesty Alignment; …☆25Nov 25, 2025Updated 10 months ago