LLM Architecture Gallery source data
☆1,401Jul 19, 2026Updated this week
Alternatives and similar repositories for llm-architecture-gallery
Users that are interested in llm-architecture-gallery are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implement a reasoning LLM in PyTorch from scratch, step by step☆4,782Jul 6, 2026Updated 2 weeks ago
- Minimal and readable coding agent harness implementation in Python to explain the core components of coding agents.☆1,033Apr 7, 2026Updated 3 months ago
- AI agents running research on single-GPU nanochat training automatically☆91,712Mar 26, 2026Updated 3 months ago
- Implement a ChatGPT-like LLM in PyTorch from scratch, step by step☆99,511Jul 11, 2026Updated last week
- Materials for the "My Workflow for Understanding LLM Architectures" tutorial☆25Apr 10, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The best ChatGPT that $100 can buy.☆56,507Jul 4, 2026Updated 2 weeks ago
- Train the smallest LM you can that fits in 16MB. Best model wins!☆5,166May 4, 2026Updated 2 months ago
- Agent Cage is a Docker-based sandbox prototype that gives AI agents complete access to an Ubuntu environment while experimenting with IT-…☆63Jun 17, 2026Updated last month
- A high-throughput and memory-efficient inference and serving engine for LLMs☆86,804Updated this week
- Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.☆68,666Updated this week
- Nano vLLM☆14,582Apr 26, 2026Updated 2 months ago
- Machine Learning Systems☆27,524Updated this week
- Official code repo for the O'Reilly Book - "Hands-On Large Language Models"☆27,718Apr 24, 2026Updated 2 months ago
- A theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.☆14,737May 23, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The absolute trainer to light up AI agents.☆17,408Updated this week
- A computer you can curl ⚡☆2,888Jun 22, 2026Updated 3 weeks ago
- Self-referential self-improving agents that can optimize for any computable task☆2,645May 9, 2026Updated 2 months ago
- Learn it. Build it. Ship it for others.☆41,296Jun 25, 2026Updated 3 weeks ago
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆61,385Nov 12, 2025Updated 8 months ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆30,583Updated this week
- An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, s…☆77,530Updated this week
- The first distributed AGI system. Thousands of autonomous AI agents collaboratively train models, share experiments via P2P gossip, and p…☆2,004Updated this week
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,691Feb 27, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆6,719Jul 9, 2026Updated last week
- The agent that grows with you☆218,250Updated this week
- This guide demonstrates how to build an Agentic AI system using Google's Agent Development Kit (ADK) and the Model Context Protocol (MCP)…☆79Mar 20, 2026Updated 4 months ago
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆10,782Updated this week
- Multi-agent systems, memory, planning, reasoning loops☆2,810Updated this week
- 100+ AI Agent & RAG apps you can actually run — clone, customize, ship.☆125,758Updated this week
- 🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models☆10,672Updated this week
- slime is an LLM post-training framework for RL Scaling.☆7,569Updated this week
- LLM Council works together to answer your hardest questions☆22,951Nov 22, 2025Updated 7 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, C…☆231,807Updated this week
- Give your agents the power of the Hugging Face ecosystem☆10,854Updated this week
- ☆3,358Mar 17, 2026Updated 4 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,587Updated this week
- 35 production-grade agentic AI architectures (Reflexion, LATS, GraphRAG, MemGPT, Voyager, BrowserAgent, ...) — a Python library and runna…☆3,911Jun 22, 2026Updated 3 weeks ago
- Run Claude Code/Codex within AgentFS, orchestrated by LlamaIndex Workflows☆325Dec 19, 2025Updated 7 months ago
- Qwen3.6 is the large language model series developed by Qwen team, Alibaba Group.☆3,705Jun 3, 2026Updated last month