LLM Architecture Gallery source data
☆1,551Sep 13, 2026Updated last week
Alternatives and similar repositories for llm-architecture-gallery
Users that are interested in llm-architecture-gallery are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implement a reasoning LLM in PyTorch from scratch, step by step☆5,253Updated this week
- Minimal and readable coding agent harness implementation in Python to explain the core components of coding agents.☆1,169Apr 7, 2026Updated 5 months ago
- AI agents running research on single-GPU nanochat training automatically☆96,417Mar 26, 2026Updated 5 months ago
- Implement a ChatGPT-like LLM in PyTorch from scratch, step by step☆105,289Updated this week
- Materials for the "My Workflow for Understanding LLM Architectures" tutorial☆30Apr 10, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The best ChatGPT that $100 can buy.☆58,164Sep 7, 2026Updated last week
- Train the smallest LM you can that fits in 16MB. Best model wins!☆5,186May 4, 2026Updated 4 months ago
- Agent Cage is a Docker-based sandbox prototype that gives AI agents complete access to an Ubuntu environment while experimenting with IT-…☆63Jun 17, 2026Updated 3 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆92,241Updated this week
- Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.☆76,481Updated this week
- Nano vLLM☆15,535Apr 26, 2026Updated 4 months ago
- Machine Learning Systems: Foundations, Scaling, Agentic AI, and Physical AI (Vols I–IV) • Harvard CS249r | https://mlsysbook.ai☆28,379Updated this week
- Official code repo for the O'Reilly Book - "Hands-On Large Language Models"☆29,206Apr 24, 2026Updated 4 months ago
- A theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.☆14,902May 23, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The absolute trainer to light up AI agents.☆18,408Updated this week
- A computer you can curl ⚡☆3,187Updated this week
- Self-referential self-improving agents that can optimize for any computable task☆2,750Jul 31, 2026Updated last month
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆63,254Nov 12, 2025Updated 10 months ago
- An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, s…☆82,751Updated this week
- SGLang is a high-performance serving framework for large language models and multimodal models.☆36,212Updated this week
- The first distributed AGI system. Thousands of autonomous AI agents collaboratively train models, share experiments via P2P gossip, and p…☆2,055Updated this week
- Learn it. Build it. Ship it for others.☆55,078Sep 7, 2026Updated last week
- Tongyi Deep Research, the Leading Open-source Deep Research Agent☆19,976Feb 27, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The agent that grows with you☆247,393Updated this week
- This guide demonstrates how to build an Agentic AI system using Google's Agent Development Kit (ADK) and the Model Context Protocol (MCP)…☆80Mar 20, 2026Updated 6 months ago
- Archived — ML Intern is no longer maintained. Continue with HuggingChat.☆10,821Updated this week
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆11,875Updated this week
- 100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.☆139,105Updated this week
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆7,138Jul 9, 2026Updated 2 months ago
- Multi-agent systems, memory, planning, reasoning loops☆2,906Updated this week
- ☆3,515Mar 17, 2026Updated 6 months ago
- LLM Council works together to answer your hardest questions☆24,924Nov 22, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- slime is an LLM post-training framework for RL Scaling.☆8,509Updated this week
- Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)☆74,919Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,503Updated this week
- The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, C…☆263,731Updated this week
- Give your agents the power of the Hugging Face ecosystem☆11,072Updated this week
- Run Claude Code/Codex within AgentFS, orchestrated by LlamaIndex Workflows☆325Dec 19, 2025Updated 9 months ago
- 35 production-grade agentic AI architectures (Reflexion, LATS, GraphRAG, MemGPT, Voyager, BrowserAgent, ...) — a Python library and runna…☆4,534Jun 22, 2026Updated 2 months ago