LLM101n: Let's build a Storyteller 中文版
☆137Aug 15, 2024Updated last year
Alternatives and similar repositories for LLM101n-CN
Users that are interested in LLM101n-CN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch Sphinx Theme☆34Jan 3, 2024Updated 2 years ago
- ☆36Sep 21, 2025Updated 10 months ago
- ☆48Jul 12, 2023Updated 3 years ago
- Skill optimization framework for LLMs — evolve system prompts via textual gradient descent with beam search, human-in-the-loop annotation…☆20Apr 1, 2026Updated 3 months ago
- tutorial for writing custom pytorch cpp+cuda kernel, applied on volume rendering (NeRF)☆29Dec 12, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- LLM&VLM Tutorial☆1,969Apr 22, 2026Updated 3 months ago
- HunyuanDiT with TensorRT and libtorch☆18May 22, 2024Updated 2 years ago
- stay tuned.☆18Jul 7, 2025Updated last year
- 基于《西游记》原文、白话文、ChatGPT生成数 据制作的,以InternLM2微调的角色扮演多LLM聊天室。 本项目将介绍关于角色扮演类 LLM 的一切,从数据获取、数据处理,到使用 XTuner 微调并部署至 OpenXLab,再到使用 LMDeploy 部署,以 op…☆109Mar 31, 2024Updated 2 years ago
- 利用大语言模型进行卧底游戏,包括谁是卧底及衍生的发现AI卧底游戏等。☆11Sep 6, 2024Updated last year
- 顾名思义:手搓的RAG☆132Feb 27, 2024Updated 2 years ago
- 😊书生·浦语 (InternLM) 开源大语言模型第一期课程笔记😊 + 🍏Horowag Model Development🍎☆44Mar 5, 2024Updated 2 years ago
- InternEvo is an open-sourced lightweight training framework aims to support model pre-training without the need for extensive dependencie…☆421Aug 21, 2025Updated 11 months ago
- ☆111Mar 8, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 使用 CUDA C++ 实现的 llama 模型推理框架☆64Nov 8, 2024Updated last year
- ☆73Mar 12, 2024Updated 2 years ago
- 《自己动手写AI编译器》☆40Oct 19, 2024Updated last year
- 本仓库在OpenVINO推理框架下部署Nanodet检测算法,并重写预处理和后处理部分,具有超高性能!让你在Intel CPU平台上的检测速度起飞! 并基于NNCF和PPQ工具将模型量化(PTQ)至int8精度,推理速度更快!☆16Jun 14, 2023Updated 3 years ago
- Llama3-Tutorial(XTuner、LMDeploy、OpenCompass)☆506May 10, 2024Updated 2 years ago
- ☆22Aug 18, 2024Updated last year
- 筱可的工程实验仓库!☆115Oct 31, 2025Updated 8 months ago
- ☆179Mar 12, 2024Updated 2 years ago
- HuixiangDou: Overcoming Group Chat Scenarios with LLM-based Technical Assistance☆2,493Nov 24, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- FlashTile is a CUDA Tile IR compiler that is compatible with NVIDIA's tileiras, targeting SM70 through SM121 NVIDIA GPUs.☆61Feb 6, 2026Updated 5 months ago
- The official implement of Freeze-Omni.☆16Jul 10, 2025Updated last year
- 尝试自己从头写一个LLM,参考llama和nanogpt☆69Apr 27, 2024Updated 2 years ago
- ☆20Jan 19, 2022Updated 4 years ago
- ToyLLM: Learning LLM from Scratch☆25Updated this week
- LMDeploy is a toolkit for compressing, deploying, and serving LLMs.☆7,972Updated this week
- A lightweight framework for building LLM-based agents☆2,272Jul 6, 2026Updated 2 weeks ago
- ☆18Nov 22, 2025Updated 8 months ago
- [ACL2024 Findings] Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models☆361Mar 22, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆21Jan 16, 2024Updated 2 years ago
- ☆14May 21, 2024Updated 2 years ago
- A lightweight, production-ready C++ library for LLM tokenization, fully compatible with HuggingFace tokenizer.json.☆33Jan 4, 2026Updated 6 months ago
- 八戒-Chat是利用《西游记》剧本中所有关于猪八戒的台词和语句,以及Chat-GPT-3.5生成的相关问题结果,基于Internlm进行QLoRA微调得到的模仿猪八戒语气的聊天语言模型。☆27Jul 30, 2025Updated 11 months ago
- ☆17Sep 6, 2022Updated 3 years ago
- This project is primarily used to deploy large language models and multimodal large models on Orin.🚀🚀🚀☆18Jun 23, 2026Updated last month
- an implementation of Deformation Graph compatible with CUDA C++ and used in warping defamations in real-time non-rigid registration☆10Jan 22, 2020Updated 6 years ago