ToyLLM: Learning LLM from Scratch
☆25Jul 27, 2026Updated this week
Alternatives and similar repositories for toyllm
Users that are interested in toyllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🎉My Collections of CUDA Kernels~☆11Jun 25, 2024Updated 2 years ago
- 🔥Keywords and URLs Censored on the Chinese Internet☆13Feb 22, 2020Updated 6 years ago
- Flash Attention in ~100 lines of CUDA (forward pass only)☆12Jun 10, 2024Updated 2 years ago
- CUDA SGEMM optimization note☆15Oct 31, 2023Updated 2 years ago
- 登录脚本☆12Nov 4, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code execution sandbox(support Open-R1), Supports multiple languages(Python/Java/C/Kotlin/Swift/OC/GO/...)☆23Mar 6, 2025Updated last year
- ☆10Jun 15, 2024Updated 2 years ago
- ☆15Nov 18, 2025Updated 8 months ago
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- Apply Iprompt on GLM with innovative new methods. Currently support Chinese QA, English QA and Chinese poem generation.☆20Jun 16, 2022Updated 4 years ago
- Few-Shot Preference Optimization (FSPO) personalizes LLMs by reframing reward modeling as a meta-learning problem, enabling rapid adaptat…☆16Feb 27, 2025Updated last year
- ☆16Apr 30, 2025Updated last year
- 中国科学院大学(国科大)研一课程☆19May 24, 2023Updated 3 years ago
- ☆11May 26, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19Oct 8, 2024Updated last year
- A light llama-like llm inference framework based on the triton kernel.☆188Jan 5, 2026Updated 6 months ago
- 轻松获取结构体大小、偏移、常量值的工具。☆12May 12, 2022Updated 4 years ago
- LLMs + Persona-Plug = Personalized LLMs☆15Oct 16, 2024Updated last year
- 我收藏的机器学习相关文集☆12Updated this week
- Building a more intelligent world.☆11Apr 29, 2024Updated 2 years ago
- WiFi-based gesture recognition using Adversarial Discriminative Domain Adaptation.☆18Jun 24, 2022Updated 4 years ago
- ☆11May 16, 2026Updated 2 months ago
- A pure C++11 based parallel programming library☆15Aug 22, 2016Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Qwen2-VL在文旅领域的LLaMA-Factory微调案例 The case for fine-tuning Qwen2-VL in the field of historical literature and museums☆15Sep 17, 2024Updated last year
- A small ADB / Fastboot installer for Windows that automatically fetches the latest version from this Github Repository☆12Jul 19, 2021Updated 5 years ago
- 🎓Automatically Update circult-eda-mlsys-tinyml Papers Daily using Github Actions (Update Every 8th hours)☆10Updated this week
- ☆15Sep 9, 2020Updated 5 years ago
- [WWW2026] The official code for paper "Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based Recommenders"☆23Jan 23, 2026Updated 6 months ago
- The Source Code Lib for BiscuitOS☆12Oct 15, 2024Updated last year
- Example code for tracing IRQs in Linux kernel☆12Oct 21, 2016Updated 9 years ago
- ☆14Nov 3, 2025Updated 8 months ago
- ☆10Jul 29, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Multi-heap-sort for many small arrays, quicksort with 3 pivots for one big array, CUDA acceleration, CUDA memory compression.☆13Sep 29, 2024Updated last year
- Linux kernel source tree☆14May 6, 2021Updated 5 years ago
- 《汇编语言一发入魂》配套代码☆15May 30, 2020Updated 6 years ago
- 美赛爬虫,美国大学生数学建模竞赛证书爬取及信息OCR识别分析☆16Jun 25, 2022Updated 4 years ago
- ☆15Jan 25, 2017Updated 9 years ago
- Efficient, Flexible, and Highly Fault-Tolerant Model Service Management Based on SGLang☆62Nov 8, 2024Updated last year
- CUDA project for uni subject☆26Oct 26, 2020Updated 5 years ago