☆385May 4, 2026Updated 3 months ago
Alternatives and similar repositories for nanowhale
Users that are interested in nanowhale are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Modal-style sandbox API on top of Hugging Face Jobs☆159Jul 6, 2026Updated last month
- Fast and memory-efficient classical machine learning operators☆555Updated this week
- Cuda kernels for leveraging LLM sparsity to improve throughput and decrease the memory requirements during inference and training.☆256Jun 29, 2026Updated last month
- 🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.☆3,718Updated this week
- TokenSpeed is a speed-of-light LLM inference engine.☆1,836Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- an implementation of paper"Retentive Network: A Successor to Transformer for Large Language Models" https://arxiv.org/pdf/2307.08621.pdf☆11Jul 25, 2023Updated 3 years ago
- DFlash: Block Diffusion for Flash Speculative Decoding☆5,579May 10, 2026Updated 3 months ago
- An unbounded n-gram language model on Tiny Shakespeare☆22Jan 21, 2026Updated 6 months ago
- The best ChatGPT that $100 can buy.☆58Updated this week
- Official code for HiLS-Attention☆138Jul 25, 2026Updated 2 weeks ago
- FlashKDA: high-performance Kimi Delta Attention kernels☆1,193Jul 30, 2026Updated last week
- HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness☆129May 9, 2026Updated 3 months ago
- Simple repository for training small reasoning models☆51Updated this week
- Everything about the SmolLM and SmolVLM family of models☆3,867May 26, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Nano vLLM☆14,923Apr 26, 2026Updated 3 months ago
- ☆25May 14, 2026Updated 2 months ago
- Tiny-R2: A hybrid architecture integrating SWA, CSA, HCA, mHC, and DSMoE under the DeepSeek V4 design paradigm, enabling single-GPU OPD p…☆46May 30, 2026Updated 2 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,820Updated this week
- 🎓Automatically Update circult-eda-mlsys-tinyml Papers Daily using Github Actions (Update Every 8th hours)☆10Aug 3, 2026Updated last week
- an open source reproduction of NVIDIA's nGPT (Normalized Transformer with Representation Learning on the Hypersphere)☆112Mar 7, 2025Updated last year
- ☆25Jan 28, 2026Updated 6 months ago
- Hugging Face's take home challenge for post-training internships, now open for you to try!☆49Apr 23, 2026Updated 3 months ago
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆4,982Oct 27, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,274Aug 26, 2025Updated 11 months ago
- A benchmark for evaluating contextual agents on realistic multimodal personal-computer environments with profiling and factual-retention …☆31Apr 2, 2026Updated 4 months ago
- 🎙️ 「大模型」从0训练0.1B能听能说能看的全模态Omni模型!A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing!☆2,298Updated this week
- A PyTorch implementation of the GPT-OSS-20B architecture. All components are coded from scratch: RoPE with YaRN, RMSNorm, SwiGLU with cla…☆238Dec 2, 2025Updated 8 months ago
- Extending the Context of Pretrained LLMs by Dropping Their Positional Embedding☆220Jan 12, 2026Updated 6 months ago
- A kernel library written in tilelang☆1,712Apr 23, 2026Updated 3 months ago
- Official Project Page for HLA: Higher-order Linear Attention (https://arxiv.org/abs/2510.27258)☆102Jun 15, 2026Updated last month
- Minimalistic large language model 3D-parallelism training☆2,779May 26, 2026Updated 2 months ago
- A Python SDK for Open Reward Standard servers and clients☆17Mar 24, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆255Nov 24, 2025Updated 8 months ago
- NanoGPT (124M) in 90 seconds☆5,654Aug 2, 2026Updated last week
- ☆402Jul 30, 2026Updated last week
- Hand-Rolled GPU communications library☆96Nov 25, 2025Updated 8 months ago
- DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms☆6,911Jul 9, 2026Updated last month
- Open source framework for evaluating AI Agents☆32Feb 24, 2026Updated 5 months ago
- ☆16Jul 30, 2026Updated last week