☆27Jun 7, 2026Updated 2 months ago
Alternatives and similar repositories for nanotok
Users that are interested in nanotok are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- a minimal paged attention implementation☆20Jan 30, 2026Updated 6 months ago
- DeepSeek 4 Flash local inference engine for Metal and CUDA with M5 optimizations.☆22May 24, 2026Updated 2 months ago
- AI-Powered Thesis Review Tool☆17Aug 8, 2025Updated last year
- Triton‑style kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Silicon☆48Mar 31, 2026Updated 4 months ago
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lets build a Deep Learning Framework!☆28Mar 12, 2026Updated 4 months ago
- So, I trained a Llama a 130M architecture I coded from ground up to build a small instruct model from scratch. Trained on FineWeb dataset…☆18Mar 26, 2025Updated last year
- An experiment in turning years of machine learning experience into a research loop that could run on its own☆53Apr 9, 2026Updated 4 months ago
- KDSS is the framework for knowledge distillation from LLMs☆12Nov 5, 2025Updated 9 months ago
- ☆15Dec 4, 2025Updated 8 months ago
- Example for a Monty-enabled RLM in DSPy☆20Feb 16, 2026Updated 5 months ago
- A series of high-performance GEMM (General Matrix Multiply) implementations Iteratively optimised for H100 GPUs in Pure CUDA.☆81Feb 18, 2026Updated 5 months ago
- MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆27May 23, 2026Updated 2 months ago
- ☆21Apr 21, 2026Updated 3 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Code companion for the RL Post-Training Handbook - training reasoning models on a single GPU☆19Jan 30, 2026Updated 6 months ago
- Code for the paper "Getting the most out of your tokenizer for pre-training and domain adaptation"☆22Feb 14, 2024Updated 2 years ago
- Generates text with diffusion models. Reproduction of the Continous Diffusion for Categorical Data paper by Deepmind☆18Dec 9, 2024Updated last year
- Low memory full parameter finetuning of LLMs☆54Jul 18, 2025Updated last year
- XoRL☆18Updated this week
- Implementation of CoLA: Compute-Efficient Pre-Training of LLMs via Low-Rank Activation☆26Feb 18, 2025Updated last year
- A set of markdown files to point Claude towards to get an amazing mandarin tutor☆16May 23, 2026Updated 2 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 10 months ago
- JAX implementation of configurable LLM distillation training☆24Nov 15, 2025Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 3 months ago
- ☆28Jul 2, 2026Updated last month
- just UI to explain product manager what dspy program does☆27Oct 26, 2025Updated 9 months ago
- A handy plugin for copying requests/responses directly from Burp, some extra magic included.☆13Oct 15, 2021Updated 4 years ago
- A curated list of resources dedicated to Code-mixed Natural Language Processing (NLP).☆15Jun 23, 2026Updated last month
- ☆10Jul 28, 2021Updated 5 years ago
- ☆66Feb 14, 2026Updated 5 months ago
- A modular framework for training and inference of (compressed) multi-vector retrieval across any modality.☆22Apr 4, 2026Updated 4 months ago
- A Difficulty-Calibrated Benchmark for Building Terminal Agents☆30Feb 20, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆81Feb 18, 2026Updated 5 months ago
- Self extensible agent harness☆22Jul 17, 2026Updated 3 weeks ago
- NanoGPT-speedrunning for the poor T4 enjoyers☆72Apr 22, 2025Updated last year
- ☆22Dec 24, 2025Updated 7 months ago
- An introduction to DSPy☆33Aug 30, 2025Updated 11 months ago
- Companion code for The Physics of LLM Inference book☆26Apr 21, 2026Updated 3 months ago
- Open source framework for evaluating AI Agents☆32Feb 24, 2026Updated 5 months ago