☆15Jan 22, 2026Updated 5 months ago
Alternatives and similar repositories for gpu-programming
Users that are interested in gpu-programming are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- qwen3 experiments☆33Jul 1, 2025Updated last year
- A Bigram Language Model from scratch with no-smoothing and add-one smoothing. Outputs bigram counts, bigram probabilities and probability…☆14Jan 12, 2021Updated 5 years ago
- As barebones as you can get the GPT, now accelerated on Macs thanks to tinygrad.☆16Oct 22, 2025Updated 8 months ago
- Simple netcat wrote in C☆16Jun 25, 2025Updated last year
- Byte-Pair Encoding (BPE) (subword-based tokenization) algorithm implementaions from scratch with python☆19Jan 30, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆16Jul 7, 2025Updated last year
- learningggggggg 🐳☆634Apr 2, 2025Updated last year
- we have ai at home☆117Jun 18, 2026Updated last month
- Implementation for IceCache: Memory-Efficient KV-cache Management for Long-Sequence LLMs (ICLR 2026).☆19Jun 9, 2026Updated last month
- Learn how Transformer models are implemented from scratch.☆22Jun 3, 2024Updated 2 years ago
- Carla worldsim envrionment for decision-making evaluation and RL☆19Apr 22, 2026Updated 2 months ago
- image captioninggg🐳☆12Aug 30, 2024Updated last year
- "My instructor was Mr. Langley, and he taught me to sing a song. If you'd like to hear it I can sing it for you..."☆10Sep 27, 2015Updated 10 years ago
- Whalegrad 🐳 is a lightweight deep learning library written in C.☆10Jan 5, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Mention any three favourite things and get recommendations in the form of a flow chart by Claude Haiku.☆14Apr 6, 2024Updated 2 years ago
- .NET library to control the MCP2221/MCP2221A, providing easy access to GPIO, DAC, ADC, and I2C functionalities using only a USB port☆15Updated this week
- An efficient and scalable attention module designed to reduce memory usage and improve inference speed in large language models. Designe…☆24Jun 25, 2025Updated last year
- ☆15Jan 26, 2025Updated last year
- NanoGPT-speedrunning for the poor T4 enjoyers☆72Apr 22, 2025Updated last year
- ☆32Jul 14, 2024Updated 2 years ago
- Inference Llama/Llama2/Llama3 Modes in NumPy☆21Nov 22, 2023Updated 2 years ago
- ☆22Jul 6, 2026Updated 2 weeks ago
- The bit-banging version of i2c for STM32F4 with possibility to create many i2c ports (limited by GPIO number). It supports clock stretchi…☆14Dec 6, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a small autograd engine, made purely from numpy and python.☆27Sep 17, 2024Updated last year
- Code companion for the RL Post-Training Handbook - training reasoning models on a single GPU☆19Jan 30, 2026Updated 5 months ago
- ☆44May 4, 2025Updated last year
- Leanstral's fork of SafeVerify, which we use for code agent training and as part of our evaluation stack.☆29Jul 3, 2026Updated 2 weeks ago
- machine learning from absolute scratch in c. gradients, linear algebra ops & everything else without using any third party library!☆26Aug 3, 2024Updated last year
- ☆93Dec 16, 2025Updated 7 months ago
- Fast Fourier Transform Frontend☆13Dec 11, 2013Updated 12 years ago
- ☆132Dec 9, 2025Updated 7 months ago
- a simple c++ inference engine for gpt based architecture☆40Dec 10, 2025Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- documenting my work in inference engineering☆25Apr 19, 2026Updated 3 months ago
- Landing repository for the paper "Softpick: No Attention Sink, No Massive Activations with Rectified Softmax"☆91Sep 12, 2025Updated 10 months ago
- WhaleSearch 🐳 - semantic search engine. (SSE)☆16Jan 4, 2025Updated last year
- Code for "Reversal Q-Learning (RQL)" for Flow RL from Prior Data☆32Jun 17, 2026Updated last month
- An opinionated extensible language for rule creation! (UNDER DEVELOPMENT)☆13Jul 12, 2026Updated last week
- 100 Days Of GPU Programming.☆45Nov 7, 2025Updated 8 months ago
- Protobuf with pizzazz☆20Sep 9, 2024Updated last year