Full Transformer into a custom chip. microGPT in RTL, generating names on a Virtex-5 FPGA at ~56k tokens/second.
☆623Jun 25, 2026Updated last month
Alternatives and similar repositories for gateGPT
Users that are interested in gateGPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- hardware implementation of transformers running microgpt at 50k+ tkps☆763May 14, 2026Updated 2 months ago
- Pipeline-parallel LLM inference across GPUs on separate machines.☆432Updated this week
- [DATE'2025, TCAD'2025] Terafly : A Multi-Node FPGA Based Accelerator Design for Efficient Cooperative Inference in LLMs☆38Nov 13, 2025Updated 8 months ago
- hardware accelerator for deep convolutional neural networks☆75Feb 25, 2026Updated 5 months ago
- A self-hosted, zero-knowledge, secure PHP notebook. Your notes, in your browser, on your server - every byte encrypted with AES-256 using…☆26Jun 20, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- microGPT benchmarks: a single M4 Max MacBook Pro P-core in C runs Karpathy's 4192-parameter transformer at ~71x the throughput of TALOS-V…☆163May 4, 2026Updated 2 months ago
- HRM-Text is a 1B text generation model based on the HRM architecture, strengthened by task completion and latent space reasoning.☆1,749Jun 17, 2026Updated last month
- Training neural networks on Apple Neural Engine via reverse-engineered private APIs☆7,056Mar 10, 2026Updated 4 months ago
- Fast LLM speculative inference server for consumer hardware.☆2,689Updated this week
- DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm☆19,356Updated this week
- A cheap multipurpose mobile robotic platform for research, development, education and fun. All the code, data, instructions and files yo…☆16Jul 1, 2026Updated 3 weeks ago
- KV260 integration lane for PCCX™ v002 LLM IP-core bring-up, validation, and board/runtime evidence.☆16Jun 3, 2026Updated last month
- ☆24Updated this week
- RISC-V Superscalar Educational Simulator based on Tomasulo's Algorithm☆36Nov 1, 2025Updated 8 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A collection of examples showcasing PyCDE and Mini RISC-V implementation.☆10Updated this week
- Self-Teaching Autoencoder learning reconstructions through latent agreement, not pixel loss.☆29May 25, 2026Updated 2 months ago
- DFlash: Block Diffusion for Flash Speculative Decoding☆5,543May 10, 2026Updated 2 months ago
- A Top-Down Profiler for GPU Applications☆23Feb 29, 2024Updated 2 years ago
- Using Feature Decomposition method to accelerate GNN inference☆13Sep 27, 2021Updated 4 years ago
- MAVLink Military Messages☆24Feb 19, 2026Updated 5 months ago
- reverse engineering the best-selling drones on Amazon to control programmatically☆527Jun 23, 2026Updated last month
- LoRa chip into a coherent linear-FM chirp generator☆107Apr 24, 2026Updated 3 months ago
- Open-source AI Accelerator Stack integrating compute, memory, and software — from RTL to PyTorch.☆26Jul 2, 2026Updated 3 weeks ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B☆1,500Jun 17, 2026Updated last month
- Native insanely fast terminal built for agents☆73Updated this week
- Ventus GPGPU ISA Simulator Based on Spike☆52Jul 21, 2026Updated last week
- ☆133Jul 12, 2026Updated 2 weeks ago
- A minimal tensor processing unit (TPU), inspired by Google's TPU V2 and V1☆1,355Apr 3, 2026Updated 3 months ago
- A vector index built on TurboQuant, written in Rust with Python bindings☆14,470Updated this week
- Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦☆20,466Updated this week
- Minimal TPU implementation with 8x8 systolic array and PyTorch integration☆66Jan 26, 2026Updated 6 months ago
- Fast and memory-efficient classical machine learning operators☆548Updated this week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A private, open-source alternative to Windows Recall — for Linux. Captures your screen, OCRs it, and makes everything you've seen instant…☆43Jun 24, 2026Updated last month
- ☆21Jul 16, 2026Updated last week
- TokenSpeed is a speed-of-light LLM inference engine.☆1,717Updated this week
- ActiveGraph/GBrain bridge proof of concept for Apprentice launch.☆24May 26, 2026Updated 2 months ago
- MathCode: A Frontier Mathematical Coding Agent☆582Jun 15, 2026Updated last month
- ☆7,003Jul 20, 2026Updated last week
- Use amaranth-to-litex to simply import Amaranth code into a Litex project.☆15Apr 22, 2024Updated 2 years ago