A minimal Tensor Processing Unit (TPU) inspired by Google's TPUv1.
☆205Aug 10, 2024Updated 2 years ago
Alternatives and similar repositories for tiny-tpu-old
Users that are interested in tiny-tpu-old are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- minimal compiler☆24Feb 19, 2026Updated 6 months ago
- A minimal tensor processing unit (TPU), inspired by Google's TPU V2 and V1☆1,373Apr 3, 2026Updated 4 months ago
- Anatomy of a powerhouse: SystemVerilog TPU based on Google TPU v1☆25Nov 9, 2025Updated 9 months ago
- Used FPGA board and System Verilog to design controller, DMA, pipelined SIMD processor, and GEMM accelerator☆13Aug 26, 2023Updated 3 years ago
- Template for project1 TPU☆23May 1, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- a highly efficient compression algorithm for the n1 implant (neuralink's compression challenge)☆47Jun 3, 2024Updated 2 years ago
- Small-scale Tensor Processing Unit built on an FPGA☆230Aug 4, 2019Updated 7 years ago
- A configurable general purpose graphics processing unit for☆12May 18, 2019Updated 7 years ago
- Systolic array based simple TPU for CNN on PYNQ-Z2☆52Jun 24, 2022Updated 4 years ago
- Hardware acceleration for transformer attention mechanisms on NVIDIA Deep Learning Accelerator (NVDLA), enabling efficient inference of…☆20Mar 30, 2025Updated last year
- ☆18Jan 8, 2023Updated 3 years ago
- Rust Primitives, Learnings, & Frameworks☆17Mar 29, 2022Updated 4 years ago
- RISC-V vector and tensor compute extensions for Vortex GPGPU acceleration for ML workloads. Optimized for transformer models, CNNs, and g…☆25Apr 25, 2025Updated last year
- IC implementation of Systolic Array for TPU☆370Oct 21, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆11Jun 28, 2020Updated 6 years ago
- a mini 2x2 systolic array and PE demo☆75Dec 21, 2025Updated 8 months ago
- A tiniest ASIC GPU that can render only two texture mapped triangles☆30Jan 2, 2026Updated 8 months ago
- Get a hack club dino☆12Mar 13, 2023Updated 3 years ago
- ☆12Oct 6, 2023Updated 2 years ago
- [HPCA 2026 Best Paper Candidate] Official implementation of "Focus: A Streaming Concentration Architecture for Efficient Vision-Language …☆62Feb 8, 2026Updated 6 months ago
- Learn and build GPU RTL from scratch☆22Aug 1, 2025Updated last year
- SystemVerilog Implementations of CUDA/TensorCore/TPU GEMM Operations☆21Apr 12, 2026Updated 4 months ago
- A minimal GPU design in Verilog to learn how GPUs work from the ground up☆12,901Aug 18, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A small Neural Network Processor for Edge devices.☆20Nov 22, 2022Updated 3 years ago
- minimal diffusion transformer in pytorch.☆17Oct 6, 2024Updated last year
- I like to learn new things☆12Feb 28, 2026Updated 6 months ago
- ☆174Jan 4, 2026Updated 7 months ago
- Formal Verification of RISC V IM Processor☆11Mar 27, 2022Updated 4 years ago
- tinyGPU: A Predicated-SIMD processor implementation in SystemVerilog☆67Jul 14, 2021Updated 5 years ago
- ☆22Jun 16, 2025Updated last year
- An AI accelerator implementation with Xilinx FPGA☆123Updated this week
- Open-source AI Accelerator Stack integrating compute, memory, and software — from RTL to PyTorch.☆28Aug 7, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 10 Gigabit Ethernet MAC Core UVM Verification☆19Oct 5, 2023Updated 2 years ago
- Matrix multiplication accelerator on ZYNQ SoC.☆13Apr 29, 2025Updated last year
- ☆86Updated this week
- Transactional Verilog design and Verilator Testbench for a RISC-V TensorCore Vector co-processor for reproducible linear algebra☆67Dec 19, 2021Updated 4 years ago
- A cross-platform monotonic clock that is suspend-unaware, written in Rust!☆81May 14, 2025Updated last year
- A simple and efficient wrapper around the OpenAI API☆29Aug 5, 2024Updated 2 years ago
- ☆31Feb 22, 2024Updated 2 years ago