Attention in SRAM on Tenstorrent Grayskull
☆38Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for grayskull-attention
Users that are interested in grayskull-attention are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simple experiments on Tenstorrent GraySkull e75 chip☆14Aug 28, 2024Updated 2 years ago
- TVM for Tenstorrent ASICs☆31Aug 28, 2026Updated last week
- Tenstorrent Firmware repository☆24Feb 25, 2026Updated 6 months ago
- tiny code to access tenstorrent blackhole☆70May 26, 2025Updated last year
- A low level hardware debugger☆21Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Tenstorrent MLIR compiler☆305Updated this week
- Buda Compiler Backend for Tenstorrent devices☆31Apr 2, 2025Updated last year
- Install the tenstorrent stack with one command☆24Updated this week
- The TT-Forge ONNX is a graph compiler designed to optimize and transform computational graphs for deep learning models, enhancing their p…☆65Updated this week
- User-Mode Driver for Tenstorrent hardware☆47Updated this week
- tenstorrent kernel from twitch☆29Mar 16, 2024Updated 2 years ago
- Tenstorrent Topology (TT-Topology) is a command line utility used to flash multiple NB cards on a system to use specific eth routing conf…☆16Jul 23, 2026Updated last month
- [ECCV 2022] SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruning☆20Jul 7, 2022Updated 4 years ago
- Tenstorrent system interface library☆34Updated this week
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Tenstorrent Firmware Update Utility☆13Updated this week
- A simple network-on-chip performance estimator (NPE) for Tenstorrent Tensix-based devices☆15Updated this week
- ☆32Jun 7, 2025Updated last year
- TT-NN operator library, and TT-Metalium low level kernel programming model.☆1,661Updated this week
- [Deprecated] ⭐️ TT-NN Compiler for PyTorch 2 ⭐️ Enables running PyTorch models on Tenstorrent hardware using eager or compile path☆62Aug 28, 2026Updated last week
- My tests and experiments with some popular dl frameworks.☆17Sep 11, 2025Updated 11 months ago
- A comprehensive tool for visualizing and analyzing model execution, offering interactive graphs, memory plots, tensor details, buffer ove…☆56Updated this week
- ☆16Sep 24, 2024Updated last year
- The official code for [ECCV2020] "HALO: Hardware-aware Learning to Optimize"☆10Mar 22, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NeurIPS 2023] ShiftAddViT: Mixture of Multiplication Primitives Towards Efficient Vision Transformer☆31Dec 6, 2023Updated 2 years ago
- ☆10May 30, 2024Updated 2 years ago
- Artifact for PPoPP20 "Understanding and Bridging the Gaps in Current GNN Performance Optimizations"☆42Nov 16, 2021Updated 4 years ago
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆13Nov 23, 2024Updated last year
- [NeurIPS'25] KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems☆18Nov 1, 2025Updated 10 months ago
- The translator that supports translating NVPTX to SPIR-V. This translator is modified from LLVM-SPIR-V Translator.☆45Oct 25, 2021Updated 4 years ago
- It's a baby compiler. (Lean btw.)☆16May 19, 2025Updated last year
- The most comprehensive community resource for the NVIDIA CMP 170HX.☆94Jul 30, 2026Updated last month
- Scripts to prepare OXFORD VGG Face dataset☆12Mar 29, 2016Updated 10 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- An open-sourced PyTorch library for developing energy efficient multiplication-less models and applications.☆14Feb 3, 2025Updated last year
- ☆11Nov 2, 2017Updated 8 years ago
- LibArgus and CUDA interaction Libraries, Samples, and Demos☆13Jan 12, 2023Updated 3 years ago
- convert a saved pytorch model to gguf and generate as much corresponding ggml c code as possible☆15Dec 19, 2023Updated 2 years ago
- Train to 94% on CIFAR-10 in 4.4 seconds on a single A100☆12Dec 30, 2023Updated 2 years ago
- ☆12May 23, 2018Updated 8 years ago
- ☆18Dec 2, 2024Updated last year