Elana: A Simple Energy & Latency Analyzer for LLMs
☆16Apr 3, 2026Updated 3 months ago
Alternatives and similar repositories for Elana
Users that are interested in Elana are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- UniQL official repository (ICLR 2026)☆16Jan 27, 2026Updated 5 months ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- The official repository of Quamba1 [ICLR 2025] & Quamba2 [ICML 2025]☆70Jun 19, 2025Updated last year
- Sequential Monte Carlo Speculative Decoding☆52Updated this week
- Model LLM inference on single-core dataflow accelerators☆19Dec 16, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025] Speculate Deep and Accurate☆21Jan 16, 2026Updated 6 months ago
- Accelerator Zoo☆21Oct 14, 2025Updated 9 months ago
- xKV: Cross-Layer SVD for KV-Cache Compression [ICML 2026]☆53Jul 7, 2026Updated 2 weeks ago
- Official implementation of EMNLP'23 paper "Revisiting Block-based Quantisation: What is Important for Sub-8-bit LLM Inference?"☆24Oct 25, 2023Updated 2 years ago
- H2-LLM: Hardware-Dataflow Co-Exploration for Heterogeneous Hybrid-Bonding-based Low-Batch LLM Inference☆113Apr 26, 2025Updated last year
- ☆18Nov 10, 2025Updated 8 months ago
- ☆14Nov 3, 2025Updated 8 months ago
- Fast, memory-efficient attention column reduction (e.g., sum, mean, max)☆49Feb 10, 2026Updated 5 months ago
- ☆27May 12, 2026Updated 2 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- PIM-DL: Expanding the Applicability of Commodity DRAM-PIMs for Deep Learning via Algorithm-System Co-Optimization☆37Feb 21, 2024Updated 2 years ago
- ☆49Apr 22, 2021Updated 5 years ago
- sozlukus.com source code☆13Feb 6, 2017Updated 9 years ago
- MICRO 2023 Evaluation Artifact for TeAAL☆11Oct 26, 2023Updated 2 years ago
- This repository contains the hardware implementation for Static BFP convolution on FPGA☆10Oct 15, 2019Updated 6 years ago
- ☆13Mar 19, 2025Updated last year
- ☆10Dec 26, 2023Updated 2 years ago
- HierCGRA: An Open-Source Framework for Large-Scale CGRA with Hierarchical Modeling and Automated Exploration☆14Mar 6, 2023Updated 3 years ago
- Processing-in Memory Architecture for Multiply-Accumulate Operations with Hybrid Memory Cube☆12Feb 13, 2017Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A PYNQ overlay demonstrating the Xilinx RFSoC SD-FEC☆13Jun 29, 2022Updated 4 years ago
- [HPCA'21] SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning☆136Aug 27, 2024Updated last year
- Framework based on Partial Reconfiguration for chip characterization utilizing ring-oscillator PUFs☆14Apr 1, 2020Updated 6 years ago
- Program to scan for malicious FPGA designs.☆17Mar 20, 2021Updated 5 years ago
- ☆11Jan 13, 2026Updated 6 months ago
- ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization☆114Oct 15, 2024Updated last year
- A fork of llama3.c used to do some R&D on inferencing☆23Dec 20, 2024Updated last year
- the xoroshiro32++ and xoroshiro64++ PRNG algorthims by David Blackman and Sebastiano Vigna in C++, Verilog, VHDL and SpinalHDL.☆16Dec 2, 2018Updated 7 years ago
- Replace original DRAM model in GPGPU-sim with Ramulator DRAM model☆22Dec 10, 2018Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- C++ library for AXI DMA with direct and scatter-gather support☆13Feb 22, 2022Updated 4 years ago
- This document adopts the method from the XAPP1230 for doing readback capture on Xilinx UltraScale devices and shows how to migrate the sa…☆18Nov 15, 2019Updated 6 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- The code for paper: Neuralpower: Predict and deploy energy-efficient convolutional neural networks☆24Jul 10, 2019Updated 7 years ago
- Official mirror of torchjpeg. Please do not open PRs here, they will be ignored. Go to the gitlab repository to contribute.☆22Jun 21, 2023Updated 3 years ago
- Real-Time Hardware Sorter, Using A Multi-Dimensional Sorting Algorithm☆20Jan 4, 2026Updated 6 months ago
- ☆14Sep 7, 2024Updated last year