Elana: A Simple Energy & Latency Analyzer for LLMs
☆16Apr 3, 2026Updated 5 months ago
Alternatives and similar repositories for Elana
Users that are interested in Elana are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- UniQL official repository (ICLR 2026)☆19Jan 27, 2026Updated 7 months ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- The official repository of Quamba1 [ICLR 2025] & Quamba2 [ICML 2025]☆70Jun 19, 2025Updated last year
- Model LLM inference on single-core dataflow accelerators☆22Dec 16, 2025Updated 9 months ago
- [NeurIPS 2025] Speculate Deep and Accurate☆25Aug 10, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Accelerator Zoo☆21Oct 14, 2025Updated 11 months ago
- xKV: Cross-Layer SVD for KV-Cache Compression [ICML 2026]☆58Jul 7, 2026Updated 2 months ago
- Official implementation of EMNLP'23 paper "Revisiting Block-based Quantisation: What is Important for Sub-8-bit LLM Inference?"☆24Oct 25, 2023Updated 2 years ago
- H2-LLM: Hardware-Dataflow Co-Exploration for Heterogeneous Hybrid-Bonding-based Low-Batch LLM Inference☆118Apr 26, 2025Updated last year
- ☆18Nov 10, 2025Updated 10 months ago
- ☆14Nov 3, 2025Updated 10 months ago
- Fast, memory-efficient attention column reduction (e.g., sum, mean, max)☆50Feb 10, 2026Updated 7 months ago
- ☆27May 12, 2026Updated 4 months ago
- PIM-DL: Expanding the Applicability of Commodity DRAM-PIMs for Deep Learning via Algorithm-System Co-Optimization☆37Feb 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- sozlukus.com source code☆13Feb 6, 2017Updated 9 years ago
- ☆49Apr 22, 2021Updated 5 years ago
- MICRO 2023 Evaluation Artifact for TeAAL☆13Oct 26, 2023Updated 2 years ago
- ☆10Dec 26, 2023Updated 2 years ago
- Processing-in Memory Architecture for Multiply-Accumulate Operations with Hybrid Memory Cube☆12Feb 13, 2017Updated 9 years ago
- HierCGRA: An Open-Source Framework for Large-Scale CGRA with Hierarchical Modeling and Automated Exploration☆15Mar 6, 2023Updated 3 years ago
- Program to scan for malicious FPGA designs.☆17Mar 20, 2021Updated 5 years ago
- ☆13Jan 13, 2026Updated 8 months ago
- ☆15Mar 19, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- the xoroshiro32++ and xoroshiro64++ PRNG algorthims by David Blackman and Sebastiano Vigna in C++, Verilog, VHDL and SpinalHDL.☆16Dec 2, 2018Updated 7 years ago
- Replace original DRAM model in GPGPU-sim with Ramulator DRAM model☆22Dec 10, 2018Updated 7 years ago
- C++ library for AXI DMA with direct and scatter-gather support☆13Feb 22, 2022Updated 4 years ago
- The code for paper: Neuralpower: Predict and deploy energy-efficient convolutional neural networks☆24Jul 10, 2019Updated 7 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- Official mirror of torchjpeg. Please do not open PRs here, they will be ignored. Go to the gitlab repository to contribute.☆23Jun 21, 2023Updated 3 years ago
- ☆14Sep 7, 2024Updated 2 years ago
- ☆15Nov 11, 2025Updated 10 months ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆120Dec 2, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains the official implementation for the ECCV'22 paper, "SPIN: An Empirical Evaluation on Sharing Parameters of Isotr…☆20Sep 11, 2026Updated last week
- ☆65Apr 26, 2025Updated last year
- Multi-task end-to-end predict-then-optimize☆15Apr 28, 2023Updated 3 years ago
- RPCNIC: A High-Performance and Reconfigurable PCIe-attached RPC Accelerator [HPCA2025]☆16Dec 9, 2024Updated last year
- An open platform for exploring scale-up network systems.☆23Mar 16, 2026Updated 6 months ago
- [HPCA'21] SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning☆141Aug 27, 2024Updated 2 years ago
- [ICLR 2025] Palu: Compressing KV-Cache with Low-Rank Projection☆163Feb 20, 2025Updated last year