Elana: A Simple Energy & Latency Analyzer for LLMs
☆16Apr 3, 2026Updated 4 months ago
Alternatives and similar repositories for Elana
Users that are interested in Elana are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- UniQL official repository (ICLR 2026)☆17Jan 27, 2026Updated 6 months ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- The official repository of Quamba1 [ICLR 2025] & Quamba2 [ICML 2025]☆70Jun 19, 2025Updated last year
- Sequential Monte Carlo Speculative Decoding☆52Jul 25, 2026Updated 2 weeks ago
- [NeurIPS 2025] Speculate Deep and Accurate☆23Jan 16, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Model LLM inference on single-core dataflow accelerators☆20Dec 16, 2025Updated 7 months ago
- Accelerator Zoo☆21Oct 14, 2025Updated 9 months ago
- This project provides provisioned HPC cluster models using underlying virtualization mechanisms.☆11Sep 10, 2022Updated 3 years ago
- Official implementation of EMNLP'23 paper "Revisiting Block-based Quantisation: What is Important for Sub-8-bit LLM Inference?"☆24Oct 25, 2023Updated 2 years ago
- H2-LLM: Hardware-Dataflow Co-Exploration for Heterogeneous Hybrid-Bonding-based Low-Batch LLM Inference☆115Apr 26, 2025Updated last year
- ☆14Nov 3, 2025Updated 9 months ago
- Fast, memory-efficient attention column reduction (e.g., sum, mean, max)☆49Feb 10, 2026Updated 6 months ago
- ☆27May 12, 2026Updated 2 months ago
- PIM-DL: Expanding the Applicability of Commodity DRAM-PIMs for Deep Learning via Algorithm-System Co-Optimization☆37Feb 21, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- sozlukus.com source code☆13Feb 6, 2017Updated 9 years ago
- ☆49Apr 22, 2021Updated 5 years ago
- MICRO 2023 Evaluation Artifact for TeAAL☆11Oct 26, 2023Updated 2 years ago
- This repository contains the hardware implementation for Static BFP convolution on FPGA☆10Oct 15, 2019Updated 6 years ago
- ☆10Dec 26, 2023Updated 2 years ago
- Processing-in Memory Architecture for Multiply-Accumulate Operations with Hybrid Memory Cube☆12Feb 13, 2017Updated 9 years ago
- A PYNQ overlay demonstrating the Xilinx RFSoC SD-FEC☆13Jun 29, 2022Updated 4 years ago
- Framework based on Partial Reconfiguration for chip characterization utilizing ring-oscillator PUFs☆14Apr 1, 2020Updated 6 years ago
- Program to scan for malicious FPGA designs.☆17Mar 20, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆11Jan 13, 2026Updated 6 months ago
- ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization☆114Oct 15, 2024Updated last year
- A fork of llama3.c used to do some R&D on inferencing☆23Dec 20, 2024Updated last year
- ☆14Mar 19, 2025Updated last year
- the xoroshiro32++ and xoroshiro64++ PRNG algorthims by David Blackman and Sebastiano Vigna in C++, Verilog, VHDL and SpinalHDL.☆16Dec 2, 2018Updated 7 years ago
- Replace original DRAM model in GPGPU-sim with Ramulator DRAM model☆22Dec 10, 2018Updated 7 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- This document adopts the method from the XAPP1230 for doing readback capture on Xilinx UltraScale devices and shows how to migrate the sa…☆18Nov 15, 2019Updated 6 years ago
- ☆14Sep 7, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Source code for XPGraph-MICRO22☆11Apr 10, 2023Updated 3 years ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆117Dec 2, 2025Updated 8 months ago
- This repository contains the official implementation for the ECCV'22 paper, "SPIN: An Empirical Evaluation on Sharing Parameters of Isotr…☆20Sep 9, 2023Updated 2 years ago
- ☆65Apr 26, 2025Updated last year
- Multi-task end-to-end predict-then-optimize☆14Apr 28, 2023Updated 3 years ago
- An open platform for exploring scale-up network systems.☆20Mar 16, 2026Updated 4 months ago
- [HPCA'21] SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning☆137Aug 27, 2024Updated last year