Elana: A Simple Energy & Latency Analyzer for LLMs
☆16Apr 3, 2026Updated 4 months ago
Alternatives and similar repositories for Elana
Users that are interested in Elana are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- UniQL official repository (ICLR 2026)☆19Jan 27, 2026Updated 7 months ago
- The official repository of Quamba1 [ICLR 2025] & Quamba2 [ICML 2025]☆70Jun 19, 2025Updated last year
- Sequential Monte Carlo Speculative Decoding☆52Updated this week
- Model LLM inference on single-core dataflow accelerators☆22Dec 16, 2025Updated 8 months ago
- Accelerator Zoo☆21Oct 14, 2025Updated 10 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- xKV: Cross-Layer SVD for KV-Cache Compression [ICML 2026]☆57Jul 7, 2026Updated last month
- This project provides provisioned HPC cluster models using underlying virtualization mechanisms.☆11Sep 10, 2022Updated 3 years ago
- Official implementation of EMNLP'23 paper "Revisiting Block-based Quantisation: What is Important for Sub-8-bit LLM Inference?"☆24Oct 25, 2023Updated 2 years ago
- ☆18Nov 10, 2025Updated 9 months ago
- ☆14Nov 3, 2025Updated 9 months ago
- Fast, memory-efficient attention column reduction (e.g., sum, mean, max)☆51Feb 10, 2026Updated 6 months ago
- ☆27May 12, 2026Updated 3 months ago
- PIM-DL: Expanding the Applicability of Commodity DRAM-PIMs for Deep Learning via Algorithm-System Co-Optimization☆37Feb 21, 2024Updated 2 years ago
- sozlukus.com source code☆13Feb 6, 2017Updated 9 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆49Apr 22, 2021Updated 5 years ago
- MICRO 2023 Evaluation Artifact for TeAAL☆12Oct 26, 2023Updated 2 years ago
- This repository contains the hardware implementation for Static BFP convolution on FPGA☆10Oct 15, 2019Updated 6 years ago
- Processing-in Memory Architecture for Multiply-Accumulate Operations with Hybrid Memory Cube☆12Feb 13, 2017Updated 9 years ago
- A PYNQ overlay demonstrating the Xilinx RFSoC SD-FEC☆13Jun 29, 2022Updated 4 years ago
- HierCGRA: An Open-Source Framework for Large-Scale CGRA with Hierarchical Modeling and Automated Exploration☆15Mar 6, 2023Updated 3 years ago
- Framework based on Partial Reconfiguration for chip characterization utilizing ring-oscillator PUFs☆14Apr 1, 2020Updated 6 years ago
- Program to scan for malicious FPGA designs.☆17Mar 20, 2021Updated 5 years ago
- ☆13Jan 13, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization☆114Oct 15, 2024Updated last year
- A fork of llama3.c used to do some R&D on inferencing☆23Dec 20, 2024Updated last year
- ☆14Mar 19, 2025Updated last year
- the xoroshiro32++ and xoroshiro64++ PRNG algorthims by David Blackman and Sebastiano Vigna in C++, Verilog, VHDL and SpinalHDL.☆16Dec 2, 2018Updated 7 years ago
- C++ library for AXI DMA with direct and scatter-gather support☆13Feb 22, 2022Updated 4 years ago
- The code for paper: Neuralpower: Predict and deploy energy-efficient convolutional neural networks☆24Jul 10, 2019Updated 7 years ago
- This document adopts the method from the XAPP1230 for doing readback capture on Xilinx UltraScale devices and shows how to migrate the sa…☆18Nov 15, 2019Updated 6 years ago
- ☆14Sep 7, 2024Updated last year
- ☆15Nov 11, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Source code for XPGraph-MICRO22☆11Apr 10, 2023Updated 3 years ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆119Dec 2, 2025Updated 8 months ago
- This repository contains the official implementation for the ECCV'22 paper, "SPIN: An Empirical Evaluation on Sharing Parameters of Isotr…☆20Sep 9, 2023Updated 2 years ago
- ☆65Apr 26, 2025Updated last year
- Multi-task end-to-end predict-then-optimize☆15Apr 28, 2023Updated 3 years ago
- RPCNIC: A High-Performance and Reconfigurable PCIe-attached RPC Accelerator [HPCA2025]☆16Dec 9, 2024Updated last year
- An open platform for exploring scale-up network systems.☆22Mar 16, 2026Updated 5 months ago