ArcLight: A Lightweight LLM Inference Framework
☆47May 30, 2026Updated 2 months ago
Alternatives and similar repositories for ArcLight
Users that are interested in ArcLight are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Supplemental materials for The ASPLOS 2025 / EuroSys 2025 Contest on Intra-Operator Parallelism for Distributed Deep Learning☆25May 12, 2025Updated last year
- ThinK: Thinner Key Cache by Query-Driven Pruning☆30Jun 2, 2026Updated 2 months ago
- ☆14Oct 3, 2024Updated last year
- ☆16Jun 15, 2026Updated last month
- 七夕孤寡助手☆13Aug 7, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Sep 19, 2024Updated last year
- FPGA Innovation Design Competition:RISC-V Processor-based Hardware and Software Design in PGL22G☆12Sep 1, 2023Updated 2 years ago
- TPAMI 2025 Survey Paper☆34Mar 31, 2025Updated last year
- We release Open Meditron, a fully open, clinician-audited medical training corpus and evaluation protocol that closes the open-vs-closed …☆15Aug 3, 2026Updated last week
- Tempo is a system for declarative, efficient, end-to-end compiled dynamic deep learning☆30Oct 21, 2025Updated 9 months ago
- CODO: An Automated Compiler for Comprehensive Dataflow Optimization☆36Jun 3, 2026Updated 2 months ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated last year
- 🫧 Code for Holistic Reasoning with Long-Context LMs: A Benchmark for Database Operations on Massive Textual Data (Maekawa*, Iso* et al.…☆12Feb 25, 2025Updated last year
- A tool allowing students of Coursera's Heterogeneous Parallel Programming to work on homework using a machine without a CUDA GPU.☆11Mar 11, 2015Updated 11 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ASPLOS'24: Optimal Kernel Orchestration for Tensor Programs with Korch☆41Mar 27, 2025Updated last year
- ☆33Mar 12, 2026Updated 4 months ago
- ☆53May 19, 2025Updated last year
- 小彭老师推出 SyCL 2020 课程(施工中,日后会在直播中放出)☆15Sep 3, 2023Updated 2 years ago
- Source code and dataset for COLING2018 paper "Adversarial Multi-lingual Neural Relation Extraction".☆21May 20, 2022Updated 4 years ago
- Some CS notes during Jiawei's undergrad.☆33Jan 6, 2022Updated 4 years ago
- Official repository for Parallax (Parameterized Local Linear Attention)☆68Jul 30, 2026Updated last week
- SC 2021, "LogECMem: Coupling Erasure-Coded In-Memory Key-Value Stores with Parity Logging"☆12Jul 12, 2021Updated 5 years ago
- Tile-based language built for AI computation across all scales☆184Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai☆176Updated this week
- [ICPP'25] TD-Pipe: Temporally-Disaggregated Pipeline Parallelism Architecture for High-Throughput LLM Inference☆52Dec 24, 2025Updated 7 months ago
- Terminal tool to share your terminal history with anyone.☆16Feb 14, 2026Updated 5 months ago
- study of Ampere' Sparse Matmul☆18Jan 10, 2021Updated 5 years ago
- ☆18Apr 21, 2024Updated 2 years ago
- TritonBench: Benchmarking Large Language Model Capabilities for Generating Triton Operators☆137Jun 14, 2025Updated last year
- ☆125May 19, 2025Updated last year
- [ACL 2025 main] FR-Spec: Frequency-Ranked Speculative Sampling☆55Jul 15, 2025Updated last year
- a cloud-native workflow engine, also known as KubeAdaptor, a docking framework able to implement workflow containerization on Kubernetes…☆13Apr 21, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆14Jun 23, 2025Updated last year
- ☆24Jan 18, 2026Updated 6 months ago
- Notes of computer science courses☆28Aug 2, 2020Updated 6 years ago
- ☆24May 2, 2026Updated 3 months ago
- Large Language Model (LLM) Serving Paper and Resource List☆29Jul 16, 2026Updated 3 weeks ago
- ☆16Oct 13, 2023Updated 2 years ago
- ☆13Aug 1, 2025Updated last year