A comprehensive benchmarking framework for evaluating and optimizing CPU-centric agentic AI systems across multiple workloads, reproducing results from the research paper: "A CPU-Centric Perspective on Agentic AI"
☆53Feb 12, 2026Updated 6 months ago
Alternatives and similar repositories for cpu-centric-agentic-ai
Users that are interested in cpu-centric-agentic-ai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The set of AI agent model implementations, benchmarks, and others used in our paper "The Cost of Dynamic Reasoning: Demystifying AI Agent…☆43Mar 26, 2026Updated 5 months ago
- [DATE 2025] Official implementation and dataset of AIrchitect v2: Learning the Hardware Accelerator Design Space through Unified Represen…☆21Jan 17, 2025Updated last year
- Agent application/benchmark/workload traces should be placed here.☆15Apr 13, 2026Updated 4 months ago
- [ECCV 2024] CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs☆20Jul 2, 2024Updated 2 years ago
- Codebase for layer wise N:M pruning pattern assignment for LLMs☆15Aug 5, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A user level library for applications to transparently use Intel DSA.☆45Aug 6, 2026Updated 3 weeks ago
- ☆16May 27, 2026Updated 3 months ago
- [ISCA 2025] Official Implementation of "MicroScopiQ: Accelerating Foundational Models through Outlier-Aware Microscaling Quantization"☆25Oct 30, 2025Updated 9 months ago
- ☆72Nov 29, 2025Updated 9 months ago
- ☆18Aug 10, 2026Updated 2 weeks ago
- Preview Code for Continuum Paper☆101Aug 13, 2026Updated 2 weeks ago
- [IEEE CAL 2025] Accelerating Page Migrations in Operating Systems with Intel DSA☆16Nov 20, 2024Updated last year
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆13Nov 23, 2024Updated last year
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆76Mar 11, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆15Updated this week
- This is the respository that holds the artifacts of ASPLOS'25 -- M5: Mastering Page Migration and Memory Management for CXL-based Tiered …☆17Apr 1, 2025Updated last year
- A simple, fast and robust program-aware agentic inference system.☆430Jul 5, 2026Updated last month
- LLM Inference analyzer for different hardware platforms☆124Jul 30, 2026Updated last month
- ☆55Dec 19, 2025Updated 8 months ago
- ☆14Apr 24, 2024Updated 2 years ago
- Artifacts of EuroSys'24 paper "Exploring Performance and Cost Optimization with ASIC-Based CXL Memory"☆31Feb 21, 2024Updated 2 years ago
- Heterogenous ML accelerator☆23May 5, 2025Updated last year
- A lightweight and fast LLM serving framework☆15Mar 5, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A framework for generating realistic LLM serving workloads☆170Updated this week
- TokenSim is a tool for simulating the behavior of large language models (LLMs) in a distributed environment.☆32Updated this week
- ☆18Mar 17, 2023Updated 3 years ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- ☆30May 20, 2026Updated 3 months ago
- code repo for GCR [FAST'26]☆16Mar 3, 2026Updated 5 months ago
- BloodCell-Detector-Yolo is a YOLOv5 implementation tailored to detect Red Blood Cells (RBC), White Blood Cells (WBC), and Platelets from …☆19Jan 23, 2025Updated last year
- Tiered memory management☆91Sep 1, 2025Updated 11 months ago
- ☆16Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [USENIX ATC 2021] Exploring the Design Space of Page Management for Multi-Tiered Memory Systems☆49Mar 31, 2022Updated 4 years ago
- Heterogeneous Memory Software Development Kit☆103Mar 9, 2026Updated 5 months ago
- Tiered Memory Management: Access Latency is the Key!☆64Mar 17, 2025Updated last year
- MAFIA: Multiple Application Framework for GPU architectures☆28Jan 21, 2022Updated 4 years ago
- ☆15Nov 7, 2024Updated last year
- IBM Platform-Independent Software Analysis☆15Mar 12, 2018Updated 8 years ago
- ☆19Feb 9, 2026Updated 6 months ago