MACKO: Sparse matrix vector multiplication for low sparsity
☆41Apr 6, 2026Updated 3 months ago
Alternatives and similar repositories for macko_spmv
Users that are interested in macko_spmv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Nanos klib for NVIDIA GPUs☆14Apr 12, 2026Updated 3 months ago
- Kolosal AI is an OpenSource and Lightweight alternative to Ollama to run LLMs 100% offline on your device.☆15Jan 2, 2026Updated 6 months ago
- A Python implementation of a sum-product network with gaussian processes leafs model (SPNGP, arXiv:1809.04400)☆19Jun 29, 2023Updated 3 years ago
- ☆19Aug 26, 2025Updated 11 months ago
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆18Jul 1, 2025Updated last year
- ☆17May 26, 2026Updated 2 months ago
- xKV: Cross-Layer SVD for KV-Cache Compression [ICML 2026]☆54Jul 7, 2026Updated 3 weeks ago
- Latex template for poster☆12Sep 6, 2023Updated 2 years ago
- Run ops on Apple ANE in NPU register with pure python on M1 Asahi Linux. No Espresso, No CoreML, no metal, no .mlmodels file, no .hwx fil…☆16Jun 28, 2026Updated last month
- ☆15Updated this week
- Proxy for Looking Glass over local networks☆28May 13, 2024Updated 2 years ago
- Live demo of hls4ml on embedded platforms such as the Pynq-Z2☆13Aug 23, 2024Updated last year
- Perceptron-based branch predictor written in C++☆14Dec 14, 2016Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An OpenVoice-based voice cloning tool, single executable file (~14M), supporting multiple formats without dependencies on ffmpeg, Python,…☆49Jan 18, 2026Updated 6 months ago
- Image synthesis using machine learning☆24May 6, 2025Updated last year
- ☆15Mar 6, 2025Updated last year
- Evaluate state-of-the-art sparse embedding models on the LIMIT dataset (`limit-small` and `limit`) from google's paper `On the Theoretica…☆16Sep 4, 2025Updated 10 months ago
- open-webui-runpod-integration☆17Jan 19, 2025Updated last year
- Efficient SAM3 (Segment Anything Model 3) inference from scratch in pure C — Metal GPU + multithreaded CPU, no Python dependencies☆17May 10, 2026Updated 2 months ago
- A novel hybrid AI architecture leveraging Titan's-like memory and HRM-like reasoning☆26Updated this week
- ☆23Jun 1, 2025Updated last year
- An Assortment of Convolutional Neural Networks☆11Mar 10, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tools for SideFX Houdini☆13Apr 2, 2020Updated 6 years ago
- A miniaturized version of the Kimi-K2 model optimized for deployment on single H100 GPUs.☆35Jul 16, 2025Updated last year
- Implementation of UltraMem, improved Product Key Memory design, from Bytedance AI labs☆28Nov 4, 2025Updated 8 months ago
- houdini hda that allows superzoom to specific coordinates☆14May 13, 2018Updated 8 years ago
- Offline-first, desktop AI assistant tailored for educators, enabling them to generate questions directly from source materials.☆24Aug 2, 2025Updated 11 months ago
- This is @BalazsJako ImGuiColorTextEdit widget improved by @dfranx with some mods (on my madX[TM] branch)☆20Jan 4, 2022Updated 4 years ago
- Qwen 3.5 in C☆24Mar 28, 2026Updated 4 months ago
- Python embedded in a Bifrost Operator☆12Nov 15, 2024Updated last year
- 🦙 Turn your idle Mac or GPU into a free public AI API — OpenAI-compatible, one command, zero config (Powered by Llama.cpp)☆24Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [TECS'23] A project on the co-design of Accelerators and CNNs.☆22Dec 10, 2022Updated 3 years ago
- Memory-bounded compressed sparse attention via streaming top-k. Triton kernels for the DeepSeek-V4 lightning indexer. 32x regime extensio…☆22May 5, 2026Updated 2 months ago
- Applying SAEs for fine-grained control☆27Dec 15, 2024Updated last year
- DOSA: Differentiable Model-Based One-Loop Search for DNN Accelerators☆20Oct 10, 2024Updated last year
- Houdini PDG nodes☆16Feb 24, 2022Updated 4 years ago
- ☆20Feb 12, 2025Updated last year
- ☆22Jul 11, 2025Updated last year