☆30Apr 18, 2024Updated 2 years ago
Alternatives and similar repositories for LibShalom
Users that are interested in LibShalom are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Apr 8, 2022Updated 4 years ago
- A direct convolution library targeting ARM multi-core CPUs.☆12Nov 27, 2024Updated last year
- Sparse kernels for GNNs based on TVM☆17Nov 18, 2020Updated 5 years ago
- Absinthe is an optimization framework to fuse and tile stencil codes in one shot☆14Jul 17, 2019Updated 7 years ago
- This is a GPU optimized version of ShengBTE.☆20Oct 3, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Nov 26, 2025Updated 9 months ago
- DietCode Code Release☆66Jul 21, 2022Updated 4 years ago
- ☆10Apr 24, 2023Updated 3 years ago
- Spack package repository maintained by Student Cluster Competition Team @ Sun Yat-sen University.☆16Aug 20, 2025Updated last year
- A High performance and tiny TVM graph executor library written in C which can compile to WebAssembly and use CUDA/WebGPU as the accelerat…☆13Aug 3, 2023Updated 3 years ago
- Paper: inexact GMRES with fast multipole method and low-p relaxation☆11Aug 23, 2023Updated 3 years ago
- GEMM by WMMA (tensor core)☆15Jul 31, 2022Updated 4 years ago
- HPC Challenge Benchmark☆70Sep 28, 2025Updated 11 months ago
- symmetric int8 gemm☆67Jun 7, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆10Jun 4, 2021Updated 5 years ago
- ☆40Feb 28, 2020Updated 6 years ago
- ☆11Mar 2, 2024Updated 2 years ago
- ☆12May 3, 2020Updated 6 years ago
- A Top-Down Profiler for GPU Applications☆24Feb 29, 2024Updated 2 years ago
- Implementation of the Barnes-Hut algorithm in C++☆12Jul 8, 2011Updated 15 years ago
- 东北大学本科毕业设计 论文latex模板 适应2021届新版书写印制规范 针对计算机类专业☆11Apr 21, 2021Updated 5 years ago
- Python package to predict deep learning execution time☆13Jul 26, 2022Updated 4 years ago
- 慕课网 thinkphp5.0 微信小程序 零食商贩项目 小程序令牌测试工具☆12Dec 13, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Aug 4, 2022Updated 4 years ago
- A comprehensive benchmarking framework for evaluating and optimizing CPU-centric agentic AI systems across multiple workloads, reproducin…☆53Feb 12, 2026Updated 6 months ago
- [COLM 2024] SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models☆25Oct 5, 2024Updated last year
- NVFP4 Flash-Attention 4 on BlackWell☆50Jul 23, 2026Updated last month
- 东北大学本科毕业设计 论文latex模板 2020 针对计算机相关专业☆12Jun 10, 2020Updated 6 years ago
- An attempt to replicate the paper "Multi-shot Pedestrian Re-identification via Sequential Decision Making (CVPR2018)"☆10Nov 16, 2019Updated 6 years ago
- study of cutlass☆22Nov 10, 2024Updated last year
- 中山大学2020年并行与分布式计算作业☆21Jul 28, 2020Updated 6 years ago
- CAKE Library for constant-bandwidth matrix multiplication on CPUs☆14Apr 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Magicube is a high-performance library for quantized sparse matrix operations (SpMM and SDDMM) of deep learning on Tensor Cores.☆92Nov 23, 2022Updated 3 years ago
- 将MNN拆解的简易前向推理框架(for study!)☆24Feb 21, 2021Updated 5 years ago
- Document image binarization for Project 3A @Mines_Nancy☆30Aug 9, 2017Updated 9 years ago
- ☆23Aug 14, 2024Updated 2 years ago
- Student teamworks summary repository for USTC Compiler H lecture in fall, 2017.☆15Jan 13, 2018Updated 8 years ago
- A source-to-source compiler for optimizing CUDA dynamic parallelism by aggregating launches☆15Jun 21, 2019Updated 7 years ago
- Friends-of-Friends via spatial hashing☆15May 26, 2023Updated 3 years ago