High-Performance Machine Learning Primitives
☆13Apr 17, 2021Updated 5 years ago
Alternatives and similar repositories for hmlp
Users that are interested in hmlp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Software library RLCM (recursively low-rank compressed matrices)☆14Apr 15, 2021Updated 5 years ago
- Quad/octree building for FMMs in Python and OpenCL☆66Jul 17, 2026Updated last week
- H2 Matrix Package☆31Jul 18, 2023Updated 3 years ago
- Software libraries that implement hierarchical matrices☆64Jun 19, 2025Updated last year
- Structured Matrix Package (LBNL)☆196Jun 6, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- hmglib - Hierarchical matrices on GPU(s) library☆13Jul 31, 2018Updated 7 years ago
- directory for randomized cholesky☆19Oct 8, 2025Updated 9 months ago
- ☆64Updated this week
- ☆21Aug 21, 2023Updated 2 years ago
- Flatiron Institute Fast Multipole Libraries --- This codebase is a set of libraries to compute N-body interactions governed by the Laplac…☆157Apr 2, 2026Updated 3 months ago
- Parallelized BBFMM3D with OpenMP☆32Oct 8, 2025Updated 9 months ago
- ☆26Sep 14, 2018Updated 7 years ago
- A conda-smithy repository for colmap.☆15Jul 15, 2026Updated last week
- Enhancing the convergence speed by 2x and improving the training success of Physics-Informed Neural Networks (PINNs).☆13Oct 14, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Integrated Interface for libraries of eigenvalue decomposition☆10Nov 29, 2024Updated last year
- A parallel kernel-independent FMM library for particle and volume potentials☆63Jul 10, 2026Updated 2 weeks ago
- ☆38Jun 8, 2026Updated last month
- Arrow Matrix Decomposition - Communication-Efficient Distributed Sparse Matrix Multiplication☆15Mar 25, 2024Updated 2 years ago
- Distributed-memory, double-precision, polar decomposition (QDWH/ZOLO-PD) of a dense matrix, svd (QDWH/ZOLOPD-SVD) of a dense matrix☆14Jun 3, 2020Updated 6 years ago
- Strassen's Algorithm for Tensor Contraction☆15Jul 7, 2017Updated 9 years ago
- A shared-memory FFT for the Kokkos ecosystem☆61Jul 10, 2026Updated 2 weeks ago
- A little library for using SIMD instructions for x86 and ARM, wrapping Agner Fog's vectorclass for x86 and filling some of its functional…☆17May 13, 2026Updated 2 months ago
- Catamount is a compute graph analysis tool to load, construct, and modify deep learning models and to symbolically analyze their compute …☆14May 18, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- devector and batch_deque containers for C++. See more at: http://erenon.hu/double_ended☆16Oct 7, 2017Updated 8 years ago
- Repo for Duchamp Chess Set STL files http://www.thingiverse.com/thing:305639/#files☆15Dec 15, 2014Updated 11 years ago
- ☆17Apr 8, 2021Updated 5 years ago
- Volume Manipulation Library☆17Jul 13, 2023Updated 3 years ago
- Parallel implementation of k-means clustering using MPI4PY and PyCUDA.☆10Mar 11, 2019Updated 7 years ago
- Sparse Matrix-Matrix Multiplication Benchmark on Intel Xeon and Xeon Phi (KNC, KNL) from blog post:☆12Sep 25, 2016Updated 9 years ago
- ☆10Apr 24, 2023Updated 3 years ago
- FLOPS counter for all your GPU benchmarking needs☆13Aug 8, 2024Updated last year
- Solve large instance of semi-discrete optimal transport problems and other Monge-Ampere equations☆34Jan 20, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A suite of stochastic optimization methods for solving the empirical risk minimization problem.☆17Nov 20, 2019Updated 6 years ago
- Distributed memory, MPI based SuperLU☆221Jul 11, 2026Updated last week
- CSR-based SpGEMM on nVidia and AMD GPUs☆48Apr 9, 2016Updated 10 years ago
- A terminal-based citation generator☆14Jan 21, 2023Updated 3 years ago
- Version 1.2☆13Mar 15, 2017Updated 9 years ago
- Web frontend for Myria☆12Sep 30, 2020Updated 5 years ago
- [experimental] multiplexed distributed tensor framework☆22Nov 17, 2025Updated 8 months ago