Parallel Prefix Sum (Scan) with CUDA.
☆15Jul 17, 2020Updated 6 years ago
Alternatives and similar repositories for CUDA-Parallel-Prefix-Sum
Users that are interested in CUDA-Parallel-Prefix-Sum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenGraph is an open-source graph processing benchmarking suite written in pure C/OpenMP.☆14Apr 27, 2024Updated 2 years ago
- This repository is outdated and the related functionality has been migrated to https://github.com/easysoc/easysoc-firrtl☆11Nov 3, 2021Updated 4 years ago
- Zero dependency Model generator based on the method of least squares.☆11Jul 10, 2026Updated 3 months ago
- Code for the paper: https://arxiv.org/pdf/2309.06979.pdf☆21Jul 29, 2024Updated 2 years ago
- High efficiency and embeddable MPM implementation.☆16Aug 4, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Procyon is the brightest star in the constellation of Canis Minor. But it's also the name of my RISC-V out-of-order processor.☆12Apr 6, 2023Updated 3 years ago
- Dynamic Hashed Blocks (DHB) data structure for dynamic graphs☆13Sep 23, 2026Updated 2 weeks ago
- Variadic reference-based implementation of a zip iterator in C++(>=17)☆20Jul 1, 2020Updated 6 years ago
- ☆10Mar 24, 2023Updated 3 years ago
- Extending the Neural Graph Algorithm Executor☆13Dec 8, 2022Updated 3 years ago
- GPU for OENG1167 in Verilog HDL for DE10 series boards☆15Nov 1, 2020Updated 5 years ago
- 🕒 Static Timing Analysis diagram renderer☆13Dec 13, 2023Updated 2 years ago
- Official code repo for paper: ACROSS: An Alignment-based Framework for Low-Resource Many-to-One Cross-Lingual Summarization☆12Jul 15, 2023Updated 3 years ago
- QuteRTL: A RTL Front-End Towards Intelligent Synthesis and Verification☆18Nov 8, 2016Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Apr 30, 2021Updated 5 years ago
- ☆21Nov 18, 2018Updated 7 years ago
- An React Single Page Application for explore Wikipedia articles☆11Dec 14, 2022Updated 3 years ago
- ☆19Mar 28, 2023Updated 3 years ago
- ☆26Mar 15, 2022Updated 4 years ago
- CUDA implementation of exclusive prefix sum via Blelloch's algorithm☆29Jul 19, 2017Updated 9 years ago
- Verilog AST☆19Dec 2, 2023Updated 2 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- Diffusion Monte Carlo method☆12Nov 2, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for "Deep Energy-Based Modeling of Discrete-Time Physics," NeurIPS, 2020. (Oral)☆19Jan 30, 2022Updated 4 years ago
- ☆11Oct 11, 2023Updated 3 years ago
- ☆10Jun 17, 2020Updated 6 years ago
- Cortex-M0 DesignStart Wrapper☆24Aug 11, 2019Updated 7 years ago
- Fine-Tuning Pre-trained Transformers into Decaying Fast Weights☆20Oct 9, 2022Updated 4 years ago
- Statistical discontinuous constituent parsing☆11Feb 15, 2018Updated 8 years ago
- Implementation of Hyena Hierarchy in JAX☆10Apr 30, 2023Updated 3 years ago
- Unofficial implementation of paper : Exploring the Space of Key-Value-Query Models with Intention☆12May 24, 2023Updated 3 years ago
- ☆13Jun 16, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- JAX/Flax implementation of the Hyena Hierarchy☆35Apr 27, 2023Updated 3 years ago
- Implementation of the paper "Shortest Path Distance Approximation using Deep learning Techniques" (under development)☆20May 24, 2022Updated 4 years ago
- Recursive Bayesian Networks☆11May 11, 2025Updated last year
- PyTorch implementation for PaLM: A Hybrid Parser and Language Model.☆10Jan 7, 2020Updated 6 years ago
- The FCUDA CUDA-to-RTL compiler☆23Jul 1, 2016Updated 10 years ago
- Findings of ACL 2021☆24May 8, 2021Updated 5 years ago
- A demo implementation of a MCP server (consuming a dummy API) and basic client.☆29Mar 18, 2025Updated last year