pure c/cpp cnn implementation, with CUDA accelerated.
☆21Apr 30, 2021Updated 5 years ago
Alternatives and similar repositories for SimpleCNN_Release
Users that are interested in SimpleCNN_Release are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- C++ implement a simple CNN framework to train mnist data. Done!☆10Mar 29, 2022Updated 4 years ago
- Created a simple neural network using C++17 standard and the Eigen library that supports both forward and backward propagation.☆11Jul 27, 2024Updated 2 years ago
- 方便扩展的Cuda算子理解和优化框架,仅用在学习使用☆18Jun 13, 2024Updated 2 years ago
- Official implementation of Acc-SpMM: Accelerating General-purpose Sparse Matrix-Matrix Multiplication with GPU Tensor Cores.☆38Nov 13, 2025Updated 9 months ago
- ☆14May 30, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- MATLAB code for laminography reconstruction using SIRT and CGLS☆14Jul 3, 2019Updated 7 years ago
- [AAAI 2026] This is the official implementation of the paper "ExtendAttack: Attacking Servers of LRMs via Extending Reasoning".☆26Mar 18, 2026Updated 5 months ago
- A Project dedicated to making GPU Partitioning on Windows easier!☆15Jan 10, 2022Updated 4 years ago
- Implementation of the D-Stream clustering algorithm for use in MOA. An earlier version is included as part of the MOA 17.06 release.☆14Jan 29, 2019Updated 7 years ago
- CNN accelerated by cuda. Test on mnist and finilly get 99.76%☆188Oct 15, 2017Updated 8 years ago
- 基于Qwen2.5模型、使用DISC-Law-SFT-Pair数据集微调的法律大模型☆11Dec 29, 2024Updated last year
- ☆11Sep 12, 2023Updated 2 years ago
- A D-Stream clustering algorithm implementation in Python☆14Mar 25, 2016Updated 10 years ago
- ☆11Oct 9, 2019Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Build CUDA Neural Network From Scratch☆22Aug 28, 2024Updated 2 years ago
- "FORB: A Flat Object Retrieval Benchmark for Universal Image Embedding", NeurIPS 2023 Datasets and Benchmarks Track☆13Jun 20, 2024Updated 2 years ago
- ☆20Sep 28, 2024Updated last year
- DiscreteTom's Blog Boilerplate.☆10Mar 6, 2023Updated 3 years ago
- Optimized Computer Graphics Matrix Library for use with the SIMD/SSE4 Instructions.☆10Mar 19, 2020Updated 6 years ago
- This repo is "NTHU Parallel Programing" course project.☆10Dec 5, 2017Updated 8 years ago
- Compress BiSeNet with Structure Knowledge Distillation for Real-time image segmentation on wali-TX2☆11Jul 29, 2020Updated 6 years ago
- FlashSparse significantly reduces the computation redundancy for unstructured sparsity (for SpMM and SDDMM) on Tensor Cores through a Swa…☆39Oct 5, 2025Updated 10 months ago
- NTHU CS6135 VLSI實體設計自動化☆11Mar 12, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fastest CUDA SIFT or other 128-float vector matcher for computer vision☆28Mar 23, 2021Updated 5 years ago
- 适配目前最新CUDA环境的SIFTGPU代码☆10May 6, 2020Updated 6 years ago
- A fast, small, efficient pthreads based threadpool in c☆16Mar 2, 2021Updated 5 years ago
- a simple API to use CUPTI☆10Aug 19, 2025Updated last year
- Pose-Based View Synthesis for Vehicles: A Perspective Aware Method☆12Nov 22, 2022Updated 3 years ago
- ☆14May 23, 2024Updated 2 years ago
- 一步步实现c++中的智能指针☆10Jun 6, 2021Updated 5 years ago
- Semi-Tenser Product based SAT and AllSAT solver, where it can solve CNF and circuit input.☆17Aug 2, 2023Updated 3 years ago
- ☆32May 1, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Sep 13, 2020Updated 5 years ago
- a vue-demo:vue仿网易新闻m站☆10Jul 26, 2017Updated 9 years ago
- Triton to TVM transpiler.☆24Oct 14, 2024Updated last year
- 计算机网络微课堂笔记☆16Jul 1, 2023Updated 3 years ago
- Fastest CUDA RGB to grayscale: 5-30x faster than OpenCV. For image processing/computer vision.☆16Mar 23, 2021Updated 5 years ago
- 用C++实现的一个简单的线程池,支持任务队列,实际任务继承自taskbase。☆12Apr 15, 2015Updated 11 years ago
- An ATPG tool using PODEM algorithm in C++ that generates a test to detect any given list of Single-Stuck-at Faults☆11Oct 29, 2017Updated 8 years ago