☆36Aug 25, 2023Updated 2 years ago
Alternatives and similar repositories for conv2d_direct
Users that are interested in conv2d_direct are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- A CUDA kernel for NHWC GroupNorm for PyTorch☆26Nov 15, 2024Updated last year
- A direct convolution library targeting ARM multi-core CPUs.☆12Nov 27, 2024Updated last year
- Some "Formula Translations" for Yousef Saad's book "Iterative Methods for Sparse Linear Systems (2nd Edition)"☆13Jan 14, 2018Updated 8 years ago
- This project is primarily used to deploy large language models and multimodal large models on Orin.🚀🚀🚀☆18Jun 23, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MINT, Multiplier-less INTeger Quantization for Energy Efficient Spiking Neural Networks, ASP-DAC 2024, Nominated for Best Paper Award☆16Apr 12, 2024Updated 2 years ago
- A set of examples around MegEngine☆31Dec 8, 2023Updated 2 years ago
- ☆10Sep 23, 2023Updated 2 years ago
- ☆14Jul 16, 2020Updated 6 years ago
- CUDA GPU implementation of GMRES iterative Solver☆10Apr 16, 2012Updated 14 years ago
- Provides json/csv/protobuf/arrow streaming support for reqwest HTTP client☆18Jun 18, 2026Updated 2 months ago
- ☆12Jan 19, 2020Updated 6 years ago
- ☆23Aug 20, 2025Updated 11 months ago
- OpenFOAM right wmake at the right time☆11Mar 10, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Experimental pipeline for FedFace.☆10Jul 6, 2021Updated 5 years ago
- implement llava using candle☆15Jun 9, 2024Updated 2 years ago
- 对 YOLOv3 做模型剪枝(network slimming),对于 oxford hand 数据集(因项目需要),模型剪枝后的参数量减少 80%,Infer 的速度达到原来 2 倍,mAP 基本不变☆12Jul 12, 2019Updated 7 years ago
- MESMERIC: A Software-based NVM Emulator Supporting Read/Write Asymmetric Latencies☆10Oct 1, 2020Updated 5 years ago
- ☆12Sep 29, 2021Updated 4 years ago
- Implement custom operators in PyTorch with cuda/c++☆77Jan 1, 2023Updated 3 years ago
- ☆13Jan 18, 2020Updated 6 years ago
- tensor library☆17Jul 19, 2024Updated 2 years ago
- A rust version of the Caffe library.☆19Jun 16, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆17Apr 9, 2025Updated last year
- Gate-Level Simulation on a GPU☆10Nov 22, 2016Updated 9 years ago
- SAM and lama inpaint,包含QT的GUI交互界面,实现了交互式可实时显示结果的画点、画框进行SAM,然后通过进行Inpaint,具体操作看readme里的视频。☆54Jan 30, 2024Updated 2 years ago
- ☆15Apr 28, 2023Updated 3 years ago
- This is the official code for our paper "Differentiable Hierarchical and Surrogate Gradient Search for Spiking Neural Networks, NeurIPS 2…☆16Oct 30, 2023Updated 2 years ago
- SeekFree RT1064 Library GCC(VSCode) Porting☆12Oct 8, 2021Updated 4 years ago
- Massively Scalable Parallel GMRES C-code for Sparse System of Equations☆13Feb 16, 2016Updated 10 years ago
- ☆32Aug 25, 2023Updated 2 years ago
- CUDA SGEMM optimization note☆15Oct 31, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆92Mar 31, 2026Updated 4 months ago
- CPU Memory Compiler and Parallel programing☆26Nov 18, 2024Updated last year
- ☆19Oct 3, 2022Updated 3 years ago
- lightning-lm ROS1 noetic 旧时代的残党☆28Jul 28, 2026Updated 3 weeks ago
- This repository contains video datasets that can be used for training coarse to fine-grained (phase, step and action) temporal classifica…☆16Oct 26, 2021Updated 4 years ago
- ☆11Aug 8, 2018Updated 8 years ago
- NetHCF: Enabling Line-rate and Adaptive Spoofed IP Traffic Filtering☆13Mar 17, 2022Updated 4 years ago