Documentation for StreamExecutor open source proposal
☆83Mar 28, 2016Updated 10 years ago
Alternatives and similar repositories for streamexecutordoc
Users that are interested in streamexecutordoc are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆371Oct 23, 2017Updated 8 years ago
- tensorflow源码阅读笔记☆191Sep 18, 2018Updated 7 years ago
- Documentation for the entire CGRAFlow☆19Sep 17, 2021Updated 4 years ago
- ☆24Jul 31, 2017Updated 9 years ago
- ☆21Oct 15, 2015Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Minimal numerical computation library with TensorFlow APIs☆303Jan 2, 2019Updated 7 years ago
- An IR for efficiently simulating distributed ML computation.☆33Jan 13, 2024Updated 2 years ago
- Compiler toolkit for neuFlow.☆27Jul 7, 2013Updated 13 years ago
- ☆24Apr 25, 2023Updated 3 years ago
- TF2 implementation of DLRM (inherited and modified from openrec's initial implementation)☆15Jul 14, 2020Updated 6 years ago
- GPU-specialized parameter server for GPU machine learning.☆102Apr 5, 2018Updated 8 years ago
- Enhanced networking support for TensorFlow. Maintained by SIG-networking.☆99Nov 19, 2021Updated 4 years ago
- Reed-Solomon Erasure Coding in Haskell☆23Jan 22, 2017Updated 9 years ago
- BladeDISC is an end-to-end DynamIc Shape Compiler project for machine learning workloads.☆932Dec 30, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A fast multi-producer, multi-consumer lock-free concurrent queue for C++11☆10May 25, 2015Updated 11 years ago
- Installation scripts for CUDA, cuDNN, TensorFlow, Caffe, etc. on Ubuntu machines☆24Aug 1, 2021Updated 5 years ago
- The Tensor Algebra SuperOptimizer for Deep Learning☆743Jan 26, 2023Updated 3 years ago
- Proof-of-Concept CNN in Halide☆22Aug 4, 2016Updated 10 years ago
- Just save my record on github...☆28Feb 7, 2021Updated 5 years ago
- Sublinear memory optimization for deep learning, reduce GPU memory cost to train deeper nets☆28Apr 22, 2016Updated 10 years ago
- A benchmark framework for Tensorflow☆1,148Oct 6, 2023Updated 2 years ago
- An Efficient Pipelined Data Parallel Approach for Training Large Model☆76Dec 11, 2020Updated 5 years ago
- TVM stack: exploring the incredible explosion of deep-learning frameworks and how to bring them together☆65May 22, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Collective communications library with various primitives for multi-machine training.☆1,447Aug 21, 2026Updated last week
- A flexible and efficient deep neural network (DNN) compiler that generates high-performance executable from a DNN model description.☆1,002Sep 19, 2024Updated last year
- Voice from TUNA☆18Dec 10, 2018Updated 7 years ago
- heterogeneity-aware-lowering-and-optimization☆258Jan 20, 2024Updated 2 years ago
- ☆607Apr 6, 2018Updated 8 years ago
- TensorFlow kernels for probing memory☆15Mar 2, 2017Updated 9 years ago
- package vecf64 provides common functions and methods for slices of float64☆24Jun 14, 2023Updated 3 years ago
- MPI bindings for Haskell☆46Apr 1, 2023Updated 3 years ago
- A simple memory manager for CUDA designed to help Deep Learning frameworks manage memory☆298Nov 28, 2018Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Chaos: Scale-out Graph Processing from Secondary Storage☆52Mar 14, 2016Updated 10 years ago
- It is open source ebook about TensorFlow kernel and implementation mechanism.☆2,888May 5, 2023Updated 3 years ago
- An MLIR frontend for tensor expressions☆24Sep 5, 2020Updated 5 years ago
- Python bindings for NVTX☆67Jun 9, 2023Updated 3 years ago
- Optimized primitives for collective multi-GPU communication☆5,036Updated this week
- OneFlow->ONNX☆42Apr 19, 2023Updated 3 years ago
- The Tensor Algebra Compiler (taco) computes sparse tensor expressions on CPUs and GPUs☆1,367Apr 14, 2025Updated last year