Library for fast image convolution in neural networks on Intel Architecture
☆30Jun 25, 2017Updated 9 years ago
Alternatives and similar repositories for FALCON
Users that are interested in FALCON are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Generating Families of Practical Fast Matrix Multiplication Algorithms☆12Jul 7, 2017Updated 9 years ago
- Phase Fair and Standard Reader Writer Locks☆16Sep 16, 2019Updated 6 years ago
- Improved performance for TensorFlow on Intel hardware.☆13Jun 25, 2018Updated 8 years ago
- Winograd minimal convolution algorithm generator for convolutional neural networks.☆629Feb 9, 2026Updated 6 months ago
- ☆10Aug 4, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Portable 128-bit SIMD intrinsics☆61Jul 4, 2023Updated 3 years ago
- Torch FFI-bindings for NNPACK☆31May 26, 2017Updated 9 years ago
- Library and accelerator backend☆15Updated this week
- C99/C++ header-only library for division via fixed-point multiplication by inverse☆61Apr 14, 2024Updated 2 years ago
- flexible-gemm conv of deepcore☆17Dec 2, 2019Updated 6 years ago
- TensorFlow frozen forward model to plain C++ converter☆10Aug 7, 2018Updated 8 years ago
- Greentea LibDNN - a universal convolution implementation supporting CUDA and OpenCL☆137Apr 20, 2017Updated 9 years ago
- ☆18Apr 8, 2022Updated 4 years ago
- ☆21Jan 21, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- detect facial landmark with mini-caffe☆18Feb 23, 2017Updated 9 years ago
- AnyDSL traversal code☆15Feb 18, 2019Updated 7 years ago
- ☆14Feb 7, 2020Updated 6 years ago
- A minimalist Deep Learning framework for embedded Computer Vision☆47Dec 31, 2019Updated 6 years ago
- CUDA and OpenMP implementations of C2R/R2C inplace transposition☆49Feb 10, 2015Updated 11 years ago
- Library for specialized dense and sparse matrix operations, and deep learning primitives.☆973Updated this week
- ☆16Jul 29, 2022Updated 4 years ago
- The SparseX sparse kernel optimization library☆43Jan 16, 2019Updated 7 years ago
- DelugeNets: Deep Networks with Efficient and Flexible Cross-layer Information Inflows☆26Mar 20, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Caffe implementation of the paper "Deep Pyramidal Residual Networks" (https://arxiv.org/abs/1610.02915).☆27Jul 18, 2017Updated 9 years ago
- Acceleration package for neural networks on multi-core CPUs☆1,710Jun 11, 2024Updated 2 years ago
- The HPC toolbox: fused matrix multiplication, convolution, data-parallel strided tensor primitives, OpenMP facilities, SIMD, JIT Assemble…☆295Jan 4, 2024Updated 2 years ago
- Read audio with FFmpeg into NumPy/PyTorch via ctypes (standard library module)☆11Aug 12, 2020Updated 5 years ago
- Depict GPU memory footprint during DNN training of PyTorch☆11Nov 17, 2022Updated 3 years ago
- An MPI-based C++ or Python library for easy distributed pipeline processing☆33Jul 30, 2018Updated 8 years ago
- A Winograd based kernel for convolutions in deep learning framework☆15Jul 22, 2017Updated 9 years ago
- GPT Demo with hybrid distributed training☆10Dec 1, 2022Updated 3 years ago
- a fork of clang with Sierra patches☆20Aug 16, 2018Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- pycaffe version of RSA 'Recurrent Scale Approximation for Object Detection in CNN'☆32Dec 5, 2017Updated 8 years ago
- BlockCIrculantRNN (LSTM and GRU) using TensorFlow☆14Oct 30, 2018Updated 7 years ago
- Graphics buffer bridge for Maru OS.☆10Nov 21, 2020Updated 5 years ago
- UME::SIMD A library for explicit simd vectorization.☆90Jan 19, 2018Updated 8 years ago
- Proof-of-Concept CNN in Halide☆22Aug 4, 2016Updated 10 years ago
- GPU-accelerated AES encryption project☆11Feb 13, 2015Updated 11 years ago
- Hanzi to Pinyin engine in Swift 拼音输入法引擎☆13Mar 29, 2024Updated 2 years ago