π€ Optimum ExecuTorch
β142Oct 7, 2026Updated this week
Alternatives and similar repositories for optimum-executorch
Users that are interested in optimum-executorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Example apps and demos using PyTorch's ExecuTorch frameworkβ83Updated this week
- On-device AI across mobile, embedded and edge for PyTorchβ5,087Updated this week
- Rust bindings for ExecuTorch - On-device AI across mobile, embedded and edge for PyTorchβ62Oct 1, 2026Updated last week
- π€ Optimum ONNX: Export your model to ONNX and run inference with ONNX Runtimeβ161Updated this week
- β19May 7, 2026Updated 5 months ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Automatically derive Python dunder methods for your Rust codeβ25Sep 24, 2026Updated 2 weeks ago
- π· Build compute kernelsβ213Apr 6, 2026Updated 6 months ago
- Efficient in-memory representation for ONNX, in Pythonβ50Updated this week
- A variable-frame-rate 16 kHz speech codec based on FocalCodecβ21Feb 11, 2026Updated 8 months ago
- Google TPU optimizations for transformers modelsβ136Jan 23, 2026Updated 8 months ago
- Support PyTorch model conversion with LiteRT.β1,104Updated this week
- Temporally-aligned Audio CaptiOnS for Language-Audio Pretrainingβ16Oct 12, 2025Updated 11 months ago
- Code for the examples presented in the talk "Training a Llama in your backyard: fine-tuning very large models on consumer hardware" givenβ¦β15Oct 16, 2023Updated 2 years ago
- A home for the final text of all TVM RFCs.β111Sep 24, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β188Sep 14, 2026Updated 3 weeks ago
- Benchmarking LLM Inference Speedsβ14Sep 30, 2026Updated last week
- β25Jun 12, 2025Updated last year
- β40Mar 6, 2026Updated 7 months ago
- One-size-fits-all model for mobile AI, a novel paradigm for mobile AI in which the OS and hardware co-manage a foundation model that is cβ¦β30Mar 5, 2024Updated 2 years ago
- A safetensors extension to efficiently store sparse quantized tensors on diskβ329Updated this week
- High-speed and easy-use LLM serving framework for local deploymentβ166Aug 7, 2025Updated last year
- Embedding and readout for simple multi-categorical and gaussian continuousβ20Jul 5, 2026Updated 3 months ago
- AMD related optimizations for transformer modelsβ101Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The backend behind the LLM-Perf Leaderboardβ11May 5, 2024Updated 2 years ago
- Generative AI extensions for onnxruntimeβ1,135Updated this week
- π€ Tokenizers.js: A pure JS/TS implementation of today's most used tokenizersβ57Sep 23, 2026Updated 2 weeks ago
- How to fix Ubuntu 18.04/20.04 not recognizing M-Audio Fast Track Pro USB audio interfaceβ18Aug 10, 2020Updated 6 years ago
- Dice Language Support for VS Codeβ10Sep 29, 2020Updated 6 years ago
- High-efficiency floating-point neural network inference operators for mobile, server, and Webβ2,469Updated this week
- Old implementation of the MaxTract system for re-engineering mathematical PDF documents.β13Jan 25, 2016Updated 10 years ago
- Training and inference on AWS Trainium and Inferentia chips.β270Updated this week
- Implementation of UltraMem, improved Product Key Memory design, from Bytedance AI labsβ28Nov 4, 2025Updated 11 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- π Accelerate inference and training of π€ Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimizationβ¦β3,502Updated this week
- PyTorch native quantization for training and inferenceβ2,997Updated this week
- In this project, we propose to study Vision Transformers trained using the Barlow Twins self-supervised method, and compare the results wβ¦β17Oct 3, 2023Updated 3 years ago
- [EMNLP Main '25] LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximationβ158May 18, 2025Updated last year
- Build compute kernels and load them from the Hub.β765Updated this week
- TORCH_TRACE parser for PT2β94May 11, 2026Updated 4 months ago
- [WIP] A π₯ interface for running code in the cloudβ87Sep 23, 2026Updated 2 weeks ago