Minimal repository to demonstrate fast LoRA inference with Flux family of models.
☆32Jul 23, 2025Updated last year
Alternatives and similar repositories for lora-fast
Users that are interested in lora-fast are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 7 months ago
- Quantized Attention on GPU☆45Nov 22, 2024Updated last year
- Making Flux go brrr on GPUs.☆172Jan 5, 2026Updated 7 months ago
- A forked version of flux-fast that makes flux-fast even faster with cache-dit, 3.3x speedup on NVIDIA L20.☆24Jul 18, 2025Updated last year
- ☆22May 5, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2026] Official implementation of DiCache: Let Diffusion Model Determine Its Own Cache☆64Jan 26, 2026Updated 7 months ago
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"☆15Feb 8, 2026Updated 6 months ago
- ☆10Aug 31, 2023Updated 3 years ago
- Controllable Face Generation via pretrained Conditional Adversarial Latent Autoencoder (ALAE)☆20Jun 9, 2020Updated 6 years ago
- Molecular computers with interaction combinators like graph rewrite systems☆17Nov 9, 2022Updated 3 years ago
- [ECCV 2024] Multiscale Sliced Wasserstein Distances as Perceptual Color Difference Measures☆36Jul 15, 2026Updated last month
- LoRAFusion: Efficient LoRA Fine-Tuning for LLMs☆30Jul 2, 2026Updated 2 months ago
- ☆73Dec 5, 2025Updated 8 months ago
- writing really fast kernels☆20Jul 15, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 🎨 Native AI image generation for Apple Silicon with Qwen-Image. Lightning LoRA acceleration for fast 4–8 step runs. Zero Docker, just wo…☆24Sep 15, 2025Updated 11 months ago
- ☆10Apr 24, 2023Updated 3 years ago
- Just another reasonably minimal repo for class-conditional training of pixel-space diffusion transformers.☆158May 29, 2025Updated last year
- Inference-time scaling of diffusion-based image and video generation models.☆175Dec 17, 2025Updated 8 months ago
- Agent-native Seedance 2.0 short-film studio: cli for AI, canvas for human☆16Aug 27, 2026Updated last week
- Contributions to Playwright for .NET 🎭🧪☆12Nov 20, 2023Updated 2 years ago
- Kernel sources for https://huggingface.co/kernels-community☆144Updated this week
- ⚡️Qwen-Image 4.8x🎉 speedup with Hybrid Acceleration for low VRAM GPUs☆17Oct 24, 2025Updated 10 months ago
- A GitHub repository used to collaborate on recipes☆12Sep 1, 2015Updated 11 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Performant kernels for symmetric tensors☆17Aug 22, 2024Updated 2 years ago
- ☆18Mar 18, 2024Updated 2 years ago
- ☆16Sep 4, 2024Updated last year
- SParse AcceleRation on Tensor Architecture☆18Apr 15, 2026Updated 4 months ago
- ☆15Apr 21, 2025Updated last year
- KopikatAPI is Python library for interacting with the Kopikat API.☆17Updated this week
- Diffusers reimplementation for https://rf-inversion.github.io/☆25Dec 17, 2024Updated last year
- SDK for Seldon Deploy☆15Dec 18, 2024Updated last year
- ☆16Sep 30, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 小彭老师推出 SyCL 2020 课程(施工中,日后会在直播中放出)☆15Sep 3, 2023Updated 3 years ago
- Program to calculate quantum geometry properties of parametrized quantum circuits. Determines number of redundant parameters, effective q…☆17Feb 3, 2025Updated last year
- ☆10Jul 13, 2024Updated 2 years ago
- CUDA Embedding Lookup Kernel Library☆50Jun 26, 2026Updated 2 months ago
- three-body simulation visualized via OpenGL☆13Jun 8, 2023Updated 3 years ago
- KsanaDiT: High-Performance DiT (Diffusion Transformer) Inference Framework for Video & Image Generation☆62May 13, 2026Updated 3 months ago
- ☆13Apr 22, 2024Updated 2 years ago