High-performance FlashAttention-2 for AMD, Intel, and Apple GPUs. Drop-in replacement for PyTorch SDPA. Triton backend for ROCm (MI300X, RDNA3), Vulkan backend for consumer GPUs. No CUDA required.
☆159Jan 27, 2026Updated 6 months ago
Alternatives and similar repositories for Aule-Attention
Users that are interested in Aule-Attention are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Feb 20, 2025Updated last year
- WanImageToVideo ComfyUI node, with Tiled VAE☆16Oct 22, 2025Updated 9 months ago
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 10 months ago
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated last year
- ☆25Feb 10, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repo documents my workflows and stack to run comfy ui GenANI assist under windows☆39Updated this week
- ☆34Jul 19, 2026Updated 3 weeks ago
- ☆40Apr 29, 2024Updated 2 years ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,196Updated this week
- Text-generation-webui oneclick UI + API☆12Dec 27, 2024Updated last year
- Holds releases for HammerAI.com☆18Jul 31, 2026Updated last week
- A complete package that provides you with all the components needed to get started of dive deeper into Machine Learning Workloads on Cons…☆54Jul 28, 2026Updated last week
- LLM inference in C/C++, but for GFX906!☆21Updated this week
- ☆24Jan 22, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- A vector similarity search engine for humans🥳☆18Oct 30, 2023Updated 2 years ago
- Privacy-first agentic framework with powerful reasoning & task automation capabilities. Natively distributed and fully ISO 27XXX complian…☆70Apr 1, 2025Updated last year
- add support on amd in zluda☆78Jul 21, 2025Updated last year
- ☆23Sep 27, 2024Updated last year
- ☆15Jan 27, 2023Updated 3 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 10 months ago
- ScribePal is an Open Source intelligent browser extension that leverages AI to empower your web experience by providing contextual insigh…☆22Apr 6, 2026Updated 4 months ago
- Luth is a state-of-the-art series of fine-tuned LLMs for French☆47Oct 12, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆24Jul 30, 2026Updated last week
- AI Based "Happiness Optimizer"☆12Oct 20, 2024Updated last year
- ☆27Jul 20, 2026Updated 3 weeks ago
- REAP: Router-weighted Expert Activation Pruning for SMoE compression☆474Apr 17, 2026Updated 3 months ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆124Jun 10, 2026Updated 2 months ago
- A PyTorch implementation of gradient-free optimization for directly optimizing NDCG (Normalized Discounted Cumulative Gain) in neural inf…☆19Dec 21, 2025Updated 7 months ago
- RadialAttention in ComfyUI native workflow☆120Dec 19, 2025Updated 7 months ago
- MagicNodes, it's a plug-and-play multi-pass "render-machine" for SD/SDXL models. Simple one-node start, expert-grade results. Core is ZeR…☆91May 16, 2026Updated 2 months ago
- Fast fp16-fp8 mixed precision matmul on RDNA3/3.5 GPUs without native fp8☆34Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The most powerful and modular stable diffusion GUI, api and backend with a graph/nodes interface. Now ZLUDA enhanced for better AMD GPU p…☆954Updated this week
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆24Jan 24, 2026Updated 6 months ago
- Distribute and run LLMs with a single file.☆25May 13, 2025Updated last year
- ☆96Apr 26, 2025Updated last year
- Apple's Neural Engine - bare metal access☆44Jul 2, 2026Updated last month
- (ECCV 2026): Official code for Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models☆20Jul 9, 2026Updated last month
- ☆22Jul 25, 2023Updated 3 years ago