A framework and CLI toolkit for orchestrating teams of loosely-coupled AI agents.
β18Jul 28, 2026Updated last week
Alternatives and similar repositories for houmao
Users that are interested in houmao are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [MLSys 26] π₯ Solution for Gated Delta Net Track of MLSys 26 Flash infer competitionβ36May 22, 2026Updated 2 months ago
- β47Jul 27, 2026Updated last week
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloadsβ17Mar 12, 2026Updated 4 months ago
- LoRAFusion: Efficient LoRA Fine-Tuning for LLMsβ28Jul 2, 2026Updated last month
- Anderson points-to analysis implementation based on LLVMβ12Jan 3, 2021Updated 5 years ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- The official code for [ECCV2020] "HALO: Hardware-aware Learning to Optimize"β10Mar 22, 2023Updated 3 years ago
- [KernelGYM & Dr. Kernel] A distributed GPU environment and a collection of RL training methods to support RL for Kernel Generations [ICMLβ¦β200Mar 29, 2026Updated 4 months ago
- An agent for CUDA compute-communication kernel co-designβ36May 7, 2026Updated 2 months ago
- Scripts to prepare OXFORD VGG Face datasetβ12Mar 29, 2016Updated 10 years ago
- An open-sourced PyTorch library for developing energy efficient multiplication-less models and applications.β14Feb 3, 2025Updated last year
- β40Dec 14, 2025Updated 7 months ago
- Synchronizing Claude Code conversations across machinesβ16Jul 23, 2026Updated last week
- Speed of Light Analysis for ML Model Runtimeβ107Jun 10, 2026Updated last month
- Minimal PyTorch implementation of TP, SP, FSDP and sharded-EMAβ32Nov 27, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- β13Jan 21, 2024Updated 2 years ago
- [ICML 2022] ShiftAddNAS: Hardware-Inspired Search for More Accurate and Efficient Neural Networksβ15May 18, 2022Updated 4 years ago
- β33Mar 12, 2026Updated 4 months ago
- [ACL 2026 π₯] CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmarkβ35Apr 20, 2026Updated 3 months ago
- [Archived] For the latest updates and community contribution, please visit: https://github.com/Ascend/TransferQueue or https://gitcode.coβ¦β16Jan 16, 2026Updated 6 months ago
- [NeurIPS 2022] "Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speeβ¦β17Sep 19, 2023Updated 2 years ago
- Dynamic resources changes for multi-dimensional parallelism trainingβ31Aug 22, 2025Updated 11 months ago
- PTX ISA 9.1 documentation converted to searchable markdown. Includes Claude Code skill for CUDA development.β221Dec 24, 2025Updated 7 months ago
- Automated High-Performance GPU Kernel Generationβ122Jun 1, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- a simple API to use CUPTIβ10Aug 19, 2025Updated 11 months ago
- β16Jul 18, 2026Updated 2 weeks ago
- Official repository for ICML 2024 paper "MoRe Fine-Tuning with 10x Fewer Parameters"β22Oct 14, 2025Updated 9 months ago
- [NeurIPS 2021] "Drawing Robust Scratch Tickets: Subnetworks with Inborn Robustness Are Found within Randomly Initialized Networks" by Yonβ¦β13Feb 13, 2022Updated 4 years ago
- β11Oct 27, 2022Updated 3 years ago
- An agent harness that compiles a model into one provably-correct, self-retargeting CUDA megakernel and self-tunes it past cuBLAS at batchβ¦β125Jun 29, 2026Updated last month
- β15May 29, 2018Updated 8 years ago
- Agentic Kernel Optimization β advanced & eXtensible: a closed-loop, campaign-based multi-agent system for optimizing GPU kernels (benchmaβ¦β62May 31, 2026Updated 2 months ago
- Lightweight agent multiplexer, all in one Web dashboardβ54Jul 22, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ECCV 2022] SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruningβ20Jul 7, 2022Updated 4 years ago
- [ICML 2021] "Double-Win Quant: Aggressively Winning Robustness of Quantized DeepNeural Networks via Random Precision Training and Inferenβ¦β16Feb 13, 2022Updated 4 years ago
- β13Jun 8, 2021Updated 5 years ago
- β21Jul 22, 2022Updated 4 years ago
- Autonomous GPU kernel optimization system driven by AI agents.β31Mar 29, 2026Updated 4 months ago
- A Library for intra-GPU/Inter-SM parallelsimβ12May 21, 2024Updated 2 years ago
- β20Jan 6, 2023Updated 3 years ago