A Lossless Compression Library for AI pipelines
☆324Apr 11, 2026Updated 3 months ago
Alternatives and similar repositories for zipnn
Users that are interested in zipnn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Gateway for connecting application services in different domains, networks, and cloud infrastructures☆23Feb 1, 2026Updated 5 months ago
- Python Memory Management☆18Aug 9, 2022Updated 3 years ago
- ☆110Jul 21, 2025Updated last year
- An official implementation of ProbeGen☆13Oct 20, 2024Updated last year
- RAG Templates Optimization Engine.☆18Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for "Attribute First, then Generate: Locally-attributable Grounded Text Generation", ACL 2024☆30Dec 19, 2024Updated last year
- Fybrik☆130Sep 7, 2025Updated 10 months ago
- General agent evaluation framework☆64Updated this week
- Top papers related to LLM-based agent evaluation☆102Oct 21, 2025Updated 9 months ago
- ☆11Aug 25, 2024Updated last year
- ☆14Jul 13, 2025Updated last year
- An official PyTorch implementation for CLIPPR☆31Jul 22, 2023Updated 3 years ago
- Official PyTorch Implementation for the "Recovering the Pre-Fine-Tuning Weights of Generative Models" paper (ICML 2024).☆86Apr 15, 2025Updated last year
- 🦄 Unitxt is a Python library for enterprise-grade evaluation of AI performance, offering the world's largest catalog of tools and data …☆217May 27, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Transform flow logs into metrics☆101Updated this week
- Official implementation of "DGD: Dynamic 3D Gaussians Distillation".☆69Aug 16, 2024Updated last year
- Let my Claude talk to yours.☆30Updated this week
- [AAAI 2025] Official Implementation for "Click2Mask: Local Editing with Dynamic Mask Generation" Paper.☆21Jan 22, 2026Updated 6 months ago
- TEAL: New Selection Strategy for Small Buffers in Experience Replay Class Incremental Learning☆18Jan 21, 2025Updated last year
- Official repository for the MMFM challenge☆26Jun 18, 2024Updated 2 years ago
- ☆20Updated this week
- llm-d Router: The intelligent entry point for inference requests☆266Updated this week
- SlamKit is an open source tool kit for efficient training of SpeechLMs. It was used for "Slamming: Training a Speech Language Model on On…☆229Mar 14, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Comprehensive LLM Error Analysis and Reporting☆53Updated this week
- llm-d benchmark scripts and tooling☆62Updated this week
- Distributed KV cache scheduling & offloading libraries☆162Updated this week
- Achieve state of the art inference performance with modern accelerators on Kubernetes☆3,873Updated this week
- PyTorch implementation of "Sample- and Parameter-Efficient Auto-Regressive Image Models" from CVPR 2025☆14Nov 21, 2025Updated 8 months ago
- An implementation of the Llama architecture, to instruct and delight☆21May 31, 2025Updated last year
- Boosting 4-bit inference kernels with 2:4 Sparsity☆96Sep 4, 2024Updated last year
- This repository contains code for the MicroAdam paper.☆21Dec 14, 2024Updated last year
- Repository for the paper: Teaching Structured Vision & Language Concepts to Vision & Language Models☆47Sep 25, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.☆33Updated this week
- Quantized Attention on GPU☆45Nov 22, 2024Updated last year
- ☆51Oct 29, 2024Updated last year
- ☆27Jun 11, 2026Updated last month
- Quality Controlled Paraphrase Generation (ACL 2022)☆71Sep 17, 2025Updated 10 months ago
- ☆36Apr 14, 2025Updated last year
- An intelligent orchestration system that automates the software development lifecycle from Feature ideation through code delivery using A…☆15Updated this week