☆16Jul 9, 2026Updated 2 months ago
Alternatives and similar repositories for PT2-LLM
Users that are interested in PT2-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] QuantCache:Adaptive Importance-Guided Quantization with Hierarchical Latent and Layer Caching for Video Generation☆18Sep 26, 2025Updated 11 months ago
- RobuQ: Pushing DiTS to W1.58A2 via Robust Activation Quantization☆17Jun 28, 2026Updated 2 months ago
- ☆17Oct 5, 2025Updated 11 months ago
- PyTorch code for our paper "Binarized Dual Residual Network for 3D Whole-body Human Mesh Recovery"☆15Dec 2, 2023Updated 2 years ago
- ⛵️ Replacing global cargo shipment with autonomous sailing vessels.☆13Jan 6, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [AAAI'26] PyTorch code for our paper "QuantVSR: Low-Bit Post-Training Quantization for Real-World Video Super-Resolution"☆34Jan 29, 2026Updated 7 months ago
- ☆41Sep 30, 2025Updated 11 months ago
- [ACL 2026 Main] Code for the paper "ARCQuant: Boosting NVFP4 Quantization with Augmented Residual Channels for LLMs"☆32Updated this week
- PyTorch code for our paper "2DQuant: Low-bit Post-Training Quantization for Image Super-Resolution"☆50Oct 24, 2024Updated last year
- [NeurIPS'25] OSCAR: One-Step Diffusion Codec Across Multiple Bit-rates☆41Oct 20, 2025Updated 11 months ago
- [ICML 2026] Memory-Efficient LLM Pretraining via Minimalist Optimizer Design☆23May 26, 2026Updated 3 months ago
- POS for African languages☆21Jun 25, 2025Updated last year
- ☆21Apr 3, 2025Updated last year
- ☆12May 14, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementations of growing and pruning in neural networks☆22Jul 26, 2023Updated 3 years ago
- (AAAI 2026) First-Order Error Matters: Accurate Compensation for Quantized Large Language Models☆17Apr 16, 2026Updated 5 months ago
- [CVPR 2025] AIGV-Assessor: Benchmarking and Evaluating the Perceptual Quality of Text-to-Video Generation with LMM☆19Mar 19, 2026Updated 6 months ago
- ☆12Dec 9, 2022Updated 3 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆66Sep 17, 2025Updated last year
- An Efficient Matrix Multiplication Algorithm for Accelerating Inference in Binary and Ternary Neural Networks☆17Mar 27, 2026Updated 5 months ago
- ☆16Oct 17, 2025Updated 11 months ago
- ☆15Oct 28, 2024Updated last year
- [NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.☆103Jan 3, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14Sep 16, 2017Updated 9 years ago
- ☆15Apr 18, 2026Updated 5 months ago
- A Python implementation of the "CoSyne" algorithim, as described in this paper: https://pdfs.semanticscholar.org/966e/41903b4aff42601a188…☆25Oct 23, 2018Updated 7 years ago
- Torch implementation of SRGAN (Ledig et al., Photo -Realistic Single Image Super-Resolution Using a Generative Adversarial Network, 2016)☆13Nov 17, 2016Updated 9 years ago
- This repository collects Visual Autoregressive (VAR) modeling papers from 2024 to 2026 published at top-tier conferences, as well as rele…☆22Mar 13, 2026Updated 6 months ago
- An experiment to see if chatgpt can improve the output of the stanford alpaca dataset☆12Mar 29, 2023Updated 3 years ago
- The code repository of "MBQ: Modality-Balanced Quantization for Large Vision-Language Models"☆96Mar 17, 2025Updated last year
- [RAL 2024 & IROS 2024] Enhancing Visual Place Recognition in Spatial Domain on Aerial Vehicle Platforms.☆16Oct 21, 2024Updated last year
- A sample app to debug and validate cellular modems on balena devices☆13Jun 5, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2025] CommVQ: Commutative Vector Quantization for KV Cache Compression☆27Sep 2, 2025Updated last year
- React 0.13 with ES6, Immutable.js and Flux, Isomorphic as well☆11Mar 10, 2015Updated 11 years ago
- ☆17Mar 10, 2025Updated last year
- AI Generated Tees☆39Mar 18, 2020Updated 6 years ago
- This repository contains low-bit quantization papers from 2020 to 2026 on top conference.☆221Aug 22, 2026Updated 3 weeks ago
- ☆15Jun 28, 2023Updated 3 years ago
- [Poster; ICLR 2026] [Oral; Neurips OPT2024] μLO: Compute-Efficient Meta-Generalization of Learned Optimizers☆16Apr 15, 2026Updated 5 months ago