Unlocking Lossless Speedups in LLMs via Discrete Diffusion
☆97Sep 20, 2026Updated 2 weeks ago
Alternatives and similar repositories for uno
Users that are interested in uno are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 6 months ago
- [ICCV 2025] What Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models☆16Nov 3, 2025Updated 11 months ago
- [ICCAD 2025] Squant☆16Jul 3, 2025Updated last year
- dUltra: Ultra-Fast Diffusion Large Language Models via Reinforcement Learning☆18Jul 11, 2026Updated 2 months ago
- Django App per l'autocompilazione dei moduli missione☆32Jan 7, 2026Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Learning-Recurrent-Binary-Ternary-Weights☆13Dec 4, 2018Updated 7 years ago
- ☆14Mar 21, 2020Updated 6 years ago
- Code for the paper *Attention Drift: What Speculative Decoding Models Learn*.☆32May 12, 2026Updated 4 months ago
- A simple Python sandbox for helpful LLM data agents☆16May 4, 2025Updated last year
- [NeurIPS 2026] Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models☆24Updated this week
- [ICML25] Agentic Compression Benchmark (ACBench)☆21Jul 2, 2025Updated last year
- [CVPR 2025] QuartDepth☆18Mar 24, 2025Updated last year
- ☆10Sep 10, 2023Updated 3 years ago
- [ICCV 2025] Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes☆22Nov 9, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 2023/12/22 电三 420 每周会议技术分享:「容器」的 slides 和附件☆10Dec 22, 2023Updated 2 years ago
- [ICLR25] STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs☆19Jun 3, 2025Updated last year
- ☆20Apr 19, 2021Updated 5 years ago
- LatentMAS with kNN kv cache pruning | up to 40% more memory efficient and 30% faster☆19Dec 10, 2025Updated 9 months ago
- Reinforcement Learning Finetunes Small Subnetworks in Large Language Models☆15Oct 20, 2025Updated 11 months ago
- Spearmint uses Gaussian Processes to automatically optimize hyper parameter. This is a fork of Spearmint for the deep learning community.…☆11Nov 30, 2016Updated 9 years ago
- FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models. FS-DFM accepted for ICLR 2026☆51Sep 11, 2026Updated 3 weeks ago
- ☆19Nov 11, 2024Updated last year
- ☆43Oct 6, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This is the repository for codes in paper "ShaderPerFormer: Platform-independent Context-aware Shader Performance Predictor"☆13May 16, 2024Updated 2 years ago
- recipe for training fully-featured self supervised image jepa models☆14Jun 4, 2025Updated last year
- ☆25Apr 30, 2026Updated 5 months ago
- AI-Powered Command-Line Photo Tagging Tool☆17Mar 15, 2026Updated 6 months ago
- ☆21Jul 3, 2026Updated 3 months ago
- Python InfluxDB to VictoriaMetrics exporter script☆16Jan 3, 2021Updated 5 years ago
- Implementation of Flash-DLM (paper: FlashDLM: Accelerating Diffusion Language Models via Efficient KV Caching and Guided Diffusion). Prov…☆25Nov 25, 2025Updated 10 months ago
- Federated reconnaissance mini-ImageNet benchmark and baseline models☆13Sep 2, 2021Updated 5 years ago
- ☆27Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- One command · One Microsoft login · Zero repeated auth Hours of uninterrupted access to NYU Torch from your terminal and IDE.☆16Jul 17, 2026Updated 2 months ago
- Research into the HTTP APIs from various LLM providers.☆30Apr 5, 2026Updated 6 months ago
- DFVG: A Heterogeneous Architecture for Speculative Decoding with Draft-on-FPGA and Verify-on-GPU.☆28Nov 26, 2025Updated 10 months ago
- ☆14Mar 15, 2021Updated 5 years ago
- 写的快糙猛的课程作业(躺☆11May 2, 2020Updated 6 years ago
- A convenient framework for developing LLM- and LMM-based web agents. (ACL'24 Demo)☆29Aug 11, 2024Updated 2 years ago
- ☆16Mar 1, 2025Updated last year