☆40Jun 23, 2026Updated 2 months ago
Alternatives and similar repositories for MineDraft
Users that are interested in MineDraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Draft-Target Disaggregation LLM Serving System via Parallel Speculative Decoding.☆217Mar 18, 2026Updated 5 months ago
- [ICLR 2025] PEARL: Parallel Speculative Decoding with Adaptive Draft Length☆172Dec 23, 2025Updated 8 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆53May 12, 2026Updated 3 months ago
- NCU-driven iterative optimization workflow for CUDA/CUTLASS/Triton/CuTe DSL kernels.☆25Apr 10, 2026Updated 4 months ago
- Artifact for "Fail Fast, Win Big: Rethinking the Drafting Strategy in Speculative Decoding via Diffusion LLMs" [arXiv '25]☆21Jul 26, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆20Dec 24, 2024Updated last year
- A PyTorch native library for training speculative decoding models☆236Updated this week
- Evaluating GPT-OSS on BrowseComp-Plus with Native Browsering Tools☆20Oct 17, 2025Updated 10 months ago
- ☆19Jul 17, 2026Updated last month
- 吴恩达深度学习课程笔记PyTorch版☆17Nov 2, 2025Updated 9 months ago
- Spec-Bench: A Comprehensive Benchmark and Unified Evaluation Platform for Speculative Decoding (ACL 2024 Findings)☆407Apr 22, 2025Updated last year
- [ACL 2025 main] FR-Spec: Frequency-Ranked Speculative Sampling☆55Jul 15, 2025Updated last year
- A selective knowledge distillation algorithm for efficient speculative decoders☆39Nov 27, 2025Updated 9 months ago
- ☆12May 19, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- a distributed computation platform for running Python and Bash computation tasks on multiple nodes☆12Jul 28, 2026Updated last month
- [ACL2025 Oral🔥]Turning Trash into Treasure: Accelerating Inference of Large Language Models with Token Recycling☆30Nov 11, 2025Updated 9 months ago
- QAQ: Quality Adaptive Quantization for LLM KV Cache☆55Mar 27, 2024Updated 2 years ago
- Это full-stack веб-приложение, которое я спроектировал и реализовал самостоятельно. Оно напоминает крупные маркетплейсы вроде Ozon или Wi…☆15Nov 11, 2025Updated 9 months ago
- [ASPLOS'26] HILOS: A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs☆22Jan 18, 2026Updated 7 months ago
- MIPS R10000 architecture simulator with C++☆11Jun 8, 2023Updated 3 years ago
- Virtual Decoupled Cores: Composable Programming Framework and Runtime for Async GPUs☆21Updated this week
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆18Updated this week
- ☆17Oct 5, 2025Updated 10 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A curated list of recent papers on efficient video attention for video diffusion models, including sparsification, quantization, and cach…☆64Oct 27, 2025Updated 10 months ago
- The first range filter to simultaneously offer dynamicity, fast operations, and a robust false positive rate for any workload.☆13Jul 15, 2025Updated last year
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.☆1,131Updated this week
- [ICML 2025] AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism☆20Jul 14, 2025Updated last year
- Make FP4 on 5090 Great Again☆24Jul 20, 2026Updated last month
- DeeperGEMM: crazy optimized version☆85May 5, 2025Updated last year
- Simulation, multi-path estimation, and CBR parsing code of SIGCOMM2023 BeamSense CBR-Sensing☆10Jan 14, 2024Updated 2 years ago
- Study Notes for 《AI Systems Performance Engineering》☆63Jul 24, 2026Updated last month
- fake CUTLASS to get peformance☆26Apr 28, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆23Aug 20, 2025Updated last year
- [ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models☆23Mar 15, 2024Updated 2 years ago
- ☆55Dec 19, 2025Updated 8 months ago
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆76Mar 11, 2026Updated 5 months ago
- Intelligent Resource Requirement Estimation and Scheduling for Deep Learning Jobs on Distributed GPU Clusters☆16Nov 18, 2021Updated 4 years ago
- ☆88Oct 9, 2024Updated last year
- Official implementation of “Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding”.☆137Jul 25, 2026Updated last month