☆41Jun 23, 2026Updated 2 months ago
Alternatives and similar repositories for MineDraft
Users that are interested in MineDraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] PEARL: Parallel Speculative Decoding with Adaptive Draft Length☆172Dec 23, 2025Updated 8 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆54May 12, 2026Updated 4 months ago
- NCU-driven iterative optimization workflow for CUDA/CUTLASS/Triton/CuTe DSL kernels.☆25Apr 10, 2026Updated 5 months ago
- Artifact for "Fail Fast, Win Big: Rethinking the Drafting Strategy in Speculative Decoding via Diffusion LLMs" [arXiv '25]☆21Jul 26, 2026Updated last month
- ☆20Dec 24, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A PyTorch native library for training speculative decoding models☆247Sep 12, 2026Updated last week
- ☆20Updated this week
- ☆19Sep 10, 2026Updated last week
- Spec-Bench: A Comprehensive Benchmark and Unified Evaluation Platform for Speculative Decoding (ACL 2024 Findings)☆412Apr 22, 2025Updated last year
- [ACL 2025 main] FR-Spec: Frequency-Ranked Speculative Sampling☆56Jul 15, 2025Updated last year
- A selective knowledge distillation algorithm for efficient speculative decoders☆39Nov 27, 2025Updated 9 months ago
- a distributed computation platform for running Python and Bash computation tasks on multiple nodes☆12Jul 28, 2026Updated last month
- [ACL2025 Oral🔥]Turning Trash into Treasure: Accelerating Inference of Large Language Models with Token Recycling☆30Nov 11, 2025Updated 10 months ago
- QAQ: Quality Adaptive Quantization for LLM KV Cache☆55Mar 27, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ASPLOS'26] HILOS: A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs☆22Jan 18, 2026Updated 8 months ago
- MIPS R10000 architecture simulator with C++☆11Jun 8, 2023Updated 3 years ago
- 吴恩达深度学习课程笔记PyTorch版☆20Nov 2, 2025Updated 10 months ago
- Virtual Decoupled Cores: Composable Programming Framework and Runtime for Async GPUs☆22Sep 8, 2026Updated last week
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆18Updated this week
- The first range filter to simultaneously offer dynamicity, fast operations, and a robust false positive rate for any workload.☆13Jul 15, 2025Updated last year
- Agent-native Seedance 2.0 short-film studio: cli for AI, canvas for human☆16Sep 8, 2026Updated last week
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.☆1,177Updated this week
- [ICML 2025] AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism☆20Jul 14, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Make FP4 on 5090 Great Again☆24Jul 20, 2026Updated last month
- DeeperGEMM: crazy optimized version☆85May 5, 2025Updated last year
- fake CUTLASS to get peformance☆25Apr 28, 2026Updated 4 months ago
- ☆23Aug 20, 2025Updated last year
- [EMNLP26] Welcome! 😊 This is the official code release of EviNote-RAG, and we’re happy to share it with the community.☆49Aug 24, 2026Updated 3 weeks ago
- [ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models☆23Mar 15, 2024Updated 2 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Multi-lingual AudioCaps☆14Nov 20, 2023Updated 2 years ago
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆77Mar 11, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆56Dec 19, 2025Updated 9 months ago
- Intelligent Resource Requirement Estimation and Scheduling for Deep Learning Jobs on Distributed GPU Clusters☆16Nov 18, 2021Updated 4 years ago
- ☆89Oct 9, 2024Updated last year
- Get away from UCAS!!!☆14Aug 24, 2022Updated 4 years ago
- [EMNLP2026 Main] Official implementation of “Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding”.☆138Jul 25, 2026Updated last month
- Continuous Pipelined Speculative Decoding☆22May 25, 2026Updated 3 months ago
- ☆30Jul 22, 2024Updated 2 years ago