☆38Jun 23, 2026Updated last month
Alternatives and similar repositories for MineDraft
Users that are interested in MineDraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Draft-Target Disaggregation LLM Serving System via Parallel Speculative Decoding.☆214Mar 18, 2026Updated 4 months ago
- Native Android app for Hyperliquid. DEX trading☆227Dec 30, 2025Updated 7 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆52May 12, 2026Updated 2 months ago
- NCU-driven iterative optimization workflow for CUDA/CUTLASS/Triton/CuTe DSL kernels.☆24Apr 10, 2026Updated 3 months ago
- Artifact for "Fail Fast, Win Big: Rethinking the Drafting Strategy in Speculative Decoding via Diffusion LLMs" [arXiv '25]☆21Jul 26, 2026Updated last week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆20Dec 24, 2024Updated last year
- A PyTorch native library for training speculative decoding models☆219Updated this week
- Evaluating GPT-OSS on BrowseComp-Plus with Native Browsering Tools☆20Oct 17, 2025Updated 9 months ago
- ☆18Jul 17, 2026Updated 2 weeks ago
- 吴恩达深度学习课程笔记PyTorch版☆16Nov 2, 2025Updated 9 months ago
- Spec-Bench: A Comprehensive Benchmark and Unified Evaluation Platform for Speculative Decoding (ACL 2024 Findings)☆403Apr 22, 2025Updated last year
- [ACL 2025 main] FR-Spec: Frequency-Ranked Speculative Sampling☆55Jul 15, 2025Updated last year
- A selective knowledge distillation algorithm for efficient speculative decoders☆39Nov 27, 2025Updated 8 months ago
- [ACL2025 Oral🔥]Turning Trash into Treasure: Accelerating Inference of Large Language Models with Token Recycling☆29Nov 11, 2025Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- QAQ: Quality Adaptive Quantization for LLM KV Cache☆55Mar 27, 2024Updated 2 years ago
- [ASPLOS'26] HILOS: A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs☆20Jan 18, 2026Updated 6 months ago
- MIPS R10000 architecture simulator with C++☆11Jun 8, 2023Updated 3 years ago
- Virtual Decoupled Cores: Composable Programming Framework and Runtime for Async GPUs☆20Updated this week
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆17Updated this week
- A curated list of recent papers on efficient video attention for video diffusion models, including sparsification, quantization, and cach…☆61Oct 27, 2025Updated 9 months ago
- The first range filter to simultaneously offer dynamicity, fast operations, and a robust false positive rate for any workload.☆13Jul 15, 2025Updated last year
- Agent-native Seedance 2.0 short-film studio: cli for AI, canvas for human☆16Jun 14, 2026Updated last month
- Train speculative decoding models effortlessly and port them smoothly to SGLang serving.☆1,049Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2025] AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism☆20Jul 14, 2025Updated last year
- Make FP4 on 5090 Great Again☆19Jul 20, 2026Updated 2 weeks ago
- 下载emotioNet_URLs的Python脚本,实现异步并行下载。☆10Dec 22, 2017Updated 8 years ago
- DeeperGEMM: crazy optimized version☆86May 5, 2025Updated last year
- Simulation, multi-path estimation, and CBR parsing code of SIGCOMM2023 BeamSense CBR-Sensing☆10Jan 14, 2024Updated 2 years ago
- recent audio generation papers (including speech, music and general audios)☆13Mar 14, 2023Updated 3 years ago
- fake CUTLASS to get peformance☆26Apr 28, 2026Updated 3 months ago
- ☆23Aug 20, 2025Updated 11 months ago
- [ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models☆23Mar 15, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- Multi-lingual AudioCaps☆14Nov 20, 2023Updated 2 years ago
- ☆52Dec 19, 2025Updated 7 months ago
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆76Mar 11, 2026Updated 4 months ago
- Intelligent Resource Requirement Estimation and Scheduling for Deep Learning Jobs on Distributed GPU Clusters☆16Nov 18, 2021Updated 4 years ago
- ☆88Oct 9, 2024Updated last year
- Get away from UCAS!!!☆14Aug 24, 2022Updated 3 years ago