A minimal PyTorch re-implementation of Qwen 3.5
☆430Jun 15, 2026Updated last month
Alternatives and similar repositories for tiny-qwen
Users that are interested in tiny-qwen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DeepSeek R1 distilled into smaller OSS models for hobbyist☆17Dec 2, 2025Updated 7 months ago
- Benchmark tests supporting the TiledCUDA library.☆19Nov 19, 2024Updated last year
- Nano vLLM☆14,557Apr 26, 2026Updated 2 months ago
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆4,957Oct 27, 2025Updated 8 months ago
- mnn asr demo.☆27Mar 24, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆102Feb 11, 2026Updated 5 months ago
- Static suckless single batch CUDA-only qwen3-0.6B mini inference engine☆557Sep 8, 2025Updated 10 months ago
- Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)☆726Sep 24, 2025Updated 9 months ago
- Local Qwen3 LLM inference. One easy-to-understand file of C source with no dependencies.☆183Jul 5, 2025Updated last year
- EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL☆5,071Updated this week
- Gensis is a lightweight deep learning framework written from scratch in Python, with Triton as its backend for high-performance computing…☆35Jan 15, 2026Updated 6 months ago
- qwen-nsa☆87Oct 14, 2025Updated 9 months ago
- slime is an LLM post-training framework for RL Scaling.☆7,551Updated this week
- ☆148Aug 18, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆26May 30, 2025Updated last year
- High performance inference engine for diffusion models☆107Sep 5, 2025Updated 10 months ago
- ☆37Aug 7, 2025Updated 11 months ago
- An open-source implementaion for fine-tuning Qwen-VL series by Alibaba Cloud.☆1,938Updated this week
- 🚀 Efficient implementations for emerging model architectures☆5,379Updated this week
- This is a training method to produce a split brain model