☆243Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for opensloth
Users that are interested in opensloth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- unsloth-5090-multiple☆62May 21, 2025Updated last year
- Bản dịch tiếng Việt của 100 bài luyện tập NLP (cập nhật bản 2020) dịch từ 言語処理100本ノック 2020 (https://nlp100.github.io/ja)☆24Jun 8, 2020Updated 6 years ago
- ☆10Mar 24, 2025Updated last year
- ☆30Mar 29, 2024Updated 2 years ago
- A list of articles outside of the official MLIR docs that I've found useful for learning MLIR☆13Aug 16, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluati…☆54May 7, 2025Updated last year
- ☆11Jan 1, 2024Updated 2 years ago
- Train your own SOTA deductive reasoning model☆111Mar 6, 2025Updated last year
- Code for the paper "FinRLlama: A Solution to LLM-Engineered Signals Challenge at FinRL Contest 2024"☆12Feb 14, 2025Updated last year
- Quy Nhon AI Hackathon 2022 - Challenge 2: Review Analytics - Top 1 Solution☆11Sep 21, 2022Updated 3 years ago
- Small, simple agent task environments for training and evaluation☆20Nov 1, 2024Updated last year
- Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM☆3,743Updated this week
- ☆14Oct 31, 2023Updated 2 years ago
- Fused Qwen3 MoE layer for faster training, compatible with Transformers, LoRA, bnb 4-bit quant, Unsloth. Also possible to train LoRA over…☆259Jul 24, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Go ahead and axolotl questions☆12,425Updated this week
- Alternate LLM Architecture Experiment with O(n) time complexity☆11Jul 18, 2026Updated last month
- Creating a chatbot from your facebook data with GPT☆23May 10, 2022Updated 4 years ago
- ☆28Oct 7, 2025Updated 10 months ago
- Baseline achieving 0.8 accuracy on the private test set in the ZaloAI Challenge 2023 Elementary Math Solving☆24May 1, 2024Updated 2 years ago
- Qwen2 VL Fine Tuning using Llama Factory☆19Sep 7, 2024Updated last year
- Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.☆75,251Updated this week
- ☆44Nov 12, 2025Updated 9 months ago
- Efficient kernel for RMS normalization with fused operations, includes both forward and backward passes, compatibility with PyTorch.☆13Jun 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A pipeline parallel training script for LLMs.☆168Apr 30, 2025Updated last year
- vietnamese-ready-to-production RAG☆19Jul 17, 2024Updated 2 years ago
- Llama 3 ORPO Fine Tuning on A100 in Colab Pro.☆12Apr 21, 2024Updated 2 years ago
- Foundation of MLOps☆17Dec 22, 2024Updated last year
- Large-scale LLM inference engine☆1,844Aug 13, 2026Updated 2 weeks ago
- A collection of riddles I've compiled. (Free Use)☆22Apr 22, 2023Updated 3 years ago
- Efficient Triton Kernels for LLM Training☆6,593Updated this week
- Extend the Conditioning of Stable Diffusion to take Audio Embeddings Instead of Text Embeddings using Wav2Vec2-BERT model☆13Sep 25, 2024Updated last year
- Pdf Query chat-bot using Gemini AI and Llma Index☆10Dec 24, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A ready to use boilerplate Flask App for Data Scientist, ML engineer...☆16Jan 31, 2023Updated 3 years ago
- Autonomously train research-agent LLMs on custom data using reinforcement learning and self-verification.☆687Mar 22, 2025Updated last year
- Tools for merging pretrained large language models.☆7,324Jun 17, 2026Updated 2 months ago
- Qwen GRPO Graph Extraction RL Finetune☆71Apr 2, 2025Updated last year
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.☆206Jul 17, 2024Updated 2 years ago
- AnyModal is a Flexible Multimodal Language Model Framework for PyTorch☆106Apr 11, 2026Updated 4 months ago
- [ACL 2025 Main] EfficientQAT: Efficient Quantization-Aware Training for Large Language Models☆348Apr 10, 2026Updated 4 months ago