Home of StarCoder2!
☆2,092Mar 21, 2024Updated 2 years ago
Alternatives and similar repositories for starcoder2
Users that are interested in starcoder2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Home of StarCoder: fine-tuning & inference!☆7,499Feb 27, 2024Updated 2 years ago
- Code for the curation of The Stack v2 and StarCoder2 training data☆144Apr 11, 2024Updated 2 years ago
- official repository of aiXcoder-7B Code Large Language Model☆2,270Jul 9, 2025Updated last year
- Inference code for CodeLlama models☆16,255Aug 12, 2024Updated 2 years ago
- [NeurIPS'24] SelfCodeAlign: Self-Alignment for Code Generation☆325Feb 24, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICML'24] Magicoder: Empowering Code Generation with OSS-Instruct☆2,096Nov 1, 2024Updated last year
- LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath☆9,478Jun 7, 2025Updated last year
- OpenCodeInterpreter is a suite of open-source code generation systems aimed at bridging the gap between large language models and sophist…☆1,767May 7, 2024Updated 2 years ago
- Modeling, training, eval, and inference code for OLMo☆6,680Nov 24, 2025Updated 9 months ago
- Official implementation for the paper: "Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering""☆3,967Nov 25, 2024Updated last year
- DeepSeek Coder: Let the Code Write Itself☆24,276Nov 11, 2025Updated 10 months ago
- open-source coding agent☆35,950Updated this week
- A framework for the evaluation of autoregressive code generation language models.☆1,062Jul 22, 2025Updated last year
- Large World Model -- Modeling Text and Video with Millions Context☆7,430Oct 19, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Large Language Model Text Generation Inference☆10,886Mar 21, 2026Updated 5 months ago
- 🙌 OpenHands: AI-Driven Development☆88,418Updated this week
- Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We als…☆18,560May 19, 2026Updated 3 months ago
- Official inference library for Mistral models☆10,825Jun 16, 2026Updated 3 months ago
- DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence☆7,014Nov 11, 2025Updated 10 months ago
- 20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.☆13,668Updated this week
- A programming framework for agentic AI☆61,044Apr 15, 2026Updated 5 months ago
- Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.☆181,209Updated this week
- This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.☆12,206Mar 8, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A series of large language models trained from scratch by developers @01-ai☆7,833Nov 27, 2024Updated last year
- Tools for merging pretrained large language models.☆7,357Sep 12, 2026Updated last week
- DeepSeek-VL: Towards Real-World Vision-Language Understanding☆4,180Apr 24, 2024Updated 2 years ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆92,107Updated this week
- The official Meta Llama 3 GitHub site☆29,224Jan 26, 2025Updated last year
- Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.☆76,370Updated this week
- ☆497Aug 15, 2024Updated 2 years ago
- EvoEval: Evolving Coding Benchmarks via LLM☆84Apr 6, 2024Updated 2 years ago
- Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.☆16,838Mar 24, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.☆9,012May 3, 2024Updated 2 years ago
- 🐙 OctoPack: Instruction Tuning Code Large Language Models☆480Feb 5, 2025Updated last year
- tiny vision language model☆10,050Apr 20, 2026Updated 4 months ago
- Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024☆1,816Oct 2, 2025Updated 11 months ago
- Accelerate your Hugging Face Transformers 7.6-9x. Native to Hugging Face and PyTorch.☆684Aug 22, 2024Updated 2 years ago
- SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersec…☆20,352Updated this week
- aider is AI pair programming in your terminal☆49,038May 22, 2026Updated 3 months ago