☆41May 26, 2026Updated last month
Alternatives and similar repositories for olmix
Users that are interested in olmix are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Organize the Web: Constructing Domains Enhances Pre-Training Data Curation☆83May 2, 2025Updated last year
- My submission for the GPUMODE/AMD fp8 mm challenge☆29Jun 4, 2025Updated last year
- Debiasing Through Data Attribution☆13May 23, 2024Updated 2 years ago
- ☆21Jun 12, 2025Updated last year
- [ICML 2024] Selecting High-Quality Data for Training Language Models☆204Dec 8, 2025Updated 7 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICLR 2026] Official Implementation of ProxyThinker: Test-Time Guidance through Small Visual Reasoners.☆22Sep 24, 2025Updated 9 months ago
- A lightweight, user-friendly data-plane for LLM training.☆40Sep 10, 2025Updated 10 months ago
- [NeurIPS-2023] The PyTorch Implementation of MoSo. The algorithms are based on our paper: "Data Pruning via Moving-one-Sample-out". MoSo …☆10May 21, 2026Updated 2 months ago
- Reinforcement Learning from Text Feedback☆49Feb 17, 2026Updated 5 months ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆33Jan 23, 2025Updated last year
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 8 months ago
- The official github repo for "Training Optimal Large Diffusion Language Models", the first-ever large-scale diffusion language models sca…☆46Nov 6, 2025Updated 8 months ago
- Codebase for ICML submission "DOGE: Domain Reweighting with Generalization Estimation"☆21Feb 29, 2024Updated 2 years ago
- The official repository for SkyLadder: Better and Faster Pretraining via Context Window Scheduling☆43Dec 29, 2025Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19Nov 4, 2025Updated 8 months ago
- An automated data pipeline scaling RL to pretraining levels☆76Jun 2, 2026Updated last month
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated 11 months ago
- Codebase for Instruction Following without Instruction Tuning☆36Sep 24, 2024Updated last year
- LLM training in simple, raw C/CUDA☆15Dec 5, 2024Updated last year
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated 11 months ago
- ☆113Jul 15, 2025Updated last year
- DSIR large-scale data selection framework for language model training☆275Apr 7, 2024Updated 2 years ago
- Adversarial Tokenization☆39Nov 21, 2025Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2025] 🧬 RegMix: Data Mixture as Regression for Language Model Pre-training (Spotlight)☆194Feb 17, 2025Updated last year
- Aioli: A unified optimization framework for language model data mixing☆33Jan 17, 2025Updated last year
- Ongoing research project for code&math LLMs☆32Jul 4, 2025Updated last year
- Official github repo for the paper "Compression Represents Intelligence Linearly" [COLM 2024]☆150Sep 20, 2024Updated last year
- The raw UserRL repo under construction☆110Jun 2, 2026Updated last month
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated 11 months ago
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 10 months ago
- ☆24Aug 20, 2025Updated 11 months ago
- Arctic Training and Inference Platform☆56Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for ICML 25 paper "Metadata Conditioning Accelerates Language Model Pre-training (MeCo)"☆51Jun 30, 2025Updated last year
- The official implementation of NOSA☆19Jun 11, 2026Updated last month
- ☆18Sep 22, 2024Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scaling☆189Jul 23, 2025Updated 11 months ago
- Tooling for exact and MinHash deduplication of large-scale text datasets☆90Mar 24, 2026Updated 3 months ago
- Reproducible, flexible LLM evaluations☆388Mar 24, 2026Updated 3 months ago
- The SAIL-VL2 series model developed by the BytedanceDouyinContent Group☆79Sep 18, 2025Updated 10 months ago