Pretraining data reconstruction scripts for Apertus
☆157Oct 27, 2025Updated 10 months ago
Alternatives and similar repositories for pretrain-data
Users that are interested in pretrain-data are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 8 months ago
- Layout Analysis Dataset with Segmonto (LADaS)☆25May 29, 2026Updated 3 months ago
- ACL24☆11Jun 7, 2024Updated 2 years ago
- Deriving steepest descent convergence bounds and hyperparameter scaling laws in machine learning optimization from first principles, form…☆17Apr 11, 2026Updated 5 months ago
- This API provides authentication and CRUD operations for data used by the Chronas application☆14Sep 13, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆19Feb 25, 2024Updated 2 years ago
- ☆24Jan 30, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 4 months ago
- Forcing Diffuse Distributions out of Language Models☆18Sep 10, 2024Updated 2 years ago
- ☆19Updated this week
- Data mapping framework for rust stuff☆59Mar 25, 2026Updated 5 months ago
- ☆16Updated this week
- MANOVA.RM☆11Sep 3, 2026Updated 2 weeks ago
- Code for "Merging Text Transformers from Different Initializations"☆20Feb 2, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Create synthetic datasets from scratch using AI-powered generation. Define topics, customize prompts, and generate high-quality reasoning…☆32Jul 23, 2026Updated last month
- A tool for detecting moral values in social discourse☆20Apr 24, 2025Updated last year
- Training codebase for K2-V2☆22Dec 17, 2025Updated 9 months ago
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆19Jul 4, 2025Updated last year
- Blender exporter for data3d - a binary format optimized for webGL.☆15Apr 22, 2020Updated 6 years ago
- A tiny easily hackable implementation of a feature dashboard.☆18Oct 21, 2025Updated 10 months ago
- PyTorch building blocks for the OLMo ecosystem☆1,538Updated this week
- Muon fsdp 2☆65Aug 8, 2025Updated last year
- ☆13Dec 12, 2025Updated 9 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A lightweight, user-friendly data-plane for LLM training.☆41Sep 10, 2025Updated last year
- UNLP 2025 Shared Task on Detecting Social Media Manipulation☆23Aug 4, 2025Updated last year
- decontamination☆38Mar 4, 2026Updated 6 months ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 5 months ago
- ☆265Oct 27, 2025Updated 10 months ago
- ☆21Jul 4, 2025Updated last year
- Source code of "What can linearized neural networks actually say about generalization?☆20Oct 21, 2021Updated 4 years ago
- Video Resource Platform☆10Nov 15, 2024Updated last year
- ☆15May 27, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- WikiVideo: Article Generation from Multiple Videos☆15Nov 14, 2025Updated 10 months ago
- Rethinking the Trust Region in LLM Reinforcement Learning☆74Mar 2, 2026Updated 6 months ago
- OLMost every training recipe you need to perform data interventions with the OLMo family of models.☆75Jul 21, 2026Updated last month
- Codes and files for the paper Are Emergent Abilities in Large Language Models just In-Context Learning☆33Jan 9, 2025Updated last year
- Code for the paper "BPE stays on SCRIPT", "Which Pieces Does Unigram Tokenization Really Need?" and MinGram☆22Updated this week
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆34Jan 23, 2025Updated last year
- ☆19Aug 17, 2022Updated 4 years ago