The repository contains code for Adaptive Data Optimization
β37Dec 9, 2024Updated last year
Alternatives and similar repositories for ado
Users that are interested in ado are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.β14Jan 9, 2024Updated 2 years ago
- Official Code Repository for [AutoScaleπ: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*β¦β14Aug 8, 2025Updated last year
- A simple and efficient baseline for data attributionβ11Nov 10, 2023Updated 2 years ago
- Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]β79Nov 14, 2024Updated last year
- Azure OpenAI ν둬νν€β11Sep 23, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β41Dec 19, 2024Updated last year
- Simple and scalable tools for data-driven pretraining data selection.β31Jun 9, 2025Updated last year
- ACL24β11Jun 7, 2024Updated 2 years ago
- Official Repository for Dataset Inference for LLMsβ41Jul 25, 2024Updated 2 years ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]β34Jan 23, 2025Updated last year
- Forcing Diffuse Distributions out of Language Modelsβ18Sep 10, 2024Updated 2 years ago
- Generating Potent Poisons and Backdoors from Scratch with Guided Diffusionβ11Apr 1, 2024Updated 2 years ago
- β18Oct 12, 2022Updated 3 years ago
- β11Oct 20, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- This is the official implementation for our ACL 2024 paper: "Causal Estimation of Memorisation Profiles".β25Mar 25, 2025Updated last year
- Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning [ICML 2024]β21May 2, 2024Updated 2 years ago
- Training vision models with full-batch gradient descent and regularizationβ39Feb 14, 2023Updated 3 years ago
- β93Aug 18, 2024Updated 2 years ago
- An empirical investigation of deep learning theoryβ16Oct 3, 2019Updated 6 years ago
- Code for the paper "The Journey, Not the Destination: How Data Guides Diffusion Models"β27Dec 12, 2023Updated 2 years ago
- β13Dec 12, 2025Updated 9 months ago
- β30Jun 19, 2023Updated 3 years ago
- PAL: Proxy-Guided Black-Box Attack on Large Language Modelsβ57Aug 17, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pytorch ImageNet1k Loader with Bounding Boxes.β13Jan 23, 2022Updated 4 years ago
- β15Oct 4, 2024Updated last year
- Source code of "What can linearized neural networks actually say about generalization?β20Oct 21, 2021Updated 4 years ago
- β14Jun 24, 2024Updated 2 years ago
- Reading comprehension based question-answering model for news articles.β11Jun 22, 2022Updated 4 years ago
- The official repository for SkyLadder: Better and Faster Pretraining via Context Window Schedulingβ43Dec 29, 2025Updated 8 months ago
- Levin tree search guided by both a policy and a heuristic functionβ19Jul 13, 2023Updated 3 years ago
- Official Pytorch repo of CVPR'23 and NeurIPS'23 papers on understanding replication in diffusion models.β114Nov 22, 2023Updated 2 years ago
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Trainingβ24Aug 18, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code and Data for the ACL 2022 paper "Rethinking Self-Supervision Objectives for Generalizable Coherence Modeling"β10Apr 5, 2022Updated 4 years ago
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.