The repository contains code for Adaptive Data Optimization
β37Dec 9, 2024Updated last year
Alternatives and similar repositories for ado
Users that are interested in ado are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.β14Jan 9, 2024Updated 2 years ago
- Official Code Repository for [AutoScaleπ: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*β¦β14Aug 8, 2025Updated 11 months ago
- A simple and efficient baseline for data attributionβ11Nov 10, 2023Updated 2 years ago
- Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]β80Nov 14, 2024Updated last year
- Azure OpenAI ν둬νν€β11Sep 23, 2024Updated last year
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- β41Dec 19, 2024Updated last year
- Simple and scalable tools for data-driven pretraining data selection.β30Jun 9, 2025Updated last year
- ACL24β11Jun 7, 2024Updated 2 years ago
- Official Repository for Dataset Inference for LLMsβ41Jul 25, 2024Updated last year
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]β33Jan 23, 2025Updated last year
- β12Oct 20, 2023Updated 2 years ago
- Forcing Diffuse Distributions out of Language Modelsβ18Sep 10, 2024Updated last year
- Generating Potent Poisons and Backdoors from Scratch with Guided Diffusionβ11Apr 1, 2024Updated 2 years ago
- β18Oct 12, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β11Oct 20, 2023Updated 2 years ago
- This is the official implementation for our ACL 2024 paper: "Causal Estimation of Memorisation Profiles".β25Mar 25, 2025Updated last year
- Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning [ICML 2024]β21May 2, 2024Updated 2 years ago
- Training vision models with full-batch gradient descent and regularizationβ40Feb 14, 2023Updated 3 years ago
- An empirical investigation of deep learning theoryβ16Oct 3, 2019Updated 6 years ago
- Code for the paper "The Journey, Not the Destination: How Data Guides Diffusion Models"β25Dec 12, 2023Updated 2 years ago
- β13Dec 12, 2025Updated 7 months ago
- β30Jun 19, 2023Updated 3 years ago
- PAL: Proxy-Guided Black-Box Attack on Large Language Modelsβ57Aug 17, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch ImageNet1k Loader with Bounding Boxes.β13Jan 23, 2022Updated 4 years ago
- β15Oct 4, 2024Updated last year
- Data for "Datamodels: Predicting Predictions with Training Data"β97May 25, 2023Updated 3 years ago
- β14Jun 24, 2024Updated 2 years ago
- Reading comprehension based question-answering model for news articles.β11Jun 22, 2022Updated 4 years ago
- Levin tree search guided by both a policy and a heuristic functionβ19Jul 13, 2023Updated 3 years ago
- Official Pytorch repo of CVPR'23 and NeurIPS'23 papers on understanding replication in diffusion models.β113Nov 22, 2023Updated 2 years ago
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Trainingβ24Aug 18, 2024Updated last year
- Implementation of experiments from The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learningβ17May 14, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.β206Jul 17, 2024Updated 2 years ago
- This repository contains the replication of the iGSM dataset generation process from the Physics of LLM paper by Zeyuan Zhu.β17Sep 13, 2024Updated last year
- β10Jul 13, 2024Updated 2 years ago
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMsβ98Nov 17, 2024Updated last year
- Learning to route instances for Human vs AI Feedback (ACL Main '25)β29Jul 23, 2025Updated last year
- β19Mar 25, 2025Updated last year
- Does Refusal Training in LLMs Generalize to the Past Tense? [ICLR 2025]β79Jan 23, 2025Updated last year