The repository contains code for Adaptive Data Optimization
β37Dec 9, 2024Updated last year
Alternatives and similar repositories for ado
Users that are interested in ado are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.β14Jan 9, 2024Updated 2 years ago
- Official Code Repository for [AutoScaleπ: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*β¦β14Aug 8, 2025Updated last year
- A simple and efficient baseline for data attributionβ11Nov 10, 2023Updated 2 years ago
- Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]β80Nov 14, 2024Updated last year
- Azure OpenAI ν둬νν€β11Sep 23, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β41Dec 19, 2024Updated last year
- Simple and scalable tools for data-driven pretraining data selection.β30Jun 9, 2025Updated last year
- ACL24β11Jun 7, 2024Updated 2 years ago
- Official Repository for Dataset Inference for LLMsβ41Jul 25, 2024Updated 2 years ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]β33Jan 23, 2025Updated last year
- Code for our paper "AMR-DA: Data augmentation by abstract meaning representation" in ACL 2022β13May 17, 2022Updated 4 years ago
- β12Oct 20, 2023Updated 2 years ago
- Forcing Diffuse Distributions out of Language Modelsβ18Sep 10, 2024Updated last year
- Generating Potent Poisons and Backdoors from Scratch with Guided Diffusionβ11Apr 1, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β18Oct 12, 2022Updated 3 years ago
- β11Oct 20, 2023Updated 2 years ago
- Training vision models with full-batch gradient descent and regularizationβ40Feb 14, 2023Updated 3 years ago
- β93Aug 18, 2024Updated last year
- An empirical investigation of deep learning theoryβ16Oct 3, 2019Updated 6 years ago
- Code for the paper "The Journey, Not the Destination: How Data Guides Diffusion Models"β26Dec 12, 2023Updated 2 years ago
- β13Dec 12, 2025Updated 8 months ago
- β30Jun 19, 2023Updated 3 years ago
- Gemstones: A Model Suite for Multi-Faceted Scaling Laws (NeurIPS 2025)β35Sep 28, 2025Updated 10 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PAL: Proxy-Guided Black-Box Attack on Large Language Modelsβ57Aug 17, 2024Updated last year
- β15Oct 4, 2024Updated last year
- Source code of "What can linearized neural networks actually say about generalization?β20Oct 21, 2021Updated 4 years ago
- Data for "Datamodels: Predicting Predictions with Training Data"β97May 25, 2023Updated 3 years ago
- β14Jun 24, 2024Updated 2 years ago
- Reading comprehension based question-answering model for news articles.β11Jun 22, 2022Updated 4 years ago
- Code base for the EMNLP 2021 Findings paper: Cartography Active Learningβ14Jun 3, 2025Updated last year
- Levin tree search guided by both a policy and a heuristic functionβ19Jul 13, 2023Updated 3 years ago
- Official Pytorch repo of CVPR'23 and NeurIPS'23 papers on understanding replication in diffusion models.β113Nov 22, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Trainingβ24Aug 18, 2024Updated last year
- Implementation of experiments from The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learningβ17May 14, 2023Updated 3 years ago
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.β206Jul 17, 2024Updated 2 years ago
- β16Jul 17, 2022Updated 4 years ago
- This repository contains the replication of the iGSM dataset generation process from the Physics of LLM paper by Zeyuan Zhu.β17Sep 13, 2024Updated last year
- β10Jul 13, 2024Updated 2 years ago
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMsβ98Nov 17, 2024Updated last year