Accelerate pretraining by pre-pretraining on formal languages!
☆22Feb 13, 2026Updated 6 months ago
Alternatives and similar repositories for pre-pretraining
Users that are interested in pre-pretraining are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Jun 2, 2026Updated 2 months ago
- Online Hyperparameter Optimization☆11Feb 17, 2021Updated 5 years ago
- Menagerie of video models trained on various video datasets☆10Oct 13, 2024Updated last year
- An R package for implementing and evaluating Maximum Entropy Optimality Theory models☆10Aug 14, 2026Updated 2 weeks ago
- ☆26Aug 19, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆14Feb 1, 2024Updated 2 years ago
- Forecasting scientific progress with AI☆31Aug 16, 2026Updated 2 weeks ago
- Structural Supervision & Human Psycholinguistic Data☆13Apr 16, 2021Updated 5 years ago
- The repository for the paper "When Do You Need Billions of Words of Pretraining Data?"☆21Nov 10, 2020Updated 5 years ago
- ☆10Jun 19, 2019Updated 7 years ago
- Code for the ACL 2021 paper "Structural Guidance for Transformer Language Models"☆15Sep 17, 2025Updated 11 months ago
- HELP: a dataset for Handling Entailments with Lexical and logical Phenomena (Ver.1.0)☆15Jul 20, 2023Updated 3 years ago
- Behavioral probing of language acquisition models at the lexical and syntactic level☆20Jul 17, 2023Updated 3 years ago
- Design and analyze optimal deep learning models.☆30Aug 2, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code and Results for "Universals of word order reflect optimization of grammars for efficient communication"☆14Aug 5, 2022Updated 4 years ago
- This is a repository for the paper on testing inductive bias with scaled-down RoBERTa models.☆21Jan 10, 2022Updated 4 years ago
- A list of resources dedicated to compositionality☆14Feb 21, 2019Updated 7 years ago
- The Earleyx parser was originated from Roger Levy's prefix parser, but has evolved significantly. Earleyx can generate Viterbi parses and…☆16Mar 27, 2014Updated 12 years ago
- Learning from Neighbors: Unsupervised Text Classification☆17Sep 27, 2022Updated 3 years ago
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Training☆24Aug 18, 2024Updated 2 years ago
- ☆16Aug 13, 2026Updated 2 weeks ago
- Gradient accumulation on tf.estimator☆12Dec 15, 2020Updated 5 years ago
- ☆19Apr 21, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Syntactic evaluation sets, attribute-varying grammars, and code for replicating the CLAMS paper. ACL 2020.☆18Nov 26, 2024Updated last year
- PhD/MBA-level human-annotated rubrics dataset across Physics, Chemistry, Finance and Consulting☆34Oct 30, 2025Updated 9 months ago
- Probe how GPT-n performs on statutory reasoning☆10Sep 17, 2024Updated last year
- [ICML 2026] Public repository for fine-tuning Masked Diffusion Models toward provable self-correction.☆26Jul 5, 2026Updated last month
- ☆10Mar 5, 2024Updated 2 years ago
- Bayesian Inverse Graphics for Few-Shot Concept Learning☆12Mar 16, 2025Updated last year
- Fine-tuning GPT-2 to generate research paper abstracts☆12Apr 28, 2021Updated 5 years ago
- A LaTeX document class for notes 📝 and textbooks 📚☆15Jul 14, 2021Updated 5 years ago
- ☆18May 5, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for Pushdown Layers from our EMNLP 2023 paper☆29Dec 3, 2023Updated 2 years ago
- OneStop: A 360-Participant Eye Tracking Dataset with Different Reading Regimes☆20Apr 18, 2026Updated 4 months ago
- Tools for all things related to Combinatory Categorial Grammar☆20Jul 12, 2025Updated last year
- Continual pretraining of foundation LLM using ⚡ Lightning Fabric☆37Nov 27, 2024Updated last year
- This is the official implementation for our ACL 2024 paper: "Causal Estimation of Memorisation Profiles".☆25Mar 25, 2025Updated last year
- [ACL 2025] Predicting Turn-Taking and Backchannel in Human-Machine Conversations Using Linguistic, Acoustic, and Visual Signals☆35Aug 11, 2025Updated last year
- ☆19Jun 1, 2021Updated 5 years ago